Jobs and Careers
CR

Staff DevOps Engineer, Embedded Cloud Reliability

Crunchyroll Inc.
San Francisco, United Statesfull_timeVerifiedPosted 7 Feb 2025
💰 $255,000/yr($210,000/yr$255,000/yr)

About the role

About Crunchyroll

WE HELP EVERYONE BELONG. IT’S OUR PURPOSE.

Founded by fans, Crunchyroll delivers the art and culture of anime to a passionate community. We super-serve over 100 million anime and manga fans across 200+ countries and territories, and help them connect with the stories and characters they crave. Whether that experience is online or in-person, streaming video, theatrical, games, merchandise, events and more, it’s powered by the anime content we all love.

Join our team, and help us shape the future of anime!

Who We Are

We're a cast of characters working to shine a spotlight on anime. Crunchyroll is an international business focused on creating both online and offline experiences for fans through content (licensed, co-produced, originals, distribution), merchandise, events, gaming, news, and more. Visit our About Us pages for more information about our collection of brands.

About the Team

At Crunchyroll, our platforms and infrastructure form the foundation on which our services are built and directly influence our customer experience and velocity of our engineers. The Cloud Reliability team at Crunchyroll embeds with our development teams and partners with our core platform teams to deliver the critical cloud infrastructure that enable our services.

About The Role

We are seeking a Staff Engineer to join Cloud Reliability. As part of the team, you will work closely with Crunchyroll’s service development teams on their platform and infrastructure needs with an emphasis on improving system stability, engineering efficiency, and creating mechanisms allowing for greater self-sufficiency. You will be a champion of operational excellence within these teams and a bar raiser. As a Staff Engineer, you will be a technical leader for the team, identifying issues common across various development teams, working with our core platform teams on solving them, owning major internal infrastructure projects, and driving improvements throughout our platform. This role is required to be hybrid two days per week in our San Francisco, Culver City or Dallas office and will report into our Senior Manager. 

About You

  • 12+ years of experience in building and running high volume customer facing services in highly dynamic environments  in Software Engineering, Site Reliability, or related roles. 
  • Experienced in mentoring other engineers and helping guide them towards success. 
  • Leads complex projects spanning multiple teams. Drives technical issues towards resolution and builds strong relationships across engineering. 
  • BS Degree in Computer Science or a related field. 
  • Proficient in programming in TypeScript with familiarity of other languages such as Go or Python. 
  • Experienced in automation, infra as code, and making reusable patterns.
  • Passionate about improving the reliability and performance of critical services through the use of monitoring, metrics, incident management, and proactive engineering. Has helped investigate and remediate critical issues in production services and infrastructure.
  • Knowledgeable in performance and load testing tools and methods to simulate production workloads.
  • Expert in observability tools such as DataDog and has hands-on experience instrumenting services for monitoring, logging, metrics collection, tracing. 
  • Acts with urgency, ownership, and with a mindset of continuous improvement. 
  • Able to participate in an on-call rotation to ensure issues are resolved as quickly as possible and prevented from further occurrence. 
  • Highly collaborative, team oriented, possessing excellent communication skills, and able to communicate effectively with different audiences. 
  • Experienced in GitOps practices and technologies (CI/CD, Infrastructure as Code, etc). 
  • 5+ years experience in AWS and its major services such as ECS, S3, SQS, Lambda, containerization.
  • Experienced with database technologies such as RDS and DynamoDB.

Pluses 

  • Extensive experience with scaling and managing relational databases and nosql technologies.
  • Experience with ElasticSearch. 
  • 2+ years of experience in GCP and its common services such as GKE, Artifact Repository, Google VPC

Why you will love working at Crunchyroll

Not only will you get to work with fun, passionate and inspired colleagues, you will also...

  • Receive a great compensation package including salary plus performance bonus earning potential, paid annually.

  • Enjoy flexible PTO and time off policies allowing you to take the time you need to be your whole self.

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Crunchyroll Inc.

View company profile →