Senior Database Reliability Engineer
Crunchyroll, LLCAbout the role
<div class="content-intro"><h2 data-pm-slice="1 1 []">About Crunchyroll</h2> <p>Founded by fans, Crunchyroll delivers the art and culture of anime to a passionate community. We super-serve over 100 million anime and manga fans across 200+ countries and territories, and help them connect with the stories and characters they crave. Whether that experience is online or in-person, streaming video, theatrical, games, merchandise, events and more, it’s powered by the anime content we all love.</p> <p>Join our team, and help us shape the future of anime!</p></div><p>Crunchyroll is growing and changing, presenting unique challenges and opportunities to support millions of anime fans around the world. The Database Operations Engineering team provides a seamless infrastructure foundation to our internal stakeholders, ensuring an exceptional experience for all Crunchyroll fans.</p> <p>As a Senior Database Reliability Engineer, you will be primarily responsible for operating, improving, and maintaining the reliability and operational excellence of our data infrastructure. Your core focus will be to introduce robust best practices for database production support, strengthen our global on-call rotation, and design and build reusable, database-specific Infrastructure as Code (IaC) components to ensure high availability, scalability, and 100% automation.</p> <h2><strong>Key Areas of Responsibility</strong></h2> <ul> <li><strong>Database Operational Excellence & Production Support:</strong> Drive, stabilize, and own 24x7 database production support operations, processes, and incident remediation. Responsibly track database alerts, establish clear operational procedures, and bring infrastructure alerts to rapid closure.</li> <li><strong>Database Infrastructure as Code (IaC) & Automation:</strong> Architect, implement, and maintain reusable database-specific IaC components and configurations using frameworks like Terraform, CloudFormation or Pulumi. Standardize configurations across multiple datastores to enable automated infrastructure deployment, sizing, and posture management.</li> <li><strong>Core Configuration Management:</strong> Proactively enable and standardize mission-critical database attributes and configurations by default, including automated backups, failover strategies, timeouts, and lifecycle policies.</li> <li><strong>On-Call & Platform Reliability:</strong> Strengthen and actively participate in the database on-call rotation, identifying SLAs, system vulnerabilities, and operational gaps to eliminate Single Points of Failure (SPOF).</li> <li><strong>Database SRE & Site Operations:</strong> Manage large-scale data infrastructures, execute cluster management, capacity planning, data go
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s