> Markdown version of [/jobs/ext/1134091-senior-database-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/1134091-senior-database-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Database Reliability Engineer - **Company:** Crunchyroll, Inc. - **Location:** San Francisco, CA, United States (Remote available) - **Experience:** Expert - **Salary:** $203,400.0 - $254,200.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Big Data, Configuration Management, Databases, Data Governance, Data Infrastructure, Data Stores, Data Systems, Database Storage Structures, DevOps, Amazon DynamoDB, Monitoring of Systems, Load Testing, MariaDB, NoSQL, Reliability Engineering, SQL Databases, Datadog, Pulumi, System Availability, Database Performance, Infrastructure as Code (IaC), Cloudformation, Build Management, Information Technology, Low Latency, Performance Monitor, Data Management, Cloudwatch, Terraform, Sql Tuning - **Published:** July 2, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=92eeb2967527088b ## About the Role We get excited about candidates, like you, because you possess: * Bachelor's degree in Computer Science, Information Technology, or a related field. * 8+ years of experience in database operations, site reliability engineering (SRE), or a related role with a heavy focus on data platforms and core operational infrastructure. * Strong proficiency in Automation and IaC frameworks, with extensive hands-on experience in building database IaC * Proven track record in Database Production Support and Operations, with deep practical experience managing robust, highly available 24x7 runtime systems (prior experience handling large-scale database production support at scale is highly valued). * Extensive experience with the AWS cloud platform and hands-on implementation of CI/CD pipelines and DatabaseOps workflows. * Proficiency in monitoring and observability tools (e.g., Datadog, CloudWatch, DevOps Guru, Database Performance Insights) to track metrics, latency, throughput, and system errors. * Strong understanding of various system performance metrics at a low level (such as Disk/IO saturation) and experience identifying or eliminating operational bottlenecks. * Familiarity with managing large-scale database structures across various systems (SQL and NoSQL). * Strong problem-solving skills, ownership mentality, proactive communication skills, and a baseline capability to document clear incident response playbooks and operational requirements. ## Description As a Senior Database Reliability Engineer, you will be primarily responsible for operating, improving, and maintaining the reliability and operational excellence of our data infrastructure. Your core focus will be to introduce robust best practices for database production support, strengthen our global on-call rotation, and design and build reusable, database-specific Infrastructure as Code (IaC) components to ensure high availability, scalability, and 100% automation., * Database Infrastructure as Code (IaC) & Automation: Architect, implement, and maintain reusable database-specific IaC components and configurations using frameworks like Terraform, CloudFormation or Pulumi. Standardize configurations across multiple datastores to enable automated infrastructure deployment, sizing, and posture management. * Core Configuration Management: Proactively enable and standardize mission-critical database attributes and configurations by default, including automated backups, failover strategies, timeouts, and lifecycle policies. * On-Call & Platform Reliability: Strengthen and actively participate in the database on-call rotation, identifying SLAs, system vulnerabilities, and operational gaps to eliminate Single Points of Failure (SPOF). * Database SRE & Site Operations: Manage large-scale data infrastructures, execute cluster management, capacity planning, data governance, compliance reviews, and handle complex data store migrations (such as MariaDB to Aurora/DynamoDB) and major version upgrades safely during non-US low traffic hours. * Collaborative Growth & Development: Work alongside a seasoned team of database engineering specialists (leveraging existing senior architectural depth on the team) to systematically scale platform features while executing a continuous learning roadmap to expand personal depth in native AWS database services (RDS/Aurora) and complex SQL tuning., The Database Operations Engineering team is dedicated to ensuring the reliability, scalability, and performance of our data infrastructure. We focus on standardizing and implementing monitoring and alerting across all datastores to track key metrics like errors, latency, and throughput, and to ensure critical systems are covered. Our team leads horizontal efforts to keep databases up-to-date, implements Infrastructure as Code (IaC) for high availability and performance, and automates key processes to enhance operational efficiency. We lead and evangelize the principle of 100% automation. Additionally, we define and document operational requirements, develop incident response processes, and automate monitoring and compliance checks to maintain a secure and reliable data environment. By continuously improving load testing and optimizing data governance practices, we support the overall health and efficiency of our data systems. ## Related Videos - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [How building an industry DBMS differs from building a research one](https://www.wearedevelopers.com/videos/768-how-building-an-industry-dbms-differs-from-building-a-research-one) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Why segmenting your infrastructure into tiers makes your infrastructure design better](https://www.wearedevelopers.com/videos/1960-why-segmenting-your-infrastructure-into-tiers-makes-your-infrastructure-design-better) - [Database Magic behind 40 Million operations/s](https://www.wearedevelopers.com/videos/748-database-magic-behind-40-million-operations-s) - [Tomorrow's cloud data platforms - fully managed database-as-a-service (DBaaS)](https://www.wearedevelopers.com/videos/254-tomorrow-s-cloud-data-platforms-fully-managed-database-as-a-service-dbaas) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)