> Markdown version of [/jobs/ext/1774792-devops-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/1774792-devops-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # DevOps / Site Reliability Engineer - **Company:** General Matter - **Location:** Los Angeles, CA, United States - **Salary:** $100,000.0 - $200,000.0 - **Contract:** Permanent contract - **Skills:** Clean Code Principles, Automation of Tests, Cloud Computing, Computer Networks, Continuous Integration, Software Debugging, DevOps, Programming Tools, Distributed Systems, Domain Name System (DNS), Reliability Engineering, Prometheus, Software Systems, Web Services, Datadog, SSL Certificate Management, Data Logging, Transport Layer Security, Grafana - **Published:** July 17, 2026 - **Apply:** https://www.dice.com/job-detail/87c75975-5bdd-4da5-a572-d23c09393d9d ## About the Role * Strong fundamentals in web service development and distributed systems * Solid understanding of networking concepts, DNS, TLS/certificate management, and HTTP * Experience operating and debugging production systems * Familiarity with observability tools (metrics, logging, alerting) and incident response * Ability to write clear, maintainable code and automation scripts * Demonstrated ownership, attention to detail, and sound technical judgment Preferred Skills and Experience * Experience with modern observability stacks (e.g., Prometheus, Grafana, OpenTelemetry, Datadog) * Hands-on experience with cloud infrastructure and infrastructure-as-code * Exposure to CI/CD pipelines and developer tooling at scale * Experience supporting safety-critical or high-reliability systems * Strong debugging skills across application, OS, and network boundaries * Prior on-call experience in a production environment Additional Requirements: * Ability to work extended hours and weekends as necessary. ## Description We are seeking a highly capable DevOps / Site Reliability Engineer to help build and operate the software systems underpinning uranium enrichment R&D and production infrastructure. This role is foundational to our reliability, safety, and developer velocity. You will be responsible for designing and maintaining observability, alerting, and developer productivity systems, and for ensuring that critical internal and production services are correctly instrumented and monitored. We are only interested in candidates with strong fundamentals, sound judgment, and the ability to operate with rigor in a production environment where failures matter., * Design, implement, and maintain observability and alerting systems across critical services and infrastructure * Ensure all production and internal services are properly instrumented with metrics, logs, and traces * Own and maintain developer productivity tools, CI/CD systems, and internal platforms * Participate in an on-call rotation and respond to production incidents with urgency and discipline * Lead incident reviews and drive long-term reliability improvements * Automate operational workflows to reduce manual toil and improve system resilience ## Related Videos - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Debugging in the Dark](https://www.wearedevelopers.com/videos/1658-debugging-in-the-dark) - [DevOps at Netflix](https://www.wearedevelopers.com/videos/270-devops-at-netflix) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [Software Engineering Social Connection: Yubo’s lean approach to scaling an 80M-user infrastructure](https://www.wearedevelopers.com/videos/1583-software-engineering-social-connection-yubo-s-lean-approach-to-scaling-an-80m-user-infrastructure) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [DevOps Engineer Salary [2023]](https://www.wearedevelopers.com/magazine/203-devops-engineer-salary-2023) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [The Best Software Developer Blogs to Read](https://www.wearedevelopers.com/magazine/156-the-best-software-developer-blogs-to-read) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated)