> Markdown version of [/jobs/ext/2828853-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/2828853-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** Maintainx Inc. - **Location:** San Francisco, CA, United States (Remote available) - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** DevOps, Distributed Systems, Node.Js, Reliability Engineering, Software Engineering, TypeScript, Cloud Platform System, Programming Languages - **Published:** September 10, 2026 - **Apply:** https://startup.jobs/site-reliability-engineer-getmaintainx-com-9952466 ## About the Role * Deep understanding of observability practices in a distributed system environment and how it influences system design and team behaviour * Practical experience with SRE concepts (SLOs, error budgets, incident management) * 3-5+ years in software development, SRE, DevOps, or production development roles with experience operating production systems * Proficient in cloud-native platforms and infrastructure-as-code concepts and tools * Working knowledge of at least one programming language (TypeScript/Node.js is a plus) * Excellent communication and collaboration abilities across technical and non-technical teams * Ability to translate complex reliability concepts into actionable guidance * You enjoy enabling teams to succeed independently and measuring success by reduced dependency on you ## Description In this role, you'll partner closely with product and platform development teams to improve the stability, resilience, and operational readiness of our services. You'll work alongside teams to design for reliability from the start, establish clear ownership and standards, and build shared tooling that enables teams to operate their services with confidence. You'll also contribute to company-wide initiatives that define how MaintainX approaches reliability software development, including observability standards, incident response practices, and service health metrics, helping the organization adopt proven industry practices at scale. This role is well-suited for an developer who enjoys working across teams, influencing technical direction through strong development practices, and turning reliability principles into practical, scalable systems. What You'll Do: * Assess service maturity and provide insights to development teams * Partner with development teams to implement observability best practices * Enable development teams to become autonomous with their service deployment, support, and infrastructure * Mentor developers on reliability practices, focusing on making them self-sufficient * Act as the bridge, ear and eyes of the Platform Division teams to drive tooling and practice adoption across development teams