> Markdown version of [/jobs/ext/1096316-site-reliability-engineer-senior-or-staff](https://www.wearedevelopers.com/jobs/ext/1096316-site-reliability-engineer-senior-or-staff). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer (Senior or Staff) - **Company:** MongoDB - **Location:** New York, NY, United States - **Experience:** Expert - **Salary:** $127,000.0 - $249,000.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Microsoft Azure, Business Process Modeling, Cloud Computing, Computer Networks, Continuous Delivery, Continuous Integration, Linux, Distributed Systems, Domain Name System (DNS), Python (Programming Language), MongoDB, Routing, Open Source Technology, Reliability Engineering, Software Engineering, Software Systems, TCP/IP, Transport Layer Security, Google Cloud, Load Balancing, Containerization, Kubernetes - **Published:** June 18, 2026 - **Apply:** https://dejobs.org/x/x/17AAFFD1BF9148FA80ACF2010AD737DC/job/ ## About the Role * Have 6+ years of experience in software development and operating distributed systems * Proficiency in Python, Go, or a similar language * Proven experience building and operating large-scale continuous integration and continuous deployment (CI/CD) pipelines * Possess a customer-focused mindset * Value efficiency in processes and operations * Prefer automation over manual process ("allergic to ops work"). We are a small team of software engineers with a strong bias towards software solutions to avoid toil * Experience using and extending containerization technologies, particularly Kubernetes, to enhance application agility, optimize resource utilization, and accelerate time-to-market * Expertise in cloud infrastructure platforms, including AWS, Google Cloud Platform (GCP), or Azure * Understanding of Linux operating system internals and networking concepts (e.g., TCP/IP, DNS, TLS, routing) ## Description Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, networking, load balancing (including our public-facing edge and internal service mesh), and observability and alerting systems. The Deployments team designs and maintains our continuous delivery infrastructure, ensuring reliable code deployment from development through production for all engineering teams. This infrastructure is primarily composed of Argo Workflows and ArgoCD. The team also provides tooling that enables clear system ownership and facilitates self-service onboarding for development teams., * Contribute to developing a world-class continuous deployment experience, enabling the rapid and reliable shipment of MongoDB products * This includes, but is not limited to, contributing to open-source projects, or engineering software-based approaches like Kubernetes operators to streamline processes * Own the onboarding flow other engineering teams follow when launching a new product or service * Collaborate with other teams within Platform Engineering to ensure a consistent service-onboarding experience * Provide internal support for our deployment systems, including answering questions and addressing issues * Participate in a 24/7 on-call rotation to resolve issues involving the deployment infrastructure ## Related Videos - [An Applied Introduction to eBPF with Go](https://www.wearedevelopers.com/videos/1075-an-applied-introduction-to-ebpf-with-go) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Creating a routing app with Google Maps API from scratch](https://www.wearedevelopers.com/videos/831-creating-a-routing-app-with-google-maps-api-from-scratch) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Turning Container security up to 11 with Capabilities](https://www.wearedevelopers.com/videos/718-turning-container-security-up-to-11-with-capabilities) - [Reliable scalability: How Amazon.com scales on AWS](https://www.wearedevelopers.com/videos/983-reliable-scalability-how-amazon-com-scales-on-aws) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers)