> Markdown version of [/jobs/ext/2360487-staff-technical-program-manager-site-reliability](https://www.wearedevelopers.com/jobs/ext/2360487-staff-technical-program-manager-site-reliability). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff Technical Program Manager, Site Reliability... - **Company:** MongoDB - **Location:** New York, NY, United States (Remote available) - **Experience:** Expert - **Salary:** $151,000.0 - $297,000.0 - **Contract:** Permanent contract - **Skills:** JIRA, Cloud Computing, MongoDB, Reliability Engineering, Software Engineering, Cloud Platform System, Kubernetes - **Published:** August 8, 2026 - **Apply:** https://www.juju.com/job/00000000gmb79n ## About the Role + 8+ years in technical program management, engineering management, or a comparable technical role partnering with software engineering teams + Proven track record leading large-scale, cross-team platform initiatives through ambiguity and change + Strong knowledge of production change management, software development lifecycle, and reliability metrics (SLOs, SLIs) + Skilled at shaping roadmaps and managing dependencies + Able to query and interpret metrics, logs, or other data sources to inform decisions and communicate risk + Excellent communicator-clear, concise, and calm-across engineers, cross-functional partners, and executives + Low-ego, highly collaborative, and motivated by ownership of hard problems end to end Nice to Have + Hands-on or close-partner experience with Kubernetes, cloud networking, or observability stacks (metrics, logs, tracing, alerting) + Prior experience working with or alongside SRE teams + Background in large-scale cloud infrastructure or platform engineering + Familiarity with MongoDB Atlas or other modern cloud database platforms ## Description As a TPM for SRE, you will partner with SRE leaders and engineers to scale the platform that underpins all of MongoDB's cloud products. You will drive program execution, strengthen production reliability practices, and coordinate cross-functional efforts across US and EMEA teams. Success in this role means smoother launches, clearer roadmaps, stronger reliability metrics and an SRE organization that's better-equipped to deliver predictability at scale. This role can be based remotely on the East Coast What You'll Do + Drive Program Planning & Execution - Define program scope, milestones, and success criteria with SRE engineers and leaders. Manage dependencies across platform teams, keep work clearly tracked in Jira, and deliver on time + Strengthen Production Reliability - Lead change management and launch readiness programs. Partner with SREs and product teams to define and operationalize SLOs/SLIs, and use incident data, metrics, and capacity signals to drive prioritization and continuous improvement + Lead Cross-Functional Coordination - Align SRE with Security, Compliance, Cloud platform, and other engineering teams. Coordinate cross-team incident response, ensure clear follow-through, and build trust as the go-to driver of complex, multi-team efforts + Build Scalable Systems & Processes - Design lightweight frameworks and communication patterns that help SRE deliver reliably at scale. Work yourself out of the "hero" role by leaving teams better-equipped to execute independently ## Related Videos - [Protector Of The Realm](https://www.wearedevelopers.com/videos/591-protector-of-the-realm) - [40 Minutes to Build a Serverless COVID-19 REST and GraphQL APIs](https://www.wearedevelopers.com/videos/208-40-minutes-to-build-a-serverless-covid-19-rest-and-graphql-apis) - [Improving quality with Agentic AI with Rovo Dev and Xray](https://www.wearedevelopers.com/videos/2005-improving-quality-with-agentic-ai-with-rovo-dev-and-xray) - [Understanding Kubernetes in a visual way](https://www.wearedevelopers.com/videos/100085-understanding-kubernetes-in-a-visual-way) - [Collaboration Quantified: Lessons from Open Source Developer Networks](https://www.wearedevelopers.com/videos/1422-collaboration-quantified-lessons-from-open-source-developer-networks) - [Kubernetes Maestro: Dive Deep into Custom Resources to Unleash Next-Level Orchestration Power!](https://www.wearedevelopers.com/videos/1087-kubernetes-maestro-dive-deep-into-custom-resources-to-unleash-next-level-orchestration-power) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Should Tech Managers Be Developers First? Pros and Cons](https://www.wearedevelopers.com/magazine/327-should-tech-managers-be-developers-first-pros-and-cons) - [From developer to manager – what does it take to become an engineering manager?](https://www.wearedevelopers.com/magazine/42-from-developer-to-manager-what-does-it-take-to-become-an-engineering-manager) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)