> Markdown version of [/jobs/ext/2224869-manager-site-reliability-engineering](https://www.wearedevelopers.com/jobs/ext/2224869-manager-site-reliability-engineering). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Manager, Site Reliability Engineering - **Company:** Mastercard - **Location:** O'Fallon, MO, United States - **Salary:** $145,000.0 - $200,000.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Microsoft Azure, Bash Shell, Cloud Computing, Electronic Design Automation, Monitoring of Systems, Python (Programming Language), Linux System Administration, Performance Tuning, Reliability Engineering, Prometheus, Software Engineering, Data Logging, Scripting, Grafana, Kubernetes, Terraform, New Relic (SaaS) - **Published:** August 25, 2026 - **Apply:** https://find.jobs/jobs-near-me/apply/ats-redirect/?id=2937667335-2 ## About the Role * Site Reliability Engineering (SRE) * Cloud platforms (AWS/Azure/GCP) * Kubernetes & container orchestration * Linux systems administration * CI/CD pipelines * Infrastructure as Code (Terraform/Cloud * Formation) * Monitoring & observability (Prometheus/Grafana/New Relic) * Incident management & on-call operations * Performance tuning & capacity planning * Scripting (Python/Bash) ## Description Mastercard seeks a Manager, Site Reliability Engineering to lead a team ensuring secure, scalable, and highly available platforms. You will design and implement SRE best practices, build automation for deployment and operations, and drive observability across complex cloud-native systems. Partner with development and security teams to improve reliability, performance, and incident response. You'll mentor engineers, champion continuous improvement, and help shape a culture of innovation, collaboration, and learning while working with cutting-edge technologies in a global environment., * Lead and mentor an SRE team supporting mission-critical platforms * Define and implement SRE best practices for reliability, scalability, and security * Design automation for deployments, configuration, and operations * Establish and improve monitoring, logging, and alerting for cloud-native systems * Drive incident management, root-cause analysis, and post-incident reviews * Collaborate with software engineering and security teams to improve system design * Optimize performance and capacity planning across services * Promote continuous improvement and a learning culture within the team ## Related Videos - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Infrastructure as Code: The Developer's Secret Weapon](https://www.wearedevelopers.com/videos/1221-infrastructure-as-code-the-developer-s-secret-weapon) - [JavaScript? No. Java Scripts! - Scripting with Java](https://www.wearedevelopers.com/videos/2094-javascript-no-java-scripts-scripting-with-java) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [All your telemetry data from any source in one place](https://www.wearedevelopers.com/videos/57-all-your-telemetry-data-from-any-source-in-one-place) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) ## Related Articles - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [The Best X (Twitter) Accounts for Developers](https://www.wearedevelopers.com/magazine/294-the-best-x-twitter-accounts-for-developers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [React Developer Salary [2023]](https://www.wearedevelopers.com/magazine/198-react-developer-salary-2023)