> Markdown version of [/jobs/ext/3022931-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/3022931-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** Mastercard - **Location:** O'Fallon, MO, United States - **Experience:** Expert - **Salary:** $135,000.0 - $180,000.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Microsoft Azure, Bash Shell, Cloud Computing, Configuration Management, Computer Programming, Customer Data Management, Distributed Systems, Monitoring of Systems, Python (Programming Language), Linux System Administration, Performance Tuning, Reliability Engineering, Prometheus, Datadog, Data Logging, Scripting, Grafana, Git, Containerization, Kubernetes, Infrastructure Automation Frameworks, Terraform, New Relic (SaaS), Docker, Jenkins, Golang, Microservices - **Published:** September 21, 2026 - **Apply:** https://find.jobs/jobs-near-me/apply/ats-redirect/?id=2968577147-2 ## About the Role * Site Reliability Engineering (SRE) * Cloud platforms (AWS, Azure, or GCP) * Kubernetes and containerization (Docker) * Infrastructure as Code (Terraform/Cloud * Formation) * CI/CD pipelines (Jenkins, Git * Hub Actions, Git * Lab CI) * Linux systems administration * Monitoring and observability (Prometheus, Grafana, Datadog, New Relic) * Scripting/programming (Python, Go, Bash) * Distributed systems and microservices * Incident management and on-call operations ## Description Mastercard is seeking a Senior Site Reliability Engineer to enhance reliability, scalability, and performance of our critical IT & Data Management platforms. You will design resilient architectures, automate deployments, and champion observability to ensure always-on services. Collaborating with cross-functional teams, you'll identify and resolve production issues, implement robust incident management, and drive SRE best practices. This role offers the opportunity to work with cutting-edge cloud and container technologies in a culture that values innovation, collaboration, and continuous growth., * Design and maintain highly available, scalable, and secure cloud infrastructure for Mastercard's IT & Data Management platforms. * Build and improve automation for deployments, configuration management, and infrastructure provisioning. * Implement and refine monitoring, logging, and alerting to ensure service reliability and rapid incident detection. * Lead and participate in incident response, root cause analysis, and post-incident reviews to drive continuous improvement. * Partner with development and data teams to embed SRE best practices, including SLIs/SLOs and capacity planning. * Optimize system performance and cost efficiency across distributed, cloud-native environments. * Develop tools and scripts to reduce toil and improve operational excellence. * Contribute to security, compliance, and governance standards within production environments. ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) ## Related Articles - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [The Best X (Twitter) Accounts for Developers](https://www.wearedevelopers.com/magazine/294-the-best-x-twitter-accounts-for-developers)