Site Reliability Software Engineer II

Mastercard
O'Fallon, MO, United States
1 day ago
Apply on find.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
$125,000.0 - $165,000.0
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Microsoft Azure Bash Shell Cloud Computing Monitoring of Systems Python (Programming Language) Network Security Linux System Administration Reliability Engineering Prometheus Scripting Kubernetes
+3 more
Information Technology Deployment Automation Terraform

Job description

Mastercard is seeking a Site Reliability Engineer II to enhance the reliability, scalability, and performance of critical IT systems. In this role, you will design and implement robust monitoring, automation, and incident response processes to support innovative, user focused software products. You will collaborate closely with development and infrastructure teams to build and maintain highly available, secure cloud environments. The culture emphasizes innovation, collaboration, and continuous growth, offering opportunities to work with cutting edge technologies and drive meaningful improvements in global IT services., * Design, implement, and maintain highly available, scalable production systems.

  • Develop and optimize monitoring, alerting, and observability solutions.
  • Automate deployments, configuration, and operational tasks using Infrastructure as Code and scripting.
  • Collaborate with development and infrastructure teams to improve reliability and performance.
  • Participate in on-call rotations, incident response, and post-incident reviews.
  • Identify and remediate reliability, capacity, and performance bottlenecks.
  • Implement and refine SRE best practices, including SLIs, SLOs, and error budgets.
  • Contribute to security, compliance, and resilience improvements across systems.

Requirements

  • Site Reliability Engineering (SRE) practices
  • Cloud platforms (AWS, GCP, or Azure)
  • Linux system administration
  • Kubernetes and container orchestration
  • Infrastructure as Code (Terraform, Cloud
  • Formation, etc.)
  • CI/CD pipelines and automation
  • Monitoring and observability (Prometheus, Grafana, etc.)
  • Scripting (Python, Bash, or similar)
  • Incident management and on-call operations
  • Networking and security fundamentals

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on find.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:04 min

Introduction to Bitcoin script parsing tools

Steve Shadders · LIVE

1:34 min

Essential commands for running and testing Terraform configurations

Hennie Francis · LIVE

2:12 min

Integrating tool definitions for local bash shell execution

Michał Michalczuk Michał Michalczuk · Europe 2026 Virtual

1:46 min

Introduction to the speaker and engineering background

Llywelyn Griffith-Swain · World Congress 2023

1:53 min

Evaluating traditional scripting languages for modern development tasks

Jens Knipper Jens Knipper · Europe 2026 Virtual

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

Videos

See all

Related articles

See all