Sr Engineer - Site Reliability

World Wide Technology
St. Louis, MO, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$108,400.0 - $135,500.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Computer Programming Databases Continuous Integration Data Sharing Linux DevOps Distributed Systems Github Python (Programming Language) PostgreSQL Enterprise Messaging Systems
+17 more
MongoDB OpenShift RabbitMQ Redis Reliability Engineering Prometheus Runbook Software Engineering Datadog Scripting Istio Grafana Caching Build Management Kubernetes Low Latency Jenkins

Job description

The INFOPS Application Development team focuses on enabling developers at scale. We build and evolve internal platforms, automation, and self-service capabilities that allow application teams to deploy and operate software safely and efficiently.

Our work emphasizes:

  • Platform Engineering best practices
  • Kubernetes and OpenShift-based platforms
  • GitOps-driven CI/CD automation
  • Shared data and messaging platforms (MongoDB, PostgreSQL, RabbitMQ, etc.)
  • AI-assisted and automation-driven engineering workflows

As the organization matures, we are evolving toward a model that blends Platform Engineering with Site Reliability Engineering (SRE) to ensure our platforms are not only scalable-but also highly reliable and resilient., * Improve platform availability, latency, and scalability across Kubernetes and supporting services, establishing and evolving reliability standards across shared platform capabilities, and measuring outage impact as a health signal to track progress in preventing, detecting, and recovering from failures

  • Design and implement monitoring, alerting, and observability frameworks across platform services, standardizing telemetry (metrics, logs, traces) for consistent visibility across environments, and proactively identifying risks and performance bottlenecks before they impact users
  • Lead triage and resolution of platform-related incidents, driving root cause analysis, long-term fixes, and blameless postmortem practices
  • Collaborate with Support Engineers on the incident queue to surface recurring failure patterns, develop long-term remediation plans, reduce operational toil through automation, and drive sustainable resolution of systemic issues
  • Design and build automation to reduce manual intervention, improve incident response and recovery time, and enable self-healing platform behaviors; develop reusable tools, runbooks, and operational patterns that scale across teams
  • Collaborate with INFOPS Engineering and the Support Automation team to hand off support work where appropriate
  • Partner with internal customers, application teams, and platform/DevOps engineering functions to deliver scalable, repeatable solutions, address shared reliability challenges, and promote a culture of operational excellence and continuous improvement, We strive to create an environment where all employees are empowered to succeed based on their skills, performance, and dedication. Our goal is to cultivate a culture of belonging that encourages innovation, collaboration, and respect for all team members, ensuring that WWT remains a great place to work for All!

Requirements

  • 5+ years of experience in Site Reliability Engineering, Platform Engineering, or related roles supporting production systems
  • Strong experience with Kubernetes-based platforms; OpenShift experience preferred
  • Proven experience designing and operating distributed systems at scale
  • Experience implementing monitoring, alerting, and incident response practices
  • Strong scripting or programming skills (Python, Go, or similar) with a focus on automation
  • Experience with CI/CD pipelines and GitOps workflows (GitHub, Jenkins, or similar)
  • Experience with observability tooling (Prometheus, Grafana, ThousandEyes, or similar)
  • Strong Linux and troubleshooting skills across complex systems
  • Excellent communication skills with the ability to collaborate across technical and non-technical teams, * Experience supporting internal developer platforms or platform engineering organizations
  • Familiarity with shared services such as databases, messaging systems, and caching layers (MongoDB, PostgreSQL, Redis, RabbitMQ)
  • Experience implementing SLOs and error budget frameworks
  • Exposure to service mesh, traffic management, or advanced Kubernetes networking
  • Experience mentoring engineers and influencing technical decision-making

Preferred Location: MO

Benefits & conditions

Certain states and localities require employers to post a reasonable estimate of salary range. A reasonable estimate of the current base pay range for this position is $108,400 to $135,500 annually. Actual salary will be based on a variety of factors, including shift, location, experience, skill set, performance, licensure and certification, and business needs. The range for this position in other geographic locations may differ. Certain positions may also be eligible for variable incentive compensation, such as bonuses or commissions, that is not included in the base pay.

The well-being of WWT employees is essential. So, when it comes to our benefits package, WWT has one of the best. We offer the following benefits to all full-time employees:

  • Health and Wellbeing: Health, Dental, and Vision Care, Onsite Health Centers, Employee Assistance Program, Wellness program
  • Financial Benefits: Competitive pay, Profit Sharing, 401k Plan with Company Matching, Life and Disability Insurance, Tuition Reimbursement
  • Paid Time Off: PTO and Sick Leave (starting at 20 days per year) & Holidays (10 per year), Parental Leave, Military Leave, Bereavement
  • Additional Perks: Nursing Mothers Benefits, Voluntary Legal, Pet Insurance, Employee Discount Program

About the company

At World Wide Technology, we work together to make a new world happen. Our important work benefits our clients and partners as much as it does our people and communities across the globe. WWT is dedicated to achieving its mission of creating a profitable growth company that is also a Great Place to Work for All. We achieve this through our world-class culture, generous benefits and by delivering cutting-edge technology solutions for our clients.

Founded in 1990, WWT is a global technology solutions provider leading the AI and Digital Revolution. WWT combines the power of strategy, execution and partnership to accelerate digital transformational outcomes for organizations around the globe. Through its Advanced Technology Center, a collaborative ecosystem of the world’s most advanced hardware and software solutions, WWT helps clients and partners conceptualize, test and validate innovative technology solutions for the best business outcomes and then deploys them at scale through its global warehousing, distribution and integration capabilities.

With over 12,000 employees across WWT and Softchoice and more than 60 locations around the world, WWT’s culture, built on a set of core values and established leadership philosophies, has been recognized 14 years in a row by Fortune and Great Place to Work® for its unique blend of determination, innovation and creating a great place to work for all.

Want to work with highly motivated individuals on high-performance teams? Join WWT today!

Why Join the Internal IT Team?

The Internal WWT IT team serves as the backbone of the company’s technology ecosystem. We design, build, and operate the platforms that power application development across the enterprise. Our work enables secure, scalable, and highly automated delivery of software and services that directly support WWT’s strategic objectives.

Joining this team means working on foundational technology that matters-shared platforms, developer tooling, and automation that WWT depends on every day. You’ll be part of a collaborative, forward-looking engineering culture where ownership, innovation, and continuous improvement are expected and encouraged.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.jobmonkeyjobs.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

2:53 min

Configuring dynamic proxy updates with Istio Pilot

Jan Mensch Jan Mensch · WWC Europe 2026

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

3:55 min

Demonstrating semantic routing thresholds with the Redis vector library

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · WWC 2021

7:15 min

Installing Istio programmatically with bash scripts

Thomas Südbröcker · LIVE

Videos

See all

Related articles

See all