Site Reliability Engineer (SRE)

Unique System Skills LLC
St. Louis, MO, United States
1 day ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Microsoft Azure Bash Shell Unix Computer Networks Linux DevOps Domain Name System (DNS) Github Monitoring of Systems Python (Programming Language) NoSQL
+24 more
OpenShift Windows PowerShell Reliability Engineering Ansible Prometheus SQL Databases Datadog Scripting Transport Layer Security Google Cloud Load Balancing Grafana Reliability of Systems Firewalls (Computer Science) Infrastructure as Code (IaC) Cloudformation Gitlab-ci Kubernetes Terraform Splunk Appdynamics Dynatrace Docker Jenkins

Job description

  • Design, implement, and maintain highly available, resilient, and scalable production environments.
  • Monitor application and infrastructure health using observability and APM tools.
  • Automate operational tasks using Infrastructure as Code (IaC) and scripting.
  • Collaborate with Development, DevOps, Security, and Infrastructure teams to improve system reliability.
  • Lead production incident response, troubleshooting, root cause analysis (RCA), and post-incident reviews.

Requirements

  • 10+ years of experience as a Site Reliability Engineer, DevOps Engineer, or Production Support Engineer.
  • Experience supporting Banking, Financial Services, FinTech, or Payment applications.
  • Strong knowledge of Linux/Unix administration.
  • Hands-on experience with AWS, Azure, or Google Cloud Platform.
  • Expertise in Kubernetes, Docker, OpenShift, or container orchestration platforms.
  • Experience with Terraform, Ansible, or CloudFormation.
  • Strong scripting skills in Python, Bash, PowerShell, or Go.
  • Experience with Jenkins, GitHub Actions, GitLab CI, or Azure DevOps.
  • Monitoring and observability tools such as Splunk, Dynatrace, Grafana, Prometheus, ELK, AppDynamics, or Datadog.
  • Knowledge of networking concepts, load balancing, DNS, SSL/TLS, and firewalls.
  • Experience with SQL and NoSQL databases.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:03 min

Microsoft integrating native Unix coreutils into Windows environments

Chris Heilmann +2 · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

3:50 min

Navigating specialized roles and toolsets across engineering teams

Nele Uhlemann · World Congress 2023

2:04 min

Defining timestamps and the international standard format

Denny Biasiolli Denny Biasiolli · Europe 2026 Virtual

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · World Congress 2021

Videos

See all

Related articles

See all