Site Reliability Engineer (SRE)

Application Management Services LLC
Pittsburgh, PA, United States
1 day ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Microsoft Azure Cloud Computing Monitoring of Systems Python (Programming Language) Performance Tuning Reliability Engineering Prometheus Google Cloud Grafana Kubernetes Splunk
+3 more
Appdynamics Dynatrace Docker

Job description

  • Maintain application reliability, availability, and performance
  • Monitor production environments and proactively identify issues
  • Automate operational tasks and reduce manual intervention
  • Support incident response, troubleshooting, and root cause analysis
  • Collaborate with development, infrastructure, and cloud teams
  • Improve system scalability, resilience, and operational efficiency

Requirements

  • Site Reliability Engineering (SRE)
  • Linux/Unix Administration
  • Monitoring & Observability (Splunk, Dynatrace, AppDynamics, Prometheus, Grafana)
  • CI/CD Pipelines
  • Kubernetes & Docker
  • Cloud Platforms (AWS/Azure/Google Cloud Platform)
  • Automation & Scripting (Python, Shell)
  • Incident Management & Production Support
  • Performance Tuning and Reliability Engineering

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:50 min

Navigating specialized roles and toolsets across engineering teams

Nele Uhlemann · World Congress 2023

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

1:04 min

Visualizing Keycloak performance via standard Grafana troubleshooting dashboards

Alexander Schwartz Alexander Schwartz · World Congress 2025

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

Videos

See all

Related articles

See all