Site Reliability Engineer

Generic Solutions , Inc
Blue Ash, OH, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$93,600.0 - $104,000.0
Working hours
Shift work
Job source

Tech stack

JIRA Microsoft Azure Bash Shell Cloud Computing Linux Python (Programming Language) Reliability Engineering Scripting System Availability Reliability of Systems Kubernetes Dynatrace
+1 more
Docker

Job description

  • Lead and coordinate P1/P2 production incidents as an Incident Commander/Bridge Lead.
  • Troubleshoot and resolve complex production issues across cloud, on-premises, and enterprise/POS applications.
  • Perform detailed Root Cause Analysis (RCA) using methodologies such as 5 Whys and Fishbone.
  • Implement corrective and preventive actions to improve platform stability.
  • Monitor application and infrastructure health using Dynatrace, Azure Monitor, and related tools.
  • Automate operational tasks using Bash or Python scripting.
  • Work closely with development, infrastructure, and business teams to ensure high system availability.
  • Participate in a shared on-call rotation and periodic travel as required.

Requirements

We are looking for an experienced Site Reliability Engineer (SRE) with a strong background in Production Support, Major Incident Management, and Root Cause Analysis. This is a hands-on production support role where you’ll troubleshoot critical production issues, improve system reliability, and drive automation across cloud and on-premises environments., * 5+ years of experience in Site Reliability Engineering (SRE) or Production Support

  • Strong experience leading Major Incident Management for P1/P2 outages
  • Hands-on experience with Root Cause Analysis (RCA) and Problem Management
  • Experience troubleshooting production issues in Cloud and On-Premises environments
  • Strong knowledge of:
  • Dynatrace
  • Azure Monitor
  • Linux
  • Bash and/or Python
  • Kubernetes
  • Docker
  • Azure or GCP
  • Jira
  • Excellent communication and collaboration skills

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:50 min

Navigating specialized roles and toolsets across engineering teams

Nele Uhlemann · WWC 2023

3:05 min

Integrating an assistant application with Jira software

Felix Augenstein · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

8:02 min

Integrating service level objectives into incident management

Diana Todea · LIVE

5:47 min

Integrating user stories and test automation via Jira tools

Christoph Ruggenthaler · LIVE

Videos

See all

Related articles

See all