Site Reliability Engineer
Generic Solutions , Inc
Blue Ash, OH, United States
3 months ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Apply on www.indeed.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$93,600.0 - $104,000.0
Working hours
Shift work
Job source
Tech stack
JIRA
Microsoft Azure
Bash Shell
Cloud Computing
Linux
Python (Programming Language)
Reliability Engineering
Scripting
System Availability
Reliability of Systems
Kubernetes
Dynatrace
+1 more
Docker
Job description
- Lead and coordinate P1/P2 production incidents as an Incident Commander/Bridge Lead.
- Troubleshoot and resolve complex production issues across cloud, on-premises, and enterprise/POS applications.
- Perform detailed Root Cause Analysis (RCA) using methodologies such as 5 Whys and Fishbone.
- Implement corrective and preventive actions to improve platform stability.
- Monitor application and infrastructure health using Dynatrace, Azure Monitor, and related tools.
- Automate operational tasks using Bash or Python scripting.
- Work closely with development, infrastructure, and business teams to ensure high system availability.
- Participate in a shared on-call rotation and periodic travel as required.
Requirements
We are looking for an experienced Site Reliability Engineer (SRE) with a strong background in Production Support, Major Incident Management, and Root Cause Analysis. This is a hands-on production support role where you’ll troubleshoot critical production issues, improve system reliability, and drive automation across cloud and on-premises environments., * 5+ years of experience in Site Reliability Engineering (SRE) or Production Support
- Strong experience leading Major Incident Management for P1/P2 outages
- Hands-on experience with Root Cause Analysis (RCA) and Problem Management
- Experience troubleshooting production issues in Cloud and On-Premises environments
- Strong knowledge of:
- Dynatrace
- Azure Monitor
- Linux
- Bash and/or Python
- Kubernetes
- Docker
- Azure or GCP
- Jira
- Excellent communication and collaboration skills
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.indeed.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
LM
Luis Minvielle
over 2 years ago
LM
Luis Minvielle
Is Software Engineering Over-Saturated?
over 2 years ago
LM
Luis Minvielle
Fully Remote Software Engineer Jobs
over 2 years ago
EM
Eli McGarvie
Find a Developer Job: 12 Best Job Sites For Developers
over 3 years ago
IK
Igor Khokhriakov
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again
about 1 month ago
LM
Luis Minvielle
The 12 Best Jobs for Software Engineers
over 2 years ago