Site Reliability Engineer

Largeton INC
Blue Ash, OH, United States
3 months ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Temporary to permanent
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Compensation
$82,000.0 - $113,000.0
Working hours
Regular working hours

Tech stack

Microsoft Azure Bash Shell Linux Python (Programming Language) Reliability Engineering Runbook Scripting Grafana Containerization Kubernetes Dynatrace Docker

Job description

Participate in a multi-stage interview process (automated prescreen, hiring manager, team, and potential onsite interview).

  • Must be flexible with travel (primarily Cincinnati and Louisville initially; expands as sites go live).
  • Participate in off-hours on-call rotation (escalation point, not initial response; 1 week every 6 weeks).
  • Lead and participate in Major Incident Management, including triage and cross-team response.
  • Conduct Root Cause Analysis (RCA) and track corrective actions.
  • Collaborate with engineering teams to define SLOs/SLIs and enhance observability using Dynatrace and Azure.
  • Create playbooks, runbooks, and postmortem documentation.
  • Requires 3+ years in SRE, incident management, or production support roles., Work Schedule Standard (Mon-Fri) Environmental Conditions Able to lift 40 lbs. without assistance, Adherence to all Good Manufacturing Practices (GMP) Safety Standards, Cleanroom…
  • 22 days ago

Requirements

Must have experience as incident commander and in bridge/war-room facilitation.

  • Strong knowledge of observability tools (Dynatrace, Azure Monitor) and containerization (Kubernetes, Docker).
  • Proficient in Linux and scripting languages (BASH, Python).
  • Familiarity with enterprise Point of Sale (POS) systems is preferred.

Benefits & conditions

6-month contract at $50/hr., + $82,000-113,000 per year Changing lives. Building Careers. Joining us is a chance to do important work that creates change and shapes the future of healthcare. Thinking differently is what we do best. To…

  • 10 days ago + *

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

8:02 min

Integrating service level objectives into incident management

Diana Todea · LIVE

2:50 min

Introduction and the value of runbooks

Hila Fish · World Congress 2023

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

1:32 min

Structuring automated incident workflows between runbooks and raw models

Aram Hakobyan Aram Hakobyan +1 · World Congress 2026 Europe

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

Videos

See all

Related articles

See all