> Markdown version of [/jobs/ext/3608335-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/3608335-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** Experis - **Location:** Louisville, KY, United States - **Experience:** Expert - **Salary:** $114,400.0 - **Contract:** Permanent contract - **Skills:** Continuous Integration, DevOps, Reliability Engineering, Kubernetes - **Published:** October 7, 2026 - **Apply:** https://www.experis.com/en/job/413939/senior-site-reliability-engineer ## About the Role + 6+ years in IT infrastructure, with 3+ years focused on site reliability engineering, cloud architecture, or platform engineering. + Hands-on experience with AWS services including VPC, EC2, ECS, EKS, Fargate, Lambda, IAM, RDS, and S3. + Working experience with Microsoft Azure. + Strong expertise in Kubernetes and container orchestration, including edge or distributed deployments. + Experience with Terraform for infrastructure provisioning and GitLab CI/CD pipeline design, plus automation/scripting skills (Python, Bash, Go, or equivalent). ## Description + Own reliability for restaurant technology systems by defining what "reliable" means, setting Service Level Objectives (SLOs) and Service Level Indicators (SLIs), and using error budgets to balance reliability with delivery speed. + Build and maintain monitoring, alerting, and observability that produce meaningful signal-so teams can detect issues early and respond with confidence. + Lead blameless post-incident reviews, drive root cause analysis, and ensure permanent corrective actions are implemented across the stack. + Design and implement scalable, secure, highly available infrastructure primarily on AWS, with working knowledge of Azure, including Kubernetes-based edge systems deployed in restaurant environments. + Improve operational excellence by optimizing CI/CD pipelines (GitLab CI/CD), creating internal automation and tooling, and reducing mean time to detection (MTTD) and mean time to resolution (MTTR) through runbooks and automation. What's Needed?, + Opportunity to be a senior generalist across cloud, edge, networking, mobile deployments, and operational automation-where your decisions directly improve system health. + Clear reliability ownership: define SLOs/SLIs, measure outcomes, and continuously engineer improvements. + Hands-on impact on Kubernetes edge systems, CI/CD, observability, and incident response for restaurant technology. + Collaboration with Global Reliability Engineering (GRE), DevOps, and security teams to build secure-by-design solutions.