> Markdown version of [/jobs/ext/1506989-senior-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/1506989-senior-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Site Reliability Engineer - **Company:** Prince Perelson and Associates, L.L.C. - **Location:** Salt Lake City, UT, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Cloud Engineering, Data Infrastructure, Software Debugging, DevOps, Distributed Systems, Amazon DynamoDB, Electronic Design Automation, PostgreSQL, MongoDB, Reliability Engineering, Prometheus, Software Engineering, Grafana, Kubernetes, Cloudwatch, Terraform, Splunk, New Relic (SaaS), Dynatrace - **Published:** July 30, 2026 - **Apply:** https://www.juju.com/job/00000000gkh9qz ## About the Role + 7+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or a similar infrastructure-focused engineering role. + Strong experience supporting production workloads in AWS and Kubernetes environments. + Demonstrated success implementing SLOs, SLIs, and reliability engineering best practices. + Experience designing observability solutions that provide actionable insights while minimizing operational noise. + Strong Infrastructure as Code experience, preferably with Terraform. + Experience automating operational workflows and building reusable engineering standards. + Familiarity with distributed systems and modern data platforms including technologies such as PostgreSQL, MongoDB, DynamoDB, or similar. + Experience with multiple observability platforms such as Prometheus, Grafana, New Relic, Splunk, CloudWatch, ELK, or comparable technologies. + Strong troubleshooting and debugging skills across distributed production environments. + Experience leading incident response and driving continuous operational improvements. + Excellent communication skills with the ability to influence engineering teams through collaboration, technical leadership, and practical solutions. ## Description Are you passionate about building highly reliable, scalable cloud platforms that power mission-critical applications? We're partnering with an innovative technology company that's investing heavily in platform reliability, automation, and observability. This is an opportunity to have a significant impact on the engineering organization by shaping reliability standards, improving developer experience, and helping build resilient systems that support a rapidly growing platform. You'll work at the intersection of software engineering and infrastructure, partnering with engineering teams to improve system performance, automate operations, and establish best practices across Kubernetes-based services. If you enjoy solving complex distributed systems challenges and creating tools that make engineers more productive, we'd love to talk. What You'll Do + Define and evolve Service Level Objectives (SLOs) and Service Level Indicators (SLIs) that align platform performance with customer and business expectations. + Build and standardize monitoring, alerting, and observability practices across engineering teams using Infrastructure as Code (Terraform preferred). + Develop scalable observability solutions leveraging metrics, logs, traces, and distributed tracing technologies such as OpenTelemetry. + Evaluate and implement modern observability platforms and integrations, leading proof-of-concepts and defining adoption strategies. + Establish reliability standards for Kubernetes-based applications, including scaling strategies, deployment safety, resource optimization, dashboards, and alerting. + Design automation that reduces manual effort, streamlines operations, and improves incident response and recovery. + Lead high-severity incident response efforts, facilitate postmortems, and drive long-term reliability improvements through measurable action plans. + Participate in an on-call rotation while continually improving monitoring quality and reducing unnecessary alert noise. + Partner with software engineers and platform teams to build resilient, scalable cloud infrastructure and operational best practices. ## Related Videos - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [40 Minutes to Build a Serverless COVID-19 REST and GraphQL APIs](https://www.wearedevelopers.com/videos/208-40-minutes-to-build-a-serverless-covid-19-rest-and-graphql-apis) - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) - [NoSQL Data Modeling for Front-end Developers](https://www.wearedevelopers.com/videos/297-nosql-data-modeling-for-front-end-developers) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [The Best Software Developer Blogs to Read](https://www.wearedevelopers.com/magazine/156-the-best-software-developer-blogs-to-read) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs)