> Markdown version of [/jobs/ext/1881465-sr-site-reliability-engr](https://www.wearedevelopers.com/jobs/ext/1881465-sr-site-reliability-engr). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Sr Site Reliability Engr - **Company:** Optum, Inc - **Location:** Eden Prairie, MN, United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Algorithm Design, Amazon Web Services, Amazon Elastic Compute Cloud, Amazon S3, Microsoft Azure, Cloud Computing, Cloud Engineering, Data Structures, DevOps, Github, Monitoring of Systems, Identity and Access Management, Internet Protocol, Open Web Application Security, Public Key Infrastructure, Reliability Engineering, Cloud Services, Prometheus, Software Engineering, Systems Integration, Data Logging, Cloud Platform System, Amazon Virtual Private Cloud (VPC), Gitlab, Cloudformation, Kubernetes, Infrastructure Automation Frameworks, Information Technology, Route53, Cloudwatch, Restful APIs, Terraform, Splunk, Software Version Control, Dynatrace - **Published:** July 31, 2026 - **Apply:** https://jobs.localjobnetwork.com/apply/add/87880679/1 ## About the Role * 6+ years of experience working within a cloud engineer/SRE role * Solid knowledge of AWS services (ex. VPC, EC2, S3, ECS, Cloudformation, Lambda, EKS, RDS, ELB, Route53, RedShift) * Expert knowledge and hands on production experience in Kubernetes (EKS or AKS) cluster setup and management required. * Experience with infrastructure as code (IaC) tools like Terraform. * Experience with Kubernetes deployment tools like Helm, ArgoCD, Flux * Solid awareness of networking and internet protocols. * Understanding of identity and access management (IAM) * Experience supporting infrastructure in production cloud environments. * Knowledge of Encryption (KMS), Public Key Infrastructure (PKI), understanding of OWASP * Experience working with RESTful services * Experience supporting environments adhering compliance standards like FedRAMP and NIST (800-171|53) * Some experience with monitoring tools (CloudWatch, VPC Flow Logs, Splunk, Dynatrace, Graphana, Prometheus). * Familiarity with IDEs and Source Control tools like Azure DevOps, Github or Gitlab. * Be part of 24/7 on-call rotation * United States Citizenship * If you are offered this position, you will be required to provide extensive personal information to obtain and maintain a suitability or determination of eligibility for a Confidential/Secret or Top Secret security clearance as a condition of your employment, * Bachelor's Degree in Computer Science, Information Technology, Software Engineering, Math, Physics * Master's Degree with coursework focused on advanced algorithms, mathematics in computing, data structures or related field * Expert knowledge of deploying Production grade applications in AWS * Demonstrate passion about infrastructure automation *All employees working remotely will be required to adhere to UnitedHealth Group's Telecommuter Policy ## Description Optum is a global organization that delivers care, aided by technology to help millions of people live healthier lives. The work you do with our team will directly improve health outcomes by connecting people with the care, pharmacy benefits, data and resources they need to feel their best. Here, you will find a culture guided by inclusion, talented peers, comprehensive benefits and career development opportunities. Come make an impact on the communities we serve as you help us advance health optimization on a global scale. Join us to start Caring. Connecting. Growing together. The Site Reliability Engineer will architect, develop, and maintain Optum Serve's cloud environment in both the commercial and government AWS cloud. The role will work closely with software engineers, architects, and DevOps engineers to architect and maintain a secure, resilient and high performance cloud infrastructure. To support this mission, OSIT has initiated a multi year modernization program aimed at updating and enhancing enterprise technology systems in accordance with modern design standards. You'll enjoy the flexibility to work remotely * from anywhere within the U.S. as you take on some tough challenges. For all hires in the Minneapolis or Washington, D.C. area, you will be required to work in the office a minimum of four days per week., * Build, maintain, and operate IaaS and PaaS infrastructure in AWS commercial and government clouds * Work closely with dev teams to identify and measure SLOs, SLAs and SLIs * Act a solid contributor to development of platform services including architecture, provisioning, configuration, deployment, and support * Perform integrations with central logging, metrics dashboards, instrumentation, incident monitoring and management * Build/integrate/administer systems and tools that enable engineering teams to observe their applications in production with autonomy (Dashboards, APMs). * Support software and/or cloud-infrastructure in an on-call rotation basis * Assist with identification and remediation of technical problems at the root cause by continuously implementing automation, self-healing, and real-time monitoring to production systems * Maintain and improve operational tooling, frameworks, * Build frameworks that test the performance and resiliency of our platform services/tools * Automate alerts for metrics on performance, cost, vulnerabilities, risk, compliance violations * Improve processes and champion automation of any manual items around support You'll be rewarded and recognized for your performance in an environment that will challenge you and give you clear direction on what it takes to succeed in your role as well as provide development for other roles you may be interested in. ## Related Videos - [WeAreDevelopers LIVE - Modern DevOps for IoT Devices and More](https://www.wearedevelopers.com/videos/1805-wearedevelopers-live-modern-devops-for-iot-devices-and-more) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Enabling automated 1-click customer deployments with built-in quality and security](https://www.wearedevelopers.com/videos/83-enabling-automated-1-click-customer-deployments-with-built-in-quality-and-security) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [The Best Job Search Websites of 2025](https://www.wearedevelopers.com/magazine/368-the-best-job-search-websites-of-2025) - [Best Countries for Software Engineers](https://www.wearedevelopers.com/magazine/267-best-countries-for-software-engineers)