> Markdown version of [/jobs/ext/2691689-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/2691689-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** Smiths Medical - **Location:** United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Amazon Web Services, Amazon Cloudfront, Amazon Elastic Compute Cloud, Amazon S3, ARM Architecture, Bash Shell, Ubuntu (Operating System), Cloud Computing, Cloud Computing Security, Databases, Continuous Delivery, Data Stores, Linux, DevOps, Identity and Access Management, Python (Programming Language), Linux System Administration, Network Segmentation, Reliability Engineering, Software Deployment, Zabbix, Datadog, Delivery Pipeline, Spring-boot, Cloudformation, Fastapi, Containerization, Git Flow, Kubernetes, Infrastructure Automation Frameworks, AWS Fargate, Cloud Migration, Functional Programming, Cloudwatch, Api Gateway, Serverless Computing, Docker, Jenkins - **Published:** September 3, 2026 - **Apply:** https://www.careerarc.com/job-listings/52994845/apply?campaign_id=39006&src=12 ## About the Role * Deep hands-on expertise with AWS core services (networking, compute, serverless, and database technologies) and CloudFormation IaC automation. * Strong Linux administration skills (primarily Ubuntu) along with proficiency in Python and Bash scripting for operational automation. * Experience with containerization technologies (Docker, ECS/Fargate) and familiarity with modern Kubernetes ecosystems (EKS, Helm, ArgoCD). * Solid understanding of observability tools (Datadog, CloudWatch, Zabbix) and CI/CD pipelines (Jenkins, CodePipeline, Git workflows). * Knowledge of cloud security best practices, access management (IAM), and compliance frameworks within regulated sectors (HIPAA/HiTrust). * Proven diagnostic, incident-management, and analytical troubleshooting skills for complex microservices architectures., * Must be at least 18 years of age. * High School Diploma required. * Bachelor's degree from an accredited college or university is required. * 7+ years of hands-on experience in AWS Cloud Engineering, DevOps, Site Reliability Engineering (SRE), or Infrastructure Engineering. * Practical background supporting production workloads in Linux/AWS environments, reading application logs, and making minor code fixes. * Direct experience participating in on-call rotations and incident response protocols. * Prior experience in the healthcare industry maintaining HIPAA/HiTrust-compliant infrastructure. ## Description We are seeking a Senior Site Reliability Engineer, CloudOps to support, scale, and optimize a multi-account AWS environment hosting healthcare-oriented applications and analytics platforms. In this role, you will bridge infrastructure engineering, operational reliability, production support, and cloud modernization initiatives across complex microservices architectures. The ideal candidate brings strong AWS expertise, solid Linux administration skills, and a proven track record of managing production systems in HIPAA/HiTrust regulated environments. You will participate in incident response, on-call rotations, and continuous deployment workflows while helping drive our transition toward containerized and Kubernetes-based platforms. This collaborative position is built for an analytical engineer who excels at resolving production incidents, partnering with developers, and continuously elevating operational excellence., * Manage, maintain, and troubleshoot a multi-account AWS Organization environment (35+ accounts) and core services, including EC2, ECS/Fargate, Lambda, S3, CloudFront, API Gateway, and Aurora/RDS databases. * Support production deployments, CI/CD pipelines (Jenkins, AWS CodePipeline), and infrastructure automation using Python, Bash, and AWS CloudFormation. * Monitor system health and performance using Datadog, CloudWatch, and Zabbix; investigate alerts, execute root-cause analysis, and refine monitoring coverage to reduce operational noise. * Participate in a shared on-call rotation, managing incident response and performing failover/recovery validation for production applications and data stores. * Maintain HIPAA/HiTrust compliance and security posture by managing tools like Prisma/Cortex Cloud, Security Hub, and GuardDuty, while enforcing proper IAM policies and network segmentation. * Support Java (Spring Boot) and Python applications running in containers, assisting developers during investigations and preparing for future Kubernetes (EKS) modernization initiatives. ## Related Videos - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Best Countries for Software Engineers](https://www.wearedevelopers.com/magazine/267-best-countries-for-software-engineers) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers)