> Markdown version of [/jobs/ext/2852390-sr-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/2852390-sr-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Sr. Site Reliability Engineer - **Company:** Amazon.com, Inc. - **Location:** Bellevue, WA, United States - **Experience:** Expert - **Salary:** $73,450.0 - $132,775.0 - **Contract:** Permanent contract - **Skills:** Microsoft Windows, Microsoft Azure, Bash Shell, Cloud Computing, Continuous Integration, Linux, DevOps, Disaster Recovery, Monitoring of Systems, Identity and Access Management, Python (Programming Language), Performance Tuning, Windows PowerShell, Reliability Engineering, Prometheus, Data Logging, System Availability, Grafana, Software Troubleshooting, Infrastructure as Code (IaC), Containerization, Gitlab-ci, Kubernetes, Information Technology, Deployment Automation, Bicep, Cloud Migration, Terraform, Docker, Jenkins - **Published:** September 11, 2026 - **Apply:** https://www.careerjet.com/jobad/usc1ae77ec1c8310a0056dc63120a1be08 ## About the Role * Bachelor's degree in Computer Science, Engineering, IT, or related field, or equivalent practical experience. * 4+ years of experience in SRE, DevOps, Cloud Infrastructure, or related roles. * Strong production experience with Microsoft Azure. * Hands-on experience with Kubernetes and Docker. * Strong troubleshooting and incident-response capabilities. * Experience supporting cloud infrastructure and production workloads., * Experience with Tencent Kubernetes Engine (TKE) or other managed Kubernetes platforms. * Experience migrating workloads from cloud VMs to Kubernetes. * Experience containerizing legacy applications. * Experience with Azure-to-Kubernetes cloud migration projects. * Azure certifications such as AZ-104, AZ-400, or AZ-500. * Kubernetes certifications such as CKA or CKAD. * Experience with capacity planning, high availability, disaster recovery, and performance optimization., Are you experienced in managing Windows and Linux systems with a passion for building highly-available infrastructure at massive scale? Join us in driving cloud innovation and auto… + 13 hours ago ## Description * Monitor, troubleshoot, and resolve production incidents impacting availability, latency, and performance. * Provision and manage Azure infrastructure, including VMs, networking, storage, and IAM. * Support the Azure-to-TKE migration, including containerization, Kubernetes deployments, testing, and cutover activities. * Design and maintain CI/CD pipelines for automated deployment and testing. * Build and maintain monitoring, dashboards, alerting, logging, and health-check solutions. * Implement Infrastructure as Code (IaC) and automation to reduce manual operational effort. * Support capacity planning, performance optimization, reliability, and security initiatives. * Collaborate with engineering and infrastructure teams on incident response and migration readiness. * Participate in on-call support as required., Cloud: * Microsoft Azure * Azure VMs * Azure Networking * Azure Storage * Azure IAM * Azure Monitor Containers & Kubernetes: * Kubernetes * Docker * Tencent Kubernetes Engine (TKE) preferred * Containerization and Kubernetes deployments CI/CD: * Azure DevOps * Jenkins * GitLab CI/CD Infrastructure as Code: * Terraform * ARM Templates * Bicep Scripting & Automation: * Python * Bash * PowerShell Monitoring & Observability: * Prometheus * Grafana * Azure Monitor * Logging, alerting, dashboards, and health checks ## Related Videos - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Back(end) to the Future: Embracing the continuous Evolution of Infrastructure and Code](https://www.wearedevelopers.com/videos/440-back-end-to-the-future-embracing-the-continuous-evolution-of-infrastructure-and-code) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated)