> Markdown version of [/jobs/ext/1281058-devops-site-reliability-engineer-sre](https://www.wearedevelopers.com/jobs/ext/1281058-devops-site-reliability-engineer-sre). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # DevOps / Site Reliability Engineer (SRE) - **Company:** THE NEXTAMP LLC - **Location:** United States (Remote available) - **Experience:** Experienced - **Salary:** $72,800.0 - $104,000.0 - **Contract:** Permanent contract - **Skills:** Kubernetes Security, Amazon Web Services, Microsoft Azure, Backup Devices, Bash Shell, Cloud Computing, DevOps, Disaster Recovery, Github, Monitoring of Systems, Python (Programming Language), Key Management, Linux System Administration, Windows PowerShell, Reliability Engineering, Site Reliability Engineering Practices, Prometheus, Software Vulnerability Management, Datadog, Data Logging, Scripting, Istio, System Availability, Delivery Pipeline, Grafana, Cloudformation, Containerization, Gitlab-ci, Kubernetes, Infrastructure Automation Frameworks, Hashicorp, Linkerd (Service Mesh), Cloudwatch, Terraform, Splunk, Docker, Jenkins - **Published:** July 15, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=fe7dee874f8154a9 ## About the Role * 4+ years of experience in DevOps or Site Reliability Engineering (SRE). * Strong experience building and managing CI/CD pipelines using tools such as Jenkins, GitHub Actions, Azure DevOps, or GitLab CI. * Hands-on experience with Docker and Kubernetes. * Strong knowledge of Infrastructure as Code (Terraform, CloudFormation, or similar). * Experience with AWS or Azure cloud platforms. * Experience with monitoring and observability tools such as Prometheus, Grafana, CloudWatch, Datadog, Splunk, ELK, or Azure Monitor. * Good understanding of Linux system administration, networking, and security best practices. * Experience with scripting using Bash, Python, or PowerShell. * Strong troubleshooting and production incident management skills. Preferred Skills * Experience with Helm, ArgoCD, or FluxCD. * Knowledge of container security and vulnerability management. * Familiarity with service mesh technologies (Istio, Linkerd). * Experience with secrets management tools such as HashiCorp Vault or AWS Secrets Manager. * Knowledge of high availability, disaster recovery, and backup strategies. Nice to Have * Experience supporting 24x7 production environments. * Exposure to SRE practices such as SLOs, SLIs, error budgets, and reliability engineering. * Experience implementing cost optimization and cloud governance initiatives. * Relevant certifications such as AWS Certified DevOps Engineer, AWS Solutions Architect, Certified Kubernetes Administrator (CKA), or Microsoft Azure DevOps Engineer Expert. ## Description We are looking for a highly skilled DevOps / Site Reliability Engineer (SRE) to build, automate, and maintain reliable, scalable, and secure cloud infrastructure. The ideal candidate should have hands-on experience with CI/CD pipelines, containerization, infrastructure as code, cloud platforms, and production operations., * Design, implement, and maintain CI/CD pipelines to enable reliable and automated software deployments. * Build, manage, and optimize containerized applications using Docker and Kubernetes. * Provision and manage cloud infrastructure using Infrastructure as Code (IaC) tools such as Terraform or CloudFormation. * Deploy, monitor, and maintain applications on AWS or Azure. * Ensure high availability, scalability, security, and reliability of production environments. * Monitor application and infrastructure health, respond to incidents, and perform root cause analysis (RCA). * Automate operational tasks to improve efficiency and reduce manual effort. * Collaborate with development teams to improve deployment processes and application reliability. * Implement monitoring, logging, alerting, and observability best practices. ## Related Videos - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Rate-limiting using eBPF and Istio: How to protect your SaaS customers from themselves](https://www.wearedevelopers.com/videos/100220-rate-limiting-using-ebpf-and-istio-how-to-protect-your-saas-customers-from-themselves) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) - [My journey into DevOps world - How it all started!](https://www.wearedevelopers.com/videos/545-my-journey-into-devops-world-how-it-all-started) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [DevOps Engineer Salary [2023]](https://www.wearedevelopers.com/magazine/203-devops-engineer-salary-2023) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)