> Markdown version of [/jobs/ext/739694-site-reliability-devops-engineer](https://www.wearedevelopers.com/jobs/ext/739694-site-reliability-devops-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability / DevOps Engineer - **Company:** eClerx LLC - **Location:** Raleigh, NC, United States - **Salary:** $120,000.0 - $137,500.0 - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Application Performance Management, Automation of Tests, Microsoft Azure, Bash Shell, Cloud Engineering, Continuous Integration, Linux, DevOps, Distributed Systems, Elasticsearch, Github, Monitoring of Systems, Python (Programming Language), Enterprise Messaging Systems, Windows PowerShell, Reliability Engineering, Data Streaming, Datadog, Scripting, Enterprise Software Applications, Cloud Platform System, Delivery Pipeline, Grafana, Mttr, Cloudformation, Gitlab-ci, Kubernetes, Bicep, Apache Kafka, Terraform, Docker, Jenkins - **Published:** June 29, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=daa7abd37eec3291 ## About the Role Do you have experience in Tooling?, * 5+ years of experience as a Site Reliability Engineer, DevOps Engineer, or similar role. * 5+ years of work experience with Public Cloud (Azure (preferred)or AWS) * 3+ years of hands-on experience with observability platforms such as Datadog, Elasticsearch, Grafana, or similar solutions. * 5+ years of experience with scripting languages like Python, Bash, Powershell, etc. * 3+ years of experience with containerization and orchestration technologies, including Docker and Kubernetes. * 2+ years of experience developing and managing CI/CD pipelines using tools such as Azure DevOps, GitLab CI/CD, GitHub Actions, Jenkins, or similar. * 2+ years of experience with Infrastructure-as-Code (IaC) tools such as Terraform, Azure Bicep, AWS CloudFormation, or equivalent technologies. * 1+ years of experience using site reliability and resilience testing tools such as Gremlin, Chaos Mesh, or similar platforms. * Proven experience leveraging observability best practices, end-user monitoring, application performance monitoring, and infrastructure monitoring solutions. * Experience with event streaming and messaging platforms such as Kafka or Azure Event Hubs. * Strong understanding of Linux operating systems and administration. * Preferred Qualifications + Kubernetes certification + Cloud platform certifications (Azure, AWS, or GCP). + Experience working in Azure environments and/or Azure DevOps. + Experience implementing and managing Datadog or other modern observability platforms. + Experience supporting enterprise-scale applications within financial services, capital markets, fintech, or other highly regulated industries. In the US, the target base salary for this role is $120,000-$137,500. Compensation is based on a range of factors that include relevant experience, knowledge, skills, other job-related qualifications, and geography. We expect the majority of candidates who are offered roles at our company to fall throughout the range based on these factors ## Description eClerx is seeking a motivated SRE/DevOps Engineer with strong observability experience to join our growing Platform Engineering team. This team is responsible for managing cloud infrastructure, advancing DevOps practices, improving platform reliability, and supporting highly available enterprise applications. The ideal candidate will have a deep understanding of cloud-native architectures, distributed systems, CI/CD automation, observability frameworks, and site reliability engineering principles. This individual will play a key role in improving platform resilience, operational efficiency, and system performance across a modern cloud-based technology ecosystem. Responsibilities * Design, implement, and enhance system observability and monitoring solutions. * Monitor system performance, create incident response plans, and implement observability practices to gain deeper insights into system behavior. * Define, implement, and monitor Service Level Objectives (SLOs) and Service Level Indicators (SLIs). * Improve platform reliability, scalability, and resiliency. * Conduct post-incident reviews and implement corrective actions to prevent recurring issues. * Partner with engineering teams to implement observability tooling and leverage telemetry data to troubleshoot and resolve incidents. * Utilize observability and event management capabilities to improve key operational metrics, including Mean Time to Detect (MTTD) and Mean Time to Restore (MTTR). * Continuously optimize infrastructure, architecture, automation, CI/CD processes, and operational workflows. * Collaborate closely with software engineers to ensure applications are designed and deployed following DevOps and reliability best practices. * Participate in a rotating on-call schedule, including support for production releases and critical incidents outside normal business hours when required. ## Related Videos - [What Developers Get Wrong About Application Quality](https://www.wearedevelopers.com/videos/233-what-developers-get-wrong-about-application-quality) - [Back(end) to the Future: Embracing the continuous Evolution of Infrastructure and Code](https://www.wearedevelopers.com/videos/440-back-end-to-the-future-embracing-the-continuous-evolution-of-infrastructure-and-code) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Navigating the Corporate Jungle: Life as a Developer in a large Company](https://www.wearedevelopers.com/videos/621-navigating-the-corporate-jungle-life-as-a-developer-in-a-large-company) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) ## Related Articles - [DevOps Engineer Salary [2023]](https://www.wearedevelopers.com/magazine/203-devops-engineer-salary-2023) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Top-Paying Tech Jobs (with Salaries)](https://www.wearedevelopers.com/magazine/372-top-paying-tech-jobs-with-salaries) - [How Much Does a Software Engineer Make? Realistic Software Engineering Salaries](https://www.wearedevelopers.com/magazine/425-how-much-does-a-software-engineer-make-realistic-software-engineering-salaries) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers)