> Markdown version of [/jobs/ext/3588429-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/3588429-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** Indotronix Avani Group - **Location:** Greenwood Village, CO, United States - **Experience:** Expert - **Contract:** Contract - **Skills:** Amazon Web Services, Continuous Integration, DevOps, Document-Oriented Databases, JSON, Python (Programming Language), Node.Js, NoSQL, SQL Databases, TypeScript, Datadog, ReactJS, Istio, Git, Gitlab-ci, Kubernetes, Infrastructure Automation Frameworks, Information Technology, Deployment Automation, Graphql, Terraform, Splunk, Software Version Control - **Published:** October 5, 2026 - **Apply:** https://candidateportal.ceipal.com/job-details/Qa9bIPA7EQ4h35Md6cX7AzHipUr9EIwjGSaEbOytyvk ## About the Role 6+ years of DevOps experience in large-scale, complex environments. - Proficiency with AWS cloud infrastructure and Terraform. - Strong Kubernetes expertise, including hands-on experience with containerized microservice applications. - Proven experience deploying with GitLab CI/CD or similar tools. - Advanced skills in observability and monitoring (e.g., Datadog, Splunk), including dashboard creation and alert tuning. - Demonstrated success in production incident triage, mitigation, and root-cause analysis under SLA constraints. - Solid knowledge of Git-based source control workflows. - Bachelor's degree in Computer Science, Engineering, or related field, or equivalent professional experience. Preferred Skills: - Familiarity with Python, Node.js, React, TypeScript, GraphQL application stacks. - Experience with both SQL and NoSQL/document data stores. - Working knowledge of Kubernetes internals, Helm, and Istio service mesh. - Experience with blue/green or canary/progressive deployment strategies. - Hands-on infrastructure cost optimization and AWS multi-account architecture exposure. - Master's degree or higher in a related field. ## Description Join a forward-thinking engineering team as a Site Reliability Engineer, specializing in enterprise-scale experimentation and configuration management platforms. In this operations-focused role, you'll drive platform reliability, optimize cloud architecture, and take ownership of mission-critical systems within a dynamic, hybrid work environment based in Greenwood Village, Colorado., Maintain and enhance Terraform modules to define and audit AWS infrastructure, ensuring state consistency and resolving configuration drift. - Operate and optimize AWS services and resources such as EKS, Helm, Istio, Aurora, DocumentDB, Redis, Amazon MQ, Route53, WAFv2, CloudFront, and S3. - Own and enforce deployment standards using GitLab CI/CD pipelines, including progressive promotion and pipeline-only deployment. - Build, deploy, and validate software releases across multiple environments; document and manage detailed release notes. - Right-size and scale resources to meet stringent SLAs while optimizing for cost efficiency. - Collaborate closely with developers and test engineers to elevate application performance. - Lead end-to-end monitoring and alerting using Datadog, Splunk, and related observability tools. - Serve as the first responder for incidents, handling mitigation, recovery, and root-cause analysis under SLA obligations. - Act as the subject-matter expert for infrastructure and pipeline issues, answering team questions and escalating architectural decisions. - Work cross-functionally with onshore and offshore teams to ensure platform stability and continuous improvement. ## Related Videos - [My journey into DevOps world - How it all started!](https://www.wearedevelopers.com/videos/545-my-journey-into-devops-world-how-it-all-started) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Tips and Tricks for Working with JSON](https://www.wearedevelopers.com/videos/1229-tips-and-tricks-for-working-with-json) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [The Best Software Developer Blogs to Read](https://www.wearedevelopers.com/magazine/156-the-best-software-developer-blogs-to-read) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated)