Site Reliability Engineer

Indotronix Avani Group
Greenwood Village, CO, United States
1 day ago
Apply on candidateportal.ceipal.com
Prepare application

Role details

Contract type
Contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
6 years minimum
Working hours
Regular working hours

Tech stack

Amazon Web Services Continuous Integration DevOps Document-Oriented Databases JSON Python (Programming Language) Node.Js NoSQL SQL Databases TypeScript Datadog ReactJS
+11 more
Istio Git Gitlab-ci Kubernetes Infrastructure Automation Frameworks Information Technology Deployment Automation Graphql Terraform Splunk Software Version Control

Job description

Join a forward-thinking engineering team as a Site Reliability Engineer, specializing in enterprise-scale experimentation and configuration management platforms. In this operations-focused role, you’ll drive platform reliability, optimize cloud architecture, and take ownership of mission-critical systems within a dynamic, hybrid work environment based in Greenwood Village, Colorado., Maintain and enhance Terraform modules to define and audit AWS infrastructure, ensuring state consistency and resolving configuration drift.

  • Operate and optimize AWS services and resources such as EKS, Helm, Istio, Aurora, DocumentDB, Redis, Amazon MQ, Route53, WAFv2, CloudFront, and S3.
  • Own and enforce deployment standards using GitLab CI/CD pipelines, including progressive promotion and pipeline-only deployment.
  • Build, deploy, and validate software releases across multiple environments; document and manage detailed release notes.
  • Right-size and scale resources to meet stringent SLAs while optimizing for cost efficiency.
  • Collaborate closely with developers and test engineers to elevate application performance.
  • Lead end-to-end monitoring and alerting using Datadog, Splunk, and related observability tools.
  • Serve as the first responder for incidents, handling mitigation, recovery, and root-cause analysis under SLA obligations.
  • Act as the subject-matter expert for infrastructure and pipeline issues, answering team questions and escalating architectural decisions.
  • Work cross-functionally with onshore and offshore teams to ensure platform stability and continuous improvement.

Requirements

6+ years of DevOps experience in large-scale, complex environments.

  • Proficiency with AWS cloud infrastructure and Terraform.
  • Strong Kubernetes expertise, including hands-on experience with containerized microservice applications.
  • Proven experience deploying with GitLab CI/CD or similar tools.
  • Advanced skills in observability and monitoring (e.g., Datadog, Splunk), including dashboard creation and alert tuning.
  • Demonstrated success in production incident triage, mitigation, and root-cause analysis under SLA constraints.
  • Solid knowledge of Git-based source control workflows.
  • Bachelor’s degree in Computer Science, Engineering, or related field, or equivalent professional experience.

Preferred Skills:

  • Familiarity with Python, Node.js, React, TypeScript, GraphQL application stacks.
  • Experience with both SQL and NoSQL/document data stores.
  • Working knowledge of Kubernetes internals, Helm, and Istio service mesh.
  • Experience with blue/green or canary/progressive deployment strategies.
  • Hands-on infrastructure cost optimization and AWS multi-account architecture exposure.
  • Master’s degree or higher in a related field.

Benefits & conditions

Hybrid work environment offering flexibility and work-life balance.

  • Opportunity to work on innovative, enterprise-scale platforms with cutting-edge cloud and DevOps technologies.
  • Collaborate with highly skilled engineering professionals in a supportive, growth-oriented culture.
  • Expand your expertise in AWS, Kubernetes, CI/CD, observability, and infrastructure automation.
  • Contract position with competitive compensation.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on candidateportal.ceipal.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:46 min

Introduction to the speaker and engineering background

Llywelyn Griffith-Swain · World Congress 2023

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:47 min

Exploring JSON, CBOR, and JOSE for data serialization

Aaron Russell · LIVE

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all