Site Reliability Engineer (SRE)

Insight Global
San Diego, CA, United States
11 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
6 years minimum
Compensation
$116,480.0 - $147,680.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Amazon Web Services Microsoft Azure Cloud Computing Cloud Engineering Continuous Integration DevOps Python (Programming Language) Windows PowerShell Reliability Engineering Software Engineering TypeScript
+12 more
Datadog AWS Cdk Cloud Platform System Large Language Models AWS Lambda Cloudformation Build Management Kubernetes Infrastructure Automation Frameworks Terraform Serverless Computing Golang

Job description

We are seeking a Senior Site Reliability Engineer (SRE) to join a high-impact Platform Engineering team focused on building scalable cloud infrastructure, reusable platform capabilities, and automation frameworks that enable software engineers to develop and deploy applications at scale. This is not a traditional operations role-the team focuses on platform engineering, Infrastructure as Code, developer enablement, and reliability through software and automation., You will design and build reusable cloud infrastructure across AWS and Azure, develop Infrastructure as Code from the ground up, create automation and self-service platform capabilities, enhance observability using Datadog, and contribute to modern CI/CD practices. The ideal candidate has a strong software engineering mindset and enjoys building platforms that improve engineering productivity while supporting highly available production environments.

Requirements

  • 6+ years of experience in Platform Engineering, Site Reliability Engineering (SRE), DevOps, or Cloud Engineering.
  • Hands-on experience designing, building, and supporting production cloud environments across both AWS and Azure.
  • Strong experience building reusable Infrastructure as Code using Terraform, including custom modules, shared frameworks, and cloud platform components (AWS CDK and/or CloudFormation preferred).
  • Experience designing and supporting Kubernetes-based platforms in production environments.
  • Strong software engineering mindset with experience building automation using languages such as Go, Python, PowerShell, TypeScript, or similar.
  • Strong experience with Datadog (or similar observability platforms), including monitoring, troubleshooting production issues, and improving platform reliability.
  • Experience building modern CI/CD pipelines and developer enablement capabilities through platform engineering.

Nice to Have Skills & Experience

  • Experience leveraging AI tools, AI agents, or LLM-powered workflows to improve engineering productivity or platform automation.
  • Experience working in regulated industries such as healthcare or medical devices.
  • Experience with serverless technologies (AWS Lambda, Azure Functions).
  • AWS, Azure, Terraform, or Kubernetes certifications.

Benefits & conditions

Benefit packages for this role will start on the 1st day of employment and include medical, dental, and vision insurance, as well as HSA, FSA, and DCFSA account options, and 401k retirement account access with employer matching. Employees in this role are also entitled to paid sick leave and/or other paid time off as provided by applicable law.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on jobs.insightglobal.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

3:35 min

Defining a serverless architecture using AWS CDK

Raphael Manke Raphael Manke · WWC 2023

3:50 min

Navigating specialized roles and toolsets across engineering teams

Nele Uhlemann · WWC 2023

1:33 min

Case study on adopting Kubernetes and Golang effectively

Andrew Holway · LIVE

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

Videos

See all

Related articles

See all