Site Reliability Engineer, GCP

Matlen Silver
Chandler, AZ, United States
12 days ago

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
4 years minimum
Compensation
$133,120.0 - $137,280.0
Working hours
Regular working hours
Job source

Tech stack

Agile Methodology Artificial Intelligence Cloud Computing Cloud Computing Security Cloud Engineering Configuration Management Continuous Integration Distributed Systems Identity and Access Management Machine Learning Reliability Engineering Cloud Services
+14 more
Systems Integration Policy as Code Data Logging Google Cloud Load Balancing Cloud Platform System Performance Testing Grafana Infrastructure as Code (IaC) Amazon Virtual Private Cloud (VPC) Deployment Automation Terraform Dynatrace Devsecops

Job description

We are seeking a Site Reliability Engineer (SRE) with strong Google Cloud Platform (GCP) and Terraform expertise to support the design, automation, and reliability of enterprise cloud infrastructure. This role will focus on building and maintaining scalable, secure, and highly available cloud environments using Infrastructure as Code (IaC), while driving automation, observability, and operational excellence. The ideal candidate has experience with Terraform Enterprise, CI/CD, DevSecOps, and cloud platform engineering in large-scale environments. Responsibilities

  • Design, build, and maintain Google Cloud Platform (GCP) infrastructure using Terraform and Infrastructure as Code (IaC).
  • Develop reusable Terraform modules and standardized infrastructure patterns for consistent cloud provisioning.
  • Implement and maintain scalable, secure, and compliant cloud environments aligned with enterprise best practices.
  • Build and enhance CI/CD pipelines for automated infrastructure deployment, testing, and validation.
  • Support Terraform Enterprise, including automated provisioning, policy enforcement, and infrastructure governance.
  • Implement automation for provisioning, configuration management, and cloud platform operations.
  • Integrate security and compliance controls into infrastructure delivery through DevSecOps and policy-as-code practices.
  • Design and improve observability solutions, including monitoring, logging, alerting, dashboards, and distributed tracing.
  • Define and support Service Level Objectives (SLOs), availability, latency, and performance metrics.
  • Participate in incident response, troubleshooting, root cause analysis, and continuous service improvement.
  • Optimize cloud infrastructure for scalability, performance, reliability, and cost efficiency.
  • Conduct capacity planning and performance testing to ensure resilient cloud services.
  • Collaborate with engineering, architecture, and security teams to deliver reliable and secure cloud platforms.
  • Continuously identify opportunities to automate manual processes and improve operational efficiency.
  • Evaluate emerging cloud technologies and automation tools, including AI/ML capabilities where applicable.

Requirements

  • 4-8+ years of experience in Site Reliability Engineering, Cloud Infrastructure Engineering, Platform Engineering, or Cloud Operations.
  • Hands-on experience with Google Cloud Platform (GCP).
  • Strong experience with Terraform and Infrastructure as Code (IaC); Terraform Enterprise experience is highly preferred.
  • Experience developing reusable Terraform modules and managing infrastructure through code.
  • Experience building and maintaining CI/CD pipelines for infrastructure deployments.
  • Knowledge of DevSecOps practices and integrating security into automated deployment workflows.
  • Strong understanding of GCP services, including VPC networking, IAM, load balancing, and cloud architecture fundamentals.
  • Experience with policy-as-code, governance, and cloud compliance frameworks.
  • Experience with monitoring, logging, and observability tools.
  • Hands-on experience with incident management, troubleshooting, and root cause analysis in cloud environments.
  • Experience designing dashboards, alerts, and monitoring strategies aligned with SLOs.
  • Strong scripting and automation skills.
  • Excellent problem-solving, analytical, and communication skills.
  • Ability to work effectively within cross-functional Agile teams.

Preferred Qualifications

  • Experience with Terraform Enterprise administration.
  • Experience supporting large-scale enterprise cloud environments.
  • Familiarity with AI/ML-driven operational automation and observability.
  • Experience implementing cloud governance and compliance standards.
  • Experience with cloud-native monitoring and distributed systems.

Benefits & conditions

$64 - $66 an hour - Contract

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

1:01 min

Connecting frontend application performance to user retention and revenue

Dani Coll Dani Coll · WWC 2025

1:15 min

Key lessons learned from implementing automated mobile DevSecOps

Moataz Nabil Moataz Nabil · LIVE

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet · LIVE

3:50 min

Scaling shift left practices within large engineering organizations

Chris Riley · WWC 2021

12:08 min

Comparing Keptn orchestration capabilities against alternative software operators

Thomas Schütz · LIVE

Videos

See all

Related articles

See all