Senior Site Reliability Engineer (SRE) - eCommerce & Google Cloud Platform (GCP)

Cognizant Technology Solutions Corporation
Phoenix, AZ, United States
15 days ago
Apply on www.phoenixjobs.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Compensation
$50,000.0
Working hours
Regular working hours

Tech stack

Amazon Elastic Compute Cloud Application Performance Management Bash Shell Cloud Computing Cloud Engineering Cloud Storage DevOps Distributed Systems Github Python (Programming Language) Linux System Administration Systems Development Life Cycle
+17 more
Reliability Engineering Prometheus Web Platforms Datadog Google Cloud Cloud Monitoring Grafana Reliability of Systems Infrastructure as Code (IaC) Git Kubernetes Api Design Terraform Splunk Docker Jenkins Microservices

Job description

We are seeking an experienced Senior Site Reliability Engineer (SRE) to support and enhance large-scale eCommerce platforms hosted on Google Cloud Platform (GCP). The ideal candidate will be responsible for ensuring the reliability, scalability, performance, and security of mission-critical production environments while driving automation and operational excellence., * Design, implement, and maintain highly available, resilient, and scalable cloud infrastructure on GCP.

  • Monitor application and platform health, proactively identify issues, and lead incident response and root cause analysis.
  • Develop and maintain Infrastructure as Code (IaC) solutions using tools such as Terraform.
  • Automate operational processes, deployments, monitoring, and remediation workflows.
  • Collaborate with development, architecture, security, and operations teams to improve system reliability and performance.
  • Manage CI/CD pipelines and support DevOps best practices across the software delivery lifecycle.
  • Optimize application performance, capacity planning, and cost management within GCP environments.
  • Establish and maintain SLIs, SLOs, and error budgets to improve service reliability.
  • Support on-call rotations and lead troubleshooting efforts for critical production incidents.

Requirements

  • 8+ years of experience in Site Reliability Engineering, Cloud Engineering, or DevOps roles.
  • Strong hands-on experience with Google Cloud Platform (GCP), including services such as GKE, Compute Engine, Cloud Storage, Cloud Monitoring, and Cloud Operations Suite.
  • Experience supporting high-volume eCommerce or customer-facing digital platforms.
  • Proficiency in Kubernetes, Docker, and container orchestration technologies.
  • Strong scripting and automation skills using Python, Bash, or similar languages.
  • Experience with Terraform, Git, Jenkins, GitHub Actions, or comparable CI/CD tools.
  • Deep understanding of Linux systems administration, networking, and distributed systems.
  • Experience with observability tools including Prometheus, Grafana, Datadog, Splunk, or similar platforms.
  • Strong problem-solving skills with a focus on reliability, automation, and operational efficiency., * Google Cloud Professional certifications.
  • Experience with microservices architectures and API-driven applications.
  • Knowledge of security best practices, compliance requirements, and cloud governance.
  • Experience working in agile, fast-paced enterprise environments., * Candidate must be legally authorized to work in the United States without the need for employer sponsorship, now or at any time in the future.
  • Applicants must be authorized to work in the United States. Sponsorship is not available for this role.
  • Candidates may be required to participate in multiple rounds of interviews, including business and client discussions.

Benefits & conditions

The annual salary for this position is between $50,000 and $70,000 , depending on experience and other qualifications of the successful candidate.

This position is also eligible for Cognizant’s discretionary annual incentive program, based on performance and subject to the terms of Cognizant’s applicable plans., Cognizant offers the following benefits for this position, subject to applicable eligibility requirements:

  • Medical/Dental/Vision/Life Insurance
  • Paid Holidays plus Paid Time Off
  • 401(k) Plan and Company Contributions
  • Long-Term Disability
  • Short-Term Disability
  • Paid Parental Leave
  • Employee Stock Purchase Plan

Disclaimer

The salary, other compensation, and benefits information is accurate as of the date of this posting. Cognizant reserves the right to modify this information at any time, subject to applicable law.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.phoenixjobs.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

Videos

See all

Related articles

See all