> Markdown version of [/jobs/ext/2562129-senior-site-reliability-engineer-sre-ecommerce-google-cloud-platform-gcp](https://www.wearedevelopers.com/jobs/ext/2562129-senior-site-reliability-engineer-sre-ecommerce-google-cloud-platform-gcp). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Site Reliability Engineer (SRE) - eCommerce & Google Cloud Platform (GCP) - **Company:** Cognizant Technology Solutions Corporation - **Location:** Phoenix, AZ, United States - **Experience:** Expert - **Salary:** $50,000.0 - **Contract:** Permanent contract - **Skills:** Amazon Elastic Compute Cloud, Application Performance Management, Bash Shell, Cloud Computing, Cloud Engineering, Cloud Storage, DevOps, Distributed Systems, Github, Python (Programming Language), Linux System Administration, Systems Development Life Cycle, Reliability Engineering, Prometheus, Web Platforms, Datadog, Google Cloud, Cloud Monitoring, Grafana, Reliability of Systems, Infrastructure as Code (IaC), Git, Kubernetes, Api Design, Terraform, Splunk, Docker, Jenkins, Microservices - **Published:** August 21, 2026 - **Apply:** https://www.phoenixjobs.com/job.asp?id=3361170483&tx=KL6767FFJ&pt=1&aff=0B19D771-A501-4A5E-8338-2A822B784D54&utm_source=Job%20Feed&utm_medium=textkernel&utm_campaign=DE&utm_term=0B19D771-A501-4A5E-8338-2A822B784D54 ## About the Role * 8+ years of experience in Site Reliability Engineering, Cloud Engineering, or DevOps roles. * Strong hands-on experience with Google Cloud Platform (GCP), including services such as GKE, Compute Engine, Cloud Storage, Cloud Monitoring, and Cloud Operations Suite. * Experience supporting high-volume eCommerce or customer-facing digital platforms. * Proficiency in Kubernetes, Docker, and container orchestration technologies. * Strong scripting and automation skills using Python, Bash, or similar languages. * Experience with Terraform, Git, Jenkins, GitHub Actions, or comparable CI/CD tools. * Deep understanding of Linux systems administration, networking, and distributed systems. * Experience with observability tools including Prometheus, Grafana, Datadog, Splunk, or similar platforms. * Strong problem-solving skills with a focus on reliability, automation, and operational efficiency., * Google Cloud Professional certifications. * Experience with microservices architectures and API-driven applications. * Knowledge of security best practices, compliance requirements, and cloud governance. * Experience working in agile, fast-paced enterprise environments., * Candidate must be legally authorized to work in the United States without the need for employer sponsorship, now or at any time in the future. * Applicants must be authorized to work in the United States. Sponsorship is not available for this role. * Candidates may be required to participate in multiple rounds of interviews, including business and client discussions. ## Description We are seeking an experienced Senior Site Reliability Engineer (SRE) to support and enhance large-scale eCommerce platforms hosted on Google Cloud Platform (GCP). The ideal candidate will be responsible for ensuring the reliability, scalability, performance, and security of mission-critical production environments while driving automation and operational excellence., * Design, implement, and maintain highly available, resilient, and scalable cloud infrastructure on GCP. * Monitor application and platform health, proactively identify issues, and lead incident response and root cause analysis. * Develop and maintain Infrastructure as Code (IaC) solutions using tools such as Terraform. * Automate operational processes, deployments, monitoring, and remediation workflows. * Collaborate with development, architecture, security, and operations teams to improve system reliability and performance. * Manage CI/CD pipelines and support DevOps best practices across the software delivery lifecycle. * Optimize application performance, capacity planning, and cost management within GCP environments. * Establish and maintain SLIs, SLOs, and error budgets to improve service reliability. * Support on-call rotations and lead troubleshooting efforts for critical production incidents. ## Related Videos - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [The Best Job Search Websites of 2025](https://www.wearedevelopers.com/magazine/368-the-best-job-search-websites-of-2025)