> Markdown version of [/jobs/ext/622256-site-reliability-engineer-lead](https://www.wearedevelopers.com/jobs/ext/622256-site-reliability-engineer-lead). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer Lead - **Company:** ESG - **Location:** Houston, TX, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Airflow, Amazon Elastic Compute Cloud, BigQuery, Cloud Storage, Continuous Integration, Information Engineering, Data Systems, DevOps, Data Flow Control, Github, Identity and Access Management, Python (Programming Language), Operational Data Store, Reliability Engineering, Site Reliability Engineering Practices, Cloud Services, Data Streaming, Data Logging, Google Cloud, Cloud Platform System, Cloud Monitoring, Gitlab-ci, Kubernetes, Information Technology, Data Management, Terraform - **Published:** June 24, 2026 - **Apply:** https://www.dice.com/job-detail/9f794ed2-2436-47b2-8e9a-a2984f384dab ## About the Role We are seeking an Site Reliability Engineer Lead to own and evolve the reliability, scalability, and operational excellence of cloud-native data platforms running primarily on Google Cloud Platform (Google Cloud Platform). This role supports data systems that ingest, process, and serve large volumes of operational data from oilfield and energy environments. The ideal candidate is a cloud-first SRE with deep Google Cloud Platform experience, strong Python engineering skills, and a track record of leading reliability initiatives for data-intensive systems., * 7+ years in SRE, Cloud Platform Engineering, or DevOps * Strong hands-on experience with Google Cloud Platform, including: * Google Cloud Platform: GKE, Compute Engine, Cloud Storage, Pub/Sub (or equivalents) * Cloud Monitoring & Logging * BigQuery * Dataflow * Datastream * IAM and networking * Composer/AIrflow * Kubernetes: deployment, scaling, reliability patterns * CI/CD: GitHub Actions, GitLab CI, or similar * Observability: Google Cloud Platform Cloud Monitoring, Logging * Experience supporting cloud-native data systems (batch and streaming) * Production experience with Python for automation, tooling, or services * Infrastructure as Code experience (Terraform strongly preferred) * Experience operating systems in 24/7 production environments, * Bachelor's degree in Business, Information Technology, Computer Science, or a related field. * 5+ years experience in Site Reliability Engineering, Cloud Platform Engineering, or DevOps * 3+ years operating production workloads on Google Cloud Platform (Google Cloud Platform) * Prior technical leadership experience (lead engineer, tech lead, or ownership of reliability initiatives) * Ability to understand and speak English at a level of proficiency allowing employee to issue, receive and respond to both safety and operations-related directions in English Preferred Qualifications: * Oil and Gas Industry knowledge * Technology/Digital Industry knowledge ## Description * Lead SRE practices for Google Cloud Platform-based data platforms * Design and own SLIs, SLOs, error budgets, and reliability metrics * Build and maintain cloud-native observability (monitoring, logging, alerting) * Lead incident response for production cloud systems and drive postmortems * Partner with data engineering and platform teams to design reliable architectures * Automate operational workflows using Python * Drive improvements in CI/CD, infrastructure as code, and deployment safety * Mentor engineers and set SRE best practices across the team ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Data Science, ML & AI in the Oil and Gas Industry at NDT Global - Dr. Katja Träumner](https://www.wearedevelopers.com/videos/1308-data-science-ml-ai-in-the-oil-and-gas-industry-at-ndt-global-dr-katja-traumner) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Bringing AI Model Testing and Prompt Management to Your Codebase with GitHub Models](https://www.wearedevelopers.com/videos/1536-bringing-ai-model-testing-and-prompt-management-to-your-codebase-with-github-models) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london) - [A Guide to Green Tech and Green IT Careers](https://www.wearedevelopers.com/magazine/374-a-guide-to-green-tech-and-green-it-careers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs)