> Markdown version of [/jobs/ext/1885182-sr-site-reliability-engineer-tvscientific](https://www.wearedevelopers.com/jobs/ext/1885182-sr-site-reliability-engineer-tvscientific). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Sr. Site Reliability Engineer, tvScientific - **Company:** Pinterest - **Location:** Redondo Beach, CA, United States (Remote available) - **Experience:** Expert - **Salary:** $139,764.0 - $287,749.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Bash Shell, Cloud Computing, Continuous Integration, Data Validation, Linux, DevOps, Distributed Systems, Github, Identity and Access Management, Python (Programming Language), Role-Based Access Control, Reliability Engineering, Scripting, Delivery Pipeline, Kubernetes, Infrastructure Automation Frameworks, Information Technology, Terraform - **Published:** August 1, 2026 - **Apply:** https://diversityjobs.com/main/sendform/8/8/28176/1/17769972?backUrl=%2Fcareer%2F17769972%2FSr-Site-Reliability-Engineer-Tvscientific ## About the Role * 4+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Cloud Infrastructure * Strong hands-on experience operating AWS in production environments * Deep expertise in Kubernetes, including cluster operations, troubleshooting, workload reliability, and platform administration * Proven experience with Kubernetes multi-tenancy, including namespaces, RBAC, quotas, policies, and tenant isolation patterns * Experience implementing and operating ArgoCD within a GitOps delivery model * Strong hands-on experience with Helm * Strong experience with Terraform/Terragrunt for infrastructure provisioning and environment management * Solid scripting and automation skills using Bash and/or Python * Experience building, maintaining, or supporting CI/CD pipelines, ideally using GitHub Actions * Strong troubleshooting skills across Linux, containers, IAM, networking, and distributed systems * Experience with monitoring, alerting, and observability in production environments * Demonstrated ownership mindset with experience handling incidents, resolving production issues, and driving follow-through after outages * Strong collaboration and communication skills, with the ability to work effectively across engineering, security, and platform teams * Bachelor's degree in computer science, engineering, a related field or equivalent experience * Demonstrated ability to use AI to improve speed and quality in your day-to-day workflow for relevant outputs * Strong track record of critical evaluation and verification of AI-assisted work (e.g., testing, source-checking, data validation, peer review) * High integrity and ownership: you protect sensitive data, avoid over-reliance on AI, and remain accountable for final decisions and deliverables. ## Description We are seeking a Senior Site ReliabilityEngineer to help operate, scale, and continuously improve a cloud-native platform built on AWS, Kubernetes/EKS, and ArgoCD-driven GitOps workflows. This role will be instrumental in advancing the reliability, scalability, automation, observability, and operational maturity of our infrastructure and delivery ecosystem. The ideal candidate is a highly hands-on engineer with strong production experience and a proven ability to build and support resilient platforms using infrastructure as code, automation, and modern Kubernetes operational practices. What you'll do: * Ensuring the reliability, availability, and performance of production infrastructure and platform services * Operating and scaling Kubernetes platforms, including governance and support for multi-tenant workloads * Managing GitOps-based deployment workflows using ArgoCD and Helm * Driving infrastructure provisioning and change management through Terraform/Terragrunt * Building and supporting CI/CD automation and deployment workflows using GitHub Actions * Leading incident response efforts, root cause analysis, and post-incident improvement initiatives * Reducing operational toil through scripting, tooling, and process automation * Advancing observability practices across logs, metrics, traces, dashboards, and alerting * Supporting secure secrets integration, IAM-aware operations, and platform guardrails * Partnering closely with application, security, and platform teams to improve reliability and delivery outcomes ## Related Videos - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Designing UX for SRE Agents in High-Stakes Incidents](https://www.wearedevelopers.com/videos/100003-designing-ux-for-sre-agents-in-high-stakes-incidents) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Trustworthy AI Starts at Deployment: 5 Checks Before You Ship](https://www.wearedevelopers.com/magazine/753-trustworthy-ai-starts-at-deployment-5-checks-before-you-ship) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production)