> Markdown version of [/jobs/ext/2271210-senior-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/2271210-senior-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Site Reliability Engineer - **Company:** Planet Labs - **Location:** San Francisco, CA, United States (Remote available) - **Experience:** Expert - **Salary:** $153,000.0 - $191,300.0 - **Contract:** Permanent contract - **Skills:** Proxmox, JIRA, Bash Shell, Program Optimization, Nvidia CUDA, Continuous Integration, Distributed Systems, Python (Programming Language), Octopus Deploy, Reliability Engineering, Ansible, Prometheus, Zero Trust Network Access, Systems Integration, Circleci, Cloud Platform System, Grafana, Build Management, Gitlab-ci, Kubernetes, Information Technology, Bare Metal, Terraform, Jenkins - **Published:** August 27, 2026 - **Apply:** https://jobs.localjobnetwork.com/apply/add/88154963/1 ## About the Role * 6+ years of experience building services that leverage cloud-native infrastructure and tooling * Bachelor's degree in Computer Science or similar * Experience deploying and maintaining bare-metal and cloud kubernetes through tools such as Talos, RKE2, Proxmox, or k3s * Proficiency with Terraform, Ansible, Helm, Kustomize, and/or similar IaC / GitOps tooling * Experience with CI/CD tooling, such as Jenkins, GitLab CI/CD, Argo CD, or CircleCI * Experience successfully building, releasing, and supporting highly available, consistently performant services * Knowledge of hardware and network level implications of on-prem compute * Experience with platform optimization, particularly resource optimization, management, and cluster tuning in a constrained environment * Ability to observe and troubleshoot distributed systems with tools such as Alloy, Prometheus, Grafana, and OpenTelemetry * Advanced skills in Python, Bash, and other tooling as appropriate to build services and meet product goals * Excellent communication skills and the ability to work through collaboration with cross-functional engineering teams * Experience working with Jira for task management and progress tracking What Makes You Stand Out: * Experience with CUDA-based GPU programs * Security expertise in sensitive environments, including implementing zero-trust architectures, hardening Kubernetes clusters, conducting security audits, and deploying workloads in air-gapped environments ## Description Planet designs, builds, and operates the largest constellation of imaging satellites in history. This constellation delivers an unprecedented dataset of empirical information via a revolutionary cloud-based platform to authoritative figures in commercial, environmental, and humanitarian sectors. We are both a space company and data company rolled into one. In this role, you will join Planet's Direct Access Service Infrastructure team, directly contributing to our next-generation Constellation as a Service platform. This platform represents a major new offering to our customers that goes beyond traditional cloud-based platforms and supports on-premises deployments. You will be responsible for building, deploying, and operating critical compute software that supports end-to-end imaging operations within customer on-premises and/or cloud environments. You will use your understanding of internal compute requirements as well as customers' environmental-specific constraints to help design, implement, and support a robust system for reproducible deployments across operating environments, to guarantee the reliability, scalability, and availability of our services. To do this, you will partner closely with cross-functional engineering teams to enable and empower the integration of software solutions and the troubleshooting of distributed systems. This is a full-time, remote position based in the United States or Canada. If located near an office, you are expected to work from that office 3 days per week. Impact You'll Own: * Build and deploy computing services and infrastructure in customer environments for a next-generation satellite operations and image processing end-to-end platform * Operate in a high-impact, tight knit team to architect novel systems for air-gapped deployments at scale * Clarify and surface requirements from ambiguous use cases defined by cross-functional stakeholders, including internal users and external customers * Responsible for operations such as deployments, service orchestration, and documentation for cross platform stakeholders * Scale architecture while ensuring availability of services * Improve reliability and scalability by resolving edge cases, studying failure modes, and writing tests * Participate in on-call rotations to ensure operational excellence ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [The Road to MLOps: How Verivox Transitioned to AWS](https://www.wearedevelopers.com/videos/1050-the-road-to-mlops-how-verivox-transitioned-to-aws) - [Improving quality with Agentic AI with Rovo Dev and Xray](https://www.wearedevelopers.com/videos/2005-improving-quality-with-agentic-ai-with-rovo-dev-and-xray) - [All your telemetry data from any source in one place](https://www.wearedevelopers.com/videos/57-all-your-telemetry-data-from-any-source-in-one-place) - [Our GitOps approach for deploying an Identity Provider and an API Gateway in a SaaS company](https://www.wearedevelopers.com/videos/776-our-gitops-approach-for-deploying-an-identity-provider-and-an-api-gateway-in-a-saas-company) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)