> Markdown version of [/jobs/ext/2709633-site-reliability-engineer-robotics-in-bodega-bay](https://www.wearedevelopers.com/jobs/ext/2709633-site-reliability-engineer-robotics-in-bodega-bay). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer, Robotics in Bodega Bay - **Company:** Energy Jobline - **Location:** Bodega Bay, CA, United States - **Salary:** $164,000.0 - $270,000.0 - **Contract:** Permanent contract - **Skills:** C++ (Programming Language), Linux, Middleware, EtherCAT, Python (Programming Language), Message Queuing Telemetry Transport (MQTT), RabbitMQ, Reliability Engineering, Prometheus, OPC Unified Architecture, TypeScript, Datadog, Diagnostic Tools, Computer Network Operations, Infrastructure as Code (IaC), Kubernetes, Bare Metal, Apache Kafka, Hardware Infrastructure, Golang - **Published:** September 4, 2026 - **Apply:** https://www.energyjobline.com/job/site-reliability-engineer-robotics-bodega-bay-31477329 ## About the Role * Ownership. Someone who has owned the reliability of a production system where downtime had physical or operational consequences (manufacturing line, autonomous vehicle, lab automation, network operations) * Systems Thinker. Focused on understanding the relationship among various systems to design sustainable solutions, not one-time fixes. * Problem Solver. Solving complex puzzles excites and motivates you to find an efficient solution. * T-Shaped Skill Set. Comfortable with bare metal Kubernetes, networking, GitOps workflows, and Infrastructure as Code (IaC). Also skilled in programming in TypeScript, Python, Golang, or C++. * Strong Communication. You can run a war room, write a post-mortem, and explain a reliability tradeoff to a stakeholder., * Background in edge/on-prem infrastructure. You've run Kubernetes at the edge (k3s, k0s, k0smotron), managing on-prem clusters, time-series at the edge, or air-gapped deployments. A deep understanding of Linux operating system fundamentals such as cgroups, sockets, and system tuning, is a big plus. * Deep understanding of shipping and storing telemetry data at scale. Experience with Kafka/MQTT/RabbitMQ is a plus. * Direct robotics experience. ROS/ROS2, OPC UA, EtherCAT, motion controllers, or fleet management for autonomous systems * An individual who is self-directed and can deliver with high velocity. ## Description What You'll Do * Own the reliability of our robotics systems, from PLCs through ROS2/middleware to Kubernetes. * Build interfaces to our observability system to ingest telemetry from our controls and robotics systems. Leverage solutions such as Prometheus, Telegraf, OpenTelemetry, and Datadog. * Write code frameworks and tools to support our controls and robotics systems, including diagnostic tools, shared libraries for telemetry data, and automated remediation. * Partner with controls, robotics, and platform engineering teams to bake reliability in early. Review designs, develop SLOs and SLIs, introduce reliability release gates, and push for telemetry contracts to develop production-grade services. ## Related Videos - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Debugging in the Dark](https://www.wearedevelopers.com/videos/1658-debugging-in-the-dark) - [Robots are coming into the wild! Full-Stack Robotics Engineers, be ready!](https://www.wearedevelopers.com/videos/479-robots-are-coming-into-the-wild-full-stack-robotics-engineers-be-ready) - [Software Engineering Social Connection: Yubo’s lean approach to scaling an 80M-user infrastructure](https://www.wearedevelopers.com/videos/1583-software-engineering-social-connection-yubo-s-lean-approach-to-scaling-an-80m-user-infrastructure) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this) - [Dev Digest 131 - AI'm not sure about OSS](https://www.wearedevelopers.com/magazine/472-dev-digest-131-ai-m-not-sure-about-oss)