> Markdown version of [/jobs/ext/2373259-senior-fleet-software-engineer](https://www.wearedevelopers.com/jobs/ext/2373259-senior-fleet-software-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Fleet Software Engineer - **Company:** Rhoda ai - **Location:** Mountain View, CA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Amazon Web Services, C++ (Programming Language), Data Transport Utility, Software Debugging, Linux on Embedded Systems, Fault Tolerance, Python (Programming Language), Prometheus, Software Engineering, Real Time Systems, Grafana, Kubernetes, Low Latency, Deployment Automation, Data Pipelines, Docker - **Published:** August 12, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=296ce4f3c32b29f1 ## About the Role * 3+ years of software engineering, with real ownership of internal tooling, infrastructure, or reliability systems * Strong proficiency in Python plus at least one systems language (Go, C++, or Rust) * Experience building and operating CI/CD pipelines and deployment automation * Experience with a major cloud (AWS or GCP), containers, and orchestration (Docker, Kubernetes) * Solid distributed-systems fundamentals across data transport, monitoring, and fault tolerance * Ability to debug across the full stack, from a device on the network to a service in the cloud, * Background in robotics, autonomous vehicles, or other latency- or safety-critical domains * Experience with observability stacks (e.g., Prometheus/Grafana, OpenTelemetry, Foxglove) * Experience with OTA or fleet deployment and safe-rollout patterns such as staged rollout and auto-rollback * Familiarity with ROS/ROS2, edge or embedded Linux, or low-latency data transport for real-time systems * Experience building tooling for an operations, on-call, or field team ## Description You'll build the software and infrastructure that lets a small team operate and monitor a growing fleet of deployed robots. The observability, alerting, and automation you build are what the team watches during a shift and what on-call responds to when something goes wrong. You'll contribute to the fleet's operational software end to end, from the telemetry we collect on every robot to the dashboards, pipelines, and tooling that act on it, and work closely with the operations and response teams so the fleet gets easier to run as it scales., * Build and own fleet observability: the metrics, logs, traces, and dashboards that give the team full visibility into live robots * Design telemetry and data pipelines that reliably move robot data to the cloud for monitoring, debugging, and model training * Build fleet-health dashboards and reports that make performance and regressions easy to spot, * Build alerting and on-call tooling that catches issues fast and routes them to the right responder with the right context * Automate diagnostics and incident capture so responders can start debugging instead of gathering data * Automate manual, error-prone work and reduce operational toil Deployment and release infrastructure * Contribute to CI/CD pipelines that deliver code reliably from development to the fleet * Work with software teams to automate over-the-air (OTA) software and firmware updates across the fleet, with staged rollout, monitoring, and rollback * Build provisioning and configuration tooling to keep the fleet consistent and reproducible ## Related Videos - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Data binning and understanding histograms](https://www.wearedevelopers.com/videos/2086-data-binning-and-understanding-histograms) - [Remote Driving on Plant Grounds with State-of-the-Art Cloud Technologies](https://www.wearedevelopers.com/videos/251-remote-driving-on-plant-grounds-with-state-of-the-art-cloud-technologies) - [All your telemetry data from any source in one place](https://www.wearedevelopers.com/videos/57-all-your-telemetry-data-from-any-source-in-one-place) - [Keycloak case study: Making users happy with service level indicators and observability](https://www.wearedevelopers.com/videos/1599-keycloak-case-study-making-users-happy-with-service-level-indicators-and-observability) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [How software is steering vehicle technology](https://www.wearedevelopers.com/magazine/515-how-software-is-steering-vehicle-technology) - [Résumé-Driven Development: How IT trends affect the job market for software developers](https://www.wearedevelopers.com/magazine/59-resume-driven-development-how-it-trends-affect-the-job-market-for-software-developers) - [The Best Software Developer Blogs to Read](https://www.wearedevelopers.com/magazine/156-the-best-software-developer-blogs-to-read)