> Markdown version of [/jobs/ext/48182-robot-system-qa](https://www.wearedevelopers.com/jobs/ext/48182-robot-system-qa). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Robot System QA - **Company:** Rhoda ai - **Location:** Palo Alto, CA, United States - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Automation of Tests, C++ (Programming Language), Cloud Computing, Communications Protocols, Continuous Integration, Software Debugging, Distributed Systems, EtherCAT, Failure Mode Effects Analysis, Hardware-In-The-Loop Simulation, Python (Programming Language), Machine Learning, Modbus, Networking Basics, Regression Testing, Reliability Engineering, Prometheus, System Testing, Grafana, Safety Critical Systems, Grpc, Data Pipelines - **Published:** May 15, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=bde47cb1fe775413 ## About the Role Do you have experience in System troubleshooting?, * 4+ years of experience in systems QA, reliability engineering, or a closely related field * Strong systems thinking - ability to reason about reliability and failure modes across complex, multi-component systems * Experience defining and tracking reliability KPIs, SLOs, and deployment readiness criteria for production systems * Hands-on experience designing and executing benchmarking, stress testing, and failure mode analysis programs * Strong debugging and root-cause analysis skills across software, hardware, and system boundaries * Proficiency in Python and/or C++ for test automation and tooling * Experience with CI/CD systems and automated test infrastructure * Effective communication skills - able to synthesize system health signal and communicate release readiness clearly across engineering and leadership Nice to Have (But Not Required) * Experience with hardware-in-the-loop testing or validation of physical systems * Familiarity with robotics software stacks, perception systems, or control systems * Background in reliability engineering, FMEA, or safety-critical systems validation * Experience with observability and monitoring infrastructure (e.g., Prometheus, Grafana, or similar) * Knowledge of industrial communication protocols (EtherCAT, Modbus, gRPC) and networking fundamentals * Familiarity with cloud infrastructure and distributed systems reliability ## Description * Design and own end-to-end system validation frameworks that span the full platform - robot hardware and software, cloud infrastructure, networking and communication systems, data pipelines, and operational tooling * Define and track reliability KPIs and deployment readiness criteria across all system components - establishing clear, measurable thresholds that gate production releases * Build and maintain benchmarking pipelines that systematically evaluate system performance across key dimensions: uptime, latency, throughput, fault recovery, and end-to-end task success rates * Design and execute stress testing, failure mode analysis, and fault injection programs to identify reliability risks before they surface in deployment * Investigate and root-cause system-level failures - spanning software, hardware, networking, and infrastructure boundaries - and drive corrective actions and regression tests to prevent recurrence * Collaborate closely with robotics, infrastructure, and ML teams to embed quality and testability into system design from the ground up * Build observability and reporting infrastructure that gives the team continuous, clear signal on system health and release readiness ## Related Videos - [Robots are coming into the wild! Full-Stack Robotics Engineers, be ready!](https://www.wearedevelopers.com/videos/479-robots-are-coming-into-the-wild-full-stack-robotics-engineers-be-ready) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Optimizing Land-Based Fish Feeding with Node-RED](https://www.wearedevelopers.com/videos/2032-optimizing-land-based-fish-feeding-with-node-red) - [Exploring the Power of gRPC-Gateway for Writing RESTful Services](https://www.wearedevelopers.com/videos/2072-exploring-the-power-of-grpc-gateway-for-writing-restful-services) - [All your telemetry data from any source in one place](https://www.wearedevelopers.com/videos/57-all-your-telemetry-data-from-any-source-in-one-place) - [Enabling intelligent logistics automation: home-grown Industrial IoT platform at Austrian Post](https://www.wearedevelopers.com/videos/2018-enabling-intelligent-logistics-automation-home-grown-industrial-iot-platform-at-austrian-post) ## Related Articles - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [The 8 Best Code Testing Tools](https://www.wearedevelopers.com/magazine/402-the-8-best-code-testing-tools) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Résumé-Driven Development: How IT trends affect the job market for software developers](https://www.wearedevelopers.com/magazine/59-resume-driven-development-how-it-trends-affect-the-job-market-for-software-developers) - [Dev Digest 119 - ❤️ === ❤️](https://www.wearedevelopers.com/magazine/454-dev-digest-119)