> Markdown version of [/jobs/ext/2551928-staff-platform-engineer-developer-infrastructure](https://www.wearedevelopers.com/jobs/ext/2551928-staff-platform-engineer-developer-infrastructure). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff Platform Engineer - Developer Infrastructure - **Company:** Person AI Inc. - **Location:** Houston, TX, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** X86-64, Bash Shell, C++ (Programming Language), Cloud Computing, CMake, Nvidia CUDA, Continuous Integration, Software Debugging, Linux, Github, Virtual Private Networks (VPN), Python (Programming Language), Open Source Technology, Ansible, Prometheus, Software Engineering, Management of Software Versions, Grafana, Yocto, Gitlab-ci, Kubernetes, Terraform, Jenkins - **Published:** August 28, 2026 - **Apply:** https://www.dice.com/job-detail/c306c8be-dac2-425b-9228-218eef52a272 ## About the Role * 8+ years operating production infrastructure for a software engineering org, with direct ownership of CI/CD or developer platform work. * Deep Linux systems fluency including networking, storage, systemd, kernel and driver debugging. * Hands-on ownership of a major CI system (GitHub Actions, GitLab CI, Buildkite, Jenkins) and an understanding of the build and cache layers underneath it. * Real experience with a C/C++ build system at scale e.g., Bazel, CMake, or equivalent. Should include cross-compilation and dependency pinning. * Infrastructure as code (Ansible, Terraform) and a GitOps mindset: reviewable, reproducible, version-controlled change. * Fluent Python and Bash. You automate rather than document a manual procedure. * Strong written communication. You leave behind decisions that outlast the conversation. Bonus Skills * Shipping software to embedded or edge Linux targets - ARM64, Jetson, Yocto/custom images, A/B partitions, OTA update systems. * Hybrid on-prem plus cloud, and clear judgment about which belongs where. * Open-source observability stack (Prometheus, Grafana, OpenTelemetry) and self-hosted services (registries, artifact stores, object storage). * Robotics, autonomous vehicles, aerospace, or another hardware-heavy environment where a bad release has physical consequences. * Nix, Bazel remote execution, or other tools for reproducible builds at scale. ## Description Our robots are built by a small team that ships fast and deploys onto hardware in customer facilities. Between a laptop and a running humanoid there are: a C++/Python/Rust monorepo, an ARM64 cross-compile matrix, kernel modules and a real-time patch set, container images that have to land on Jetson devices, and a rollout process that cannot brick a robot several timezones away. Today that path is held together by the engineers who also write control code. Your job is to own it so they don't have to. Success is measured in build times, time-to-first-commit for a new engineer, deployment frequency, and the number of infrastructure problems the rest of the team stops thinking about. This is not a cloud-only role. Roughly half the surface area is physical: lab networks, robot dev boxes, bench and HIL fixtures, on-site LAN infrastructure, and the mirrors that let a deployment site work with no internet at all. What You Will Be Doing Build and CI * The build graph for a mixed C++/Python/Rust/CUDA monorepo: involving incremental correctness, remote caching, and cross-compilation for AMD64 dev boxes and ARM64 targets. * Our CI capacity: self-hosted runners, GPU and hardware-attached runners, and the queue discipline that keeps PR feedback under ten minutes. * Reproducibility: A build from a tagged commit six months from now should produce a bit-identical artifact. Release and fleet delivery * Versioning, artifact promotion, and the container registry / package mirrors that back it. * Safe, resumable, bandwidth-aware rollout to robots - staged channels, canary robots, and rollback that works over a bad link. * Signing and provenance for anything that lands on a robot. Developer platform * Reproducible dev environments across laptops, shared dev machines, and robots. * Self-service tooling so an autonomy engineer can get a branch onto a robot without filing a ticket or learning Kubernetes. * Onboarding path: a new engineer builds, tests in sim, and deploys to a bench robot on day one. Infrastructure and observability * Cloud and on-prem compute, storage for multi-TB robot logs, and the training/simulation cluster's operational layer. * Fleet observability: metrics, logs, and traces from robot to dashboard. * Site infrastructure for deployments: VPN/overlay networking, local mirrors, offline-capable registry authorization. ## Related Videos - [Code to Road in < 12 hours](https://www.wearedevelopers.com/videos/1082-code-to-road-in-12-hours) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Platform Engineering vs. DevOps Why not both?](https://www.wearedevelopers.com/videos/885-platform-engineering-vs-devops-why-not-both) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) - [Discover the open source trio you didn’t expect: .NET and PostgreSQL on Linux](https://www.wearedevelopers.com/videos/2042-discover-the-open-source-trio-you-didn-t-expect-net-and-postgresql-on-linux) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this) - [The Best X (Twitter) Accounts for Developers](https://www.wearedevelopers.com/magazine/294-the-best-x-twitter-accounts-for-developers)