Developer Infrastructure Engineer

Sunday Inc
Redwood City, CA, United States
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

C++ (Programming Language) CMake Continuous Integration Linux Programming Tools Middleware Python (Programming Language) Machine Learning Software Deployment Software Engineering System Software Management of Software Versions
+5 more
Scripting Codebase Machine Learning Operations Oracle Cloud Infrastructure Docker

Job description

At Sunday, we’re developing personal robots to reclaim the hours lost to repetitive tasks. We’re focused on an ambitious goal to make generalized robots broadly accessible, enabling households to take back quality time.

We have spent the last 18 months building a talented team, securing capital, and validating our technology. We are now seeking passionate individuals to join us in the next phase of our growth. If you are ready to apply your skills to the forefront of robotics innovation, we’d love to hear from you.

What to Expect

The ML & Robotics Infra team builds the foundational systems that every part of our robot perception, ML, controls and behavior runs on, and the developer infrastructure that lets us build, ship, and update that software quickly and safely on every robot in the fleet.

As a Developer Infrastructure Engineer, you’ll own how code and model artifacts get built and delivered to the robot. From the moment a developer hits commit to the moment the new binary, middleware, or model is running on a robot in someone’s home. You’ll build the systems that let our engineers ship quickly and confidently: a fast and reliable build, a CI/CD pipeline they trust, signed and reproducible artifacts and an OTA path that gets those artifacts safely onto the fleet. You’ll work alongside teammates who own the runtime substrate and the GPU and accelerated compute layer, they build what runs on the robot; you build how it gets there.

What You’ll Do

  • Build system: Own and evolve our build system to keep builds fast, hermetic, and reliable as the codebase and team grow
  • CI/CD: Design and operate pre-merge and post-merge pipelines that catch regressions early and produce trustworthy release artifacts
  • Artifact pipeline: Build the packaging, versioning, and signing pipeline for the binaries that ship to the robot i.e. system software, application code, and exported/compiled ML models
  • Dev containers & dependency management: Define and maintain consistent development containers and package/dependency management across the team, keeping the dev environment in parity with the base Linux image that runs on the robot
  • Over-the-air updates: Build and operate the OTA system that delivers new software to robots safely - staged rollouts, canarying, rollback, and verification
  • Model delivery: Treat compiled/exported models as first-class deployable artifacts, with versioning, validation, and the same delivery guarantees as the rest of the stack
  • Developer experience: Reduce the time between “I have an idea” and “it’s running on a robot” for local dev workflows, fast feedback loops, and clear failure signals when things break

Requirements

  • 5+ years of software engineering experience, with significant time spent on build infrastructure, developer tooling, CI/CD, or software delivery
  • Hands-on experience with a modern build system (Bazel, Buck, CMake at scale, or similar) in a large multi-language codebase
  • Strong CI/CD background - designing pipelines, managing test infrastructure, and improving signal-to-noise as systems grow
  • Working knowledge of containers (Docker/OCI) and the practical tradeoffs around image building, layering, and runtime
  • Experience deploying software to production environments - whether cloud, embedded devices, fleets, or similar - with a real understanding of versioning, rollback, and safe rollout
  • Comfort working in a systems language (C++, Rust, or Go) and a scripting language (Python) - you’ll read and write both
  • Working comfort on Linux as a target deployment platform

Nice to Have

  • Direct experience with Bazel in a polyglot codebase (C++, Python, ML model toolchains)
  • Experience building or operating OTA update systems for embedded devices, vehicles, robots, or similar fleets
  • Experience packaging and deploying ML models - model registries, versioning, compiled/exported artifact pipelines
  • Familiarity with artifact signing, supply chain security, or reproducible builds
  • Background working alongside runtime, robotics or ML teams

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

1:15 min

Defining classes and building packages with pybind11

Konstantin Bespalov · WWC 2023

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · WWC Europe 2026

2:39 min

Experiencing core Linux capabilities for DevOps administration

Michael Cade · LIVE

Videos

See all

Related articles

See all