Staff Reinforcement Learning Research Engineer

Boston Dynamics
Waltham, MA, United States
3 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$155,284.0 - $200,000.0
Working hours
Regular working hours
Job source

Tech stack

Continuous Integration Data Visualization Machine Learning Reinforcement Learning Pytorch Kubernetes ONNX (Open Neural Network Exchange) Format Data Analytics TensorRT Docker

Job description

Do you want to build the scalable reinforcement learning framework that powers the next generation of humanoid and quadruped robots? As a Staff RL Research Engineer, you’ll own the RL stack, including massively parallel simulation, domain randomization, policy optimization, and on-robot deployment. Your job is to make the pipeline fast, reliable, and reproducible. You’ll work alongside world-class engineers and scientists pushing the boundaries of whole-body control and dexterous manipulation.

In this role, you will:

  • Implement on-policy and off-policy learning algorithms
  • Scale GPU-accelerated simulation to generate millions of samples per second
  • Crack sim-to-real to produce policies that transfer to the physical robot
  • Integrate RL with VLAs to fine-tune and distill large multimodal policies
  • Make deployment easy, fast, and reproducible
  • Build visualization tools that enable data-driven research

Requirements

Do you have experience in Simulation systems?, * MS with 3+ years of experience, or PhD, in ML, Robotics, or a related field

  • Deployed policies on physical robots with attention to latency, robustness, and safety
  • Expertise with RL toolboxes (RSL-RL, CleanRL, RLlib, Stable Baselines)
  • Expertise with simulation and rendering tooling (Isaac Lab, MuJoCo, MjWarp, MjLab)
  • Proficient in PyTorch and/or JAX, plus inference runtimes (ONNX, Triton, TensorRT)
  • Solid software fundamentals: Bazel, monorepos, Docker, CI/CD

The ideal candidate has:

  • Built production-grade RL training pipelines
  • Deep knowledge of GPU-accelerated physics simulation
  • Applied RL to humanoid locomotion, whole-body control, or dexterous manipulation
  • Worked on sim-to-real transfer, domain randomization, or system identification
  • Experience with heterogeneous compute clusters and Kubernetes

Benefits & conditions

Pulled from the full job description

  • 401(k)
  • Health insurance
  • Paid time off
  • Vision insurance
  • Dental insurance, * Ownership of the company wide RL tools powering all of our robots
  • Direct access to the compute infrastructure to run large-scale experiments
  • The chance to help define what’s possible in real-world robotics

The salary or hourly pay range for this position will be clearly stated in the job posting as required by Massachusetts law. The base pay range for this position is between $155,284.34- $200,000. Base pay will depend on multiple individualized factors including, but not limited to internal equity, job related knowledge, skills and experience. This range represents a good faith estimate of compensation at the time of posting. Boston Dynamics offers a generous Benefits package including medical, dental vision, 401(k), paid time off and an annual bonus structure. Additional details regarding these benefit plans will be provided if an employee receives an offer for employment.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:57 min

Advancing simulated reinforcement learning for diverse physical environments

Clemens Wasner Clemens Wasner +3 · WWC Europe 2026

4:52 min

Essential phases in building and refining language models

Anshul Jindal Anshul Jindal +1 · WWC 2025

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

2:03 min

Performing targeted reinforcement learning using specialized simulated laboratory gyms

Sergio Perez Sergio Perez · WWC Europe 2026

2:32 min

Core libraries driving inference engines and multi-GPU networking

Adolf Hohl Adolf Hohl · WWC 2024

Videos

See all

Related articles

See all