Senior Solutions Architect, Robotics Foundation Model Training

NVIDIA Corporation
Santa Clara, CA, United States
1 day ago
Apply on nvidia.wd5.myworkdayjobs.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$152,000.0 - $241,500.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Data Analysis Profiling Computer Engineering Extract Transform Load (ETL) Data Transformation Shard (Database Architecture) Distributed Systems Performance Tuning Reinforcement Learning Pytorch Large Language Models
+8 more
Deep Learning Model Validation AI Platforms Information Technology HuggingFace Decoding Data Pipelines Data Generation

Job description

We are looking for a hands-on Applied Engineer with deep expertise in training foundation models at scale and a strong background in robotics. This role operates at the intersection of innovative AI research, accelerated computing and real-world applications, offering a unique opportunity to work directly with model builders to scale cutting edge Robotics Models from experimentation to production. Collaboration spans research, engineering, and customer teams, influencing both product direction and applied AI adoption. Come join us and help shape the future of robotics foundation model training!

What You’ll Be Doing:

  • Engage with Researchers and ML engineers to architect and optimize end-to-end training workflows for robotics foundation models, like World Models, VLAs, WAMs.
  • Build proof-of-concepts, reference architectures, and agentic workflows that accelerate experimentation, benchmarking, and model improvement of NVIDIA’s Robotics Open model platforms like Cosmos and GR00T.
  • Scale pre-training, fine-tuning, and reinforcement learning workloads across multi-GPU and multi-node systems, improving utilization, throughput, and memory efficiency.
  • Identify and eliminate data pipeline bottlenecks across storage, networking, preprocessing, and data loading for multimodal datasets (video, sensor data, trajectories).
  • Collaborate with NVIDIA product and engineering teams to provide feedback that shapes future Physical AI platforms

Requirements

  • MS, PhD, or equivalent experience in Computer Science, Artificial Intelligence, Electrical or Computer Engineering, Robotics, or a related field.
  • 5+ years of industry or research experience in deep learning, distributed computing, or large-scale model training.
  • Hands-on experience training or optimizing multimodal or foundation models (e.g., VLMs, VLAs, World Models), ideally in robotics settings.
  • Experience across the AI model lifecycle, including pre-training, supervised fine-tuning, RL or other post-training methods, evaluation, and model optimization.
  • Strong expertise in distributed training techniques (data/model/pipeline parallelism, sharding, check-pointing) on multi-GPU or multi-node systems.
  • Expertise with multimodal training frameworks such as PyTorch, NVIDIA NeMo, JAX, or Hugging Face Transformers.
  • Experience building or working with high-throughput data pipelines for large-scale training, including storage bandwidth, network throughput, and preprocessing (e.g., decoding, tokenization, batching)
  • Strong communication skills with the ability to effectively collaborate across Researchers, Engineers and executives.

Ways to Stand Out From the Crowd:

  • Familiarity with NVIDIA AI and robotics platforms (e.g., Cosmos, GR00T, NeMo, Isaac Sim, Isaac Lab)
  • Experience with robotics AI workloads, including reinforcement learning in simulation and synthetic data generation.
  • Experience profiling and optimizing workloads using tools such as Nsight Systems, Nsight Compute, or PyTorch Profiler
  • Demonstrated impact improving training efficiency and scaling performance
  • Experience building agentic workflows for automated experimentation, model evaluation, data analysis, or research acceleration

Benefits & conditions

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on nvidia.wd5.myworkdayjobs.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

2:23 min

Evaluating video encoding formats for faster decoding

Magne Johansen Magne Johansen · Europe 2026 Virtual

47 sec

Profiling native execution calls with async-profiler

Gonzalo Ortiz Jaureguizar Gonzalo Ortiz Jaureguizar · World Congress 2026 Europe

2:50 min

Transitioning from deep learning models to foundation software

Marcel Scherenberg Marcel Scherenberg · World Congress 2025

1:06 min

Compiling PyTorch environments for advanced time forecasting

Christoph Lohrmann Christoph Lohrmann +1 · World Congress 2026 Europe

3:22 min

Evaluating advanced artificial intelligence platforms for daily recruitment

Rudi Bauer Rudi Bauer +1 · Cappuccino with HR

Videos

See all

Related articles

See all