Senior AI Architect, Foundation Models and SoC Co-Design - Autonomous Vehicles

NVIDIA Corporation
Santa Clara, CA, United States
1 day ago
Apply on nvidia.wd5.myworkdayjobs.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$208,000.0 - $327,750.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Computing Platforms Profiling Microprocessors Distributed Computing Environment Hardware Design Machine Learning Data Streaming Systems Architecture Graphics Processing Unit (GPU) Large Language Models Deep Learning
+5 more
AI Platforms Information Technology Low Latency Machine Learning Operations TensorRT

Job description

We are looking for a Senior AI Architect to help define the next generation of AI model paradigms for autonomous vehicles and shape how those models co-evolve with NVIDIA’s future embedded SoC architectures. This is a highly strategic role operating at the intersection of frontier AI research, hardware architecture, systems optimization, and autonomous driving. You will work with world-class AI researchers, silicon architects, and AV platform teams to identify the AI workloads that will define the next decade - and ensure NVIDIA platforms are architected to lead them.

What You’ll Be Doing:

  • Research and forecast emerging AI model architectures that are expected to shape the future autonomous vehicle stack, including Vision-Language-Action (VLA) models, Multimodal foundation models and more.
  • Drive hardware-software co-design across next-generation AI workloads and NVIDIA embedded SoCs, including GPU, CPU, DLA, memory hierarchy, interconnects, and accelerator subsystems.
  • Analyze compute, memory, bandwidth, and latency characteristics of sophisticated AI architectures such as transformers, diffusion models, or MoE systems
  • Develop architectural insights and influence future NVIDIA silicon, IP, and system-level design decisions through deep workload characterization and performance analysis.
  • Prototype and evaluate emerging model paradigms on NVIDIA DRIVE and embedded AI platforms to validate scalability, efficiency, and deployment feasibility.
  • Partner closely with AI research, autonomous driving software, compiler, runtime, and hardware architecture teams to align long-term roadmap and platform strategy.
  • Evaluate tradeoffs across latency, throughput, power efficiency, safety, and real-time constraints in production AV systems.
  • Define benchmarking methodologies and evaluation metrics for next-generation AV AI systems, including robustness, safety, calibration, and edge-case performance.

Requirements

  • MS, PhD, or equivalent experience in Computer Science, Electrical Engineering, Machine Learning, Robotics, or related field.
  • 12+ years of experience in AI/ML systems, deep learning architecture, or hardware/software co-design.
  • Deep expertise in modern AI architectures and large-scale model systems
  • Experience mapping AI workloads onto heterogeneous compute architectures including GPUs, CPUs, NPUs/DLAs, DSPs, and memory subsystems.
  • Solid understanding of distributed training systems, scaling laws, and inference optimization techniques.
  • Experience with model optimization methods such as quantization, sparsity, pruning, distillation, and memory-efficient inference.
  • Understanding of performance profiling, systems bottleneck analysis, and workload characterization., * Experience with autonomous vehicle or robotics stacks including perception, planning, prediction, or control.
  • Deep familiarity with NVIDIA platforms such as DRIVE , Jetson , CUDA®, TensorRT , Triton, or TensorRT-LLM.
  • Experience influencing silicon architecture or collaborating directly with hardware design teams.
  • Expertise in sophisticated AI efficiency techniques (e.g. FP8/FP4 inference, Mixture-of-Experts routing, Streaming attention and KV-cache optimization)
  • Strong understanding of multimodal fusion across camera, lidar, radar, HD maps, and language inputs.

We have some of the most forward-thinking and hardworking people in the world working for us. If you’re creative, autonomous, and passionate about building the future of AI and autonomous systems, we want to hear from you. Come join our team and help shape the next generation of AI computing platforms powering autonomous machines worldwide.

Benefits & conditions

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 208,000 USD - 327,750 USD.

About the company

NVIDIA is at the forefront of accelerated computing, AI, and autonomous machines. From generative AI to robotics and self-driving vehicles, our technologies are transforming some of the world’s largest industries. NVIDIA DRIVE is redefining autonomous mobility through state-of-the-art AI, high-performance compute, and scalable software-defined architectures.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on nvidia.wd5.myworkdayjobs.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · World Congress 2026 Europe

5:48 min

Balancing delivery latency with stream reliability and scale

Phil Cluff · LIVE

47 sec

Profiling native execution calls with async-profiler

Gonzalo Ortiz Jaureguizar Gonzalo Ortiz Jaureguizar · World Congress 2026 Europe

4:52 min

Essential phases in building and refining language models

Anshul Jindal Anshul Jindal +1 · World Congress 2025

6:28 min

Defining critical competencies for automotive AI engineering

Daniel Graff +1 · World Congress 2021

3:37 min

Accessing API documentation and testing remote driving latency

Alexandru Ciinaru Alexandru Ciinaru +3 · World Congress 2025

Videos

See all

Related articles

See all