Senior High Performance AI Engineer, Agentic AI

NVIDIA Ltd.
Austin, TX, United States
23 days ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
4 years minimum
Compensation
$152,000.0 - $241,500.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Systems Engineering C++ (Programming Language) Compilers Nvidia CUDA Programming Tools Python (Programming Language) Machine Learning Open Source Technology Performance Tuning Reinforcement Learning Rust (Programming Language)
+8 more
Graphics Processing Unit (GPU) High Performance Computing Large Language Models Multi-Agent Systems Deep Learning Gpu Programming Information Technology Virtual Agents

Job description

You will help develop intelligent agentic systems that can reason about, generate, optimize, and operate across NVIDIA’s accelerated computing stack. This includes advancing model and agent capabilities, building scalable agent systems and runtimes, developing high-fidelity training and evaluation environments, and co-designing with NVIDIA’s libraries, runtimes, compilers, and hardware. You will collaborate closely with internal NVIDIA software, model, and hardware teams to bring new capabilities into NVIDIA products.

What you’ll be doing:

  • Design, build, and optimize agentic AI systems for the CUDA ecosystem, including agent architectures, multi-agent workflows, tool use, memory, and orchestration.
  • Build the data, environments, verification, reward, and evaluation systems needed to continuously improve model and agent capabilities.
  • Co-design and optimize agentic systems across NVIDIA’s software and hardware stack, from models and inference through compilers, runtimes, libraries, kernels, and GPUs.

Requirements

  • Bachelor’s degree in Computer Science, Electrical Engineering, or a related field, or equivalent experience; MS or PhD preferred.
  • 4+ years of relevant industry or academic experience in AI systems, machine learning, compilers, high-performance computing, or related areas.
  • Hands-on experience in one or more of the following: agent systems, coding agents, reinforcement learning, or modern AI inference systems.
  • Strong C/C++, Rust, and Python programming skills, with solid software engineering fundamentals.
  • Experience with GPU programming and performance optimization using CUDA or comparable accelerator platforms, and the ability to work effectively across system boundaries.

Ways To Stand Out From The Crowd:

  • Track record of building high-impact coding agents, autonomous software-engineering systems, or developer tools.
  • Hands-on experience optimizing and deploying with TRT-LLM, SGLang, vLLM, or Transformer Engine.
  • Deep expertise in GPU systems and performance optimization, demonstrated through benchmark results, deployed systems, publications, or widely used software.
  • Publications or open-source leadership in deep learning, agentic AI, reinforcement learning, compilers, high-performance computing, or AI systems.

Benefits & conditions

With highly competitive salaries and a comprehensive benefits package, NVIDIA is widely considered to be one of the technology industry’s most desirable employers. We have some of the most brilliant and hardworking people in the world working with us and our product lines are growing fast in some of the hottest state of the art fields such as Virtual Reality, Artificial Intelligence, Deep Learning and Autonomous Vehicles.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD.

You will also be eligible for equity and benefits.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 Ā· World Congress 2026 Europe

1:51 min

Evolution of custom compilers and virtual machines

Florian Rappl Ā· LIVE

6:21 min

Previewing upcoming hardware acceleration capabilities for Python environments

Chris Heilmann Chris Heilmann +2 Ā· LIVE

2:47 min

Contrasting classical machine learning with deep learning approaches

Adrian Spataru +1 Ā· LIVE

1:25 min

Distinguishing artificial intelligence from deep learning

Sam Witteveen Ā· Coffee With Developers

1:37 min

Accelerating compute with focused developer tools

Julia Koch Julia Koch +1 Ā· World Congress 2026 Europe

Videos

See all

Related articles

See all