Senior Systems Software Engineer, CUDA Driver...

NVIDIA Ltd.
Santa Clara, CA, United States
3 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Compensation
$184,000.0 - $287,500.0
Working hours
Regular working hours
Job source

Tech stack

Microsoft Windows Application Programming Interfaces (APIs) Artificial Intelligence Computing Platforms C++ (Programming Language) Nvidia CUDA Computer Programming Software Debugging Linux Device Drivers Video Game Development Scientific Computating
+10 more
Software Engineering System Software Virtual Memory Multithreading Pytorch Virtual Reality Deep Learning Parallel Computation Information Technology Low Latency

Job description

As a member of our team, you will use your design abilities, coding expertise, and creativity to deliver the best compute platform in the world. You will craft elegant solutions to exciting problems and shape the future direction of CUDA as you collaborate with your peers across NVIDIA.

  • Evangelize, architect, and implement new features related to CUDA’s memory model and multi-node scalability geared towards next-gen AI applications and deployments

  • Coordinate and drive development efforts across multiple teams

  • Help define forward-looking improvements to the CUDA APIs and programming model

Requirements

Are you a motivated system software engineer with a deep understanding of device drivers, memory coherency & consistency models, phenomenal C/C++ skills, and an interest in multi-node scalability? If so, this role might be for you. We are looking for a seasoned software professional to work on the CUDA Driver, a core component of our platform for accelerating general purpose computation on the GPU. You will be an integral part of a team that delivers features and improvements to better realize the potential of NVIDIA hardware for a growing range of computational workloads, ranging from deep learning, scientific computation, data science and self-driving cars to video games and virtual reality., + BS or MS degree in Computer Science, Electrical Engineering or related field (or equivalent experience)

  • Strong C and C++ programming skills

  • Minimum of 8 years of related development experience (multiple positions for varying experience levels open)

  • Experience driving projects across multiple teams

  • Experience working with large codebases

  • Background with operating system interfaces for threads, process control, and virtual memory

  • Experience writing and debugging multithreaded programs

  • Good written communication as well as presentation skills

Ways to stand out from the crowd:

  • Prior experience with parallel computing, PyTorch, low-latency AI inference

  • Understanding of system level architecture, such as interconnects, memory hierarchy, interrupts, and memory-mapped IO

  • Knowledge of memory coherence and consistency models

  • Background with kernel mode development

  • Experience with Linux, or Windows Systems Software development

Benefits & conditions

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.

About the company

NVIDIA’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern AI - the next era of computing - with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. We’re looking to grow our company, and form teams with the smartest people in the world. Join us at the forefront of technological advancement.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on juju.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

5:48 min

Balancing delivery latency with stream reliability and scale

Phil Cluff · LIVE

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · WWC Europe 2026

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

2:27 min

Exploring the CUDA ecosystem and levels of abstraction

Paul Graham Paul Graham · WWC 2025

Videos

See all

Related articles

See all