Senior GPU Architect, Deep Learning

NVIDIA Ltd.
Santa Clara, CA, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Compensation
$184,000.0 - $287,500.0
Working hours
Regular working hours
Job source

Tech stack

C (Programming Language) C++ (Programming Language) Computer Programming Databases Computer Engineering Perl (Programming Language) Python (Programming Language) Signal Processing Software Engineering Graphics Processing Unit (GPU) High Performance Computing Deep Learning
+2 more
Parallel Computation Information Technology

Job description

The NVIDIA GPU Architecture group is looking for world class architects and software developers to join and lead our various architecture efforts. A key part of NVIDIA’s strength is to innovate in the graphics and parallel computing fields delivering the highest performance in the world for deep learning and parallel processing algorithms. We are constantly looking for ways to improve our GPU architecture, especially for deep learning workloads, both training and inference, and maintain our leadership by developing new parallel programming models, and new architectures required to make this successful. In this position, you will be responsible for developing and enhancing various features in the GPU architecture that advance the state of the art in parallel programming models or parallel computing performance. You would interact with other world-class architects and researchers to build simulators, mapping deep learning workloads to current and future hardware, and validate new architectural features.

What you’ll be doing:

  • Design new hardware features for future processing architectures targeted at deep learning workloads, for both training and inference.

  • Advance the state of parallel computation.

  • Be knowledgeable about future parallel programming models and their impact to hardware.

  • Develop software for various hardware simulators, test infrastructures or metrics systems including databases.

  • Work in a team to document, design, develop tools to analyze and simulate, validate, and verify functional or performance models.

  • Develop tests, testplans, and testing infrastructure for new graphics or parallel processing architectures

Requirements

  • MS in Computer Science, Electrical Engineering or Computer Engineering or equivalent experience.

  • Experience in working with hardware targeted at deep learning, or working on mapping deep learning algorithms to hardware.

  • 8+ years of relevant industry experience in GPU or other parallel programming architectures (or other equivalent experience).

  • Strong programming ability in C, C++, Perl and Python.

  • Background in computer architecture, parallel processing, signal processing and/or high performance computing.

  • Knowledge of state of the art in DL algorithms and attention mechanisms is a huge plus.

Benefits & conditions

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.

About the company

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hard working people in the world working for us. If you’re creative, autonomous, and love a challenge, consider joining our GPU Architecture team and help us build the real-time, cost-effective AI computing platform driving our success in this exciting and quickly growing field.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.juju.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:04 min

Database evolution and the funding behind vector databases

Erik Bamberg · LIVE

1:25 min

Distinguishing artificial intelligence from deep learning

Sam Witteveen · Coffee With Developers

1:39 min

Core skills required for robust robotic systems

Falk-Moritz Schaefer · WWC 2022

2:08 min

History and scale of NVIDIA GPU computing

Paul Graham Paul Graham · LIVE

4:01 min

Managing application isolation via pluggable database models

Wei Hu Wei Hu · WWC 2022

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · WWC Europe 2026

Videos

See all

Related articles

See all