Senior Deep Learning Systems Architect

NVIDIA Ltd.
Santa Clara, CA, United States
27 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Compensation
$224,000.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Artificial Neural Networks C++ (Programming Language) Nvidia CUDA Computer Programming General-Purpose Computing on Graphics Processing Units Python (Programming Language) Machine Learning OpenMP OpenCL Tensorflow Systems Architecture
+4 more
Graphics Processing Unit (GPU) Pytorch Deep Learning Information Technology

Job description

  • As a member of our deep learning architecture team, you will contribute to features that help next-generation GPUs and systems advancing the state of AI.

  • This position requires you to keep up with the latest DL research and collaborate with diverse teams (internal and external to NVIDIA), including DL researchers, hardware architects, and software engineers.

  • As a system architect for NVIDIA’s offerings for AI systems, you will participate in engineering projects and co-design architecture for systems from conception, specification and prototyping.

  • Understanding various AI/DL workloads and their mapping to underlying HW and Systems. Identifying potential improvements and bottlenecks, proposing solutions to address existing gaps, and accelerate/improve current systems/methods.

  • Comprehensive analyses from first principles of various deep learning techniques, system optimizations to build out analytical models as well as implementing prototypes, and benchmarking to test/prove ideas.

Requirements

  • MS (or equivalent experience) or PhD degree in computer science, computer architecture, electrical engineering or related field with 10+ years of relevant work experience. Additional equivalent experience in several of the relevant areas listed below can substitute for an advanced degree.

  • Strong background in at least a few of the following relevant areas is required in your work history: Machine learning (with focus on Deep Neural Networks), including a solid understanding of DL fundamentals; Experience adapting and training DNNs for various tasks; Experience developing code for one or more of the DNN training frameworks (such as PyTorch, TensorFlow or JAX): Numerical analysis, Performance analysis and optimization & Computer architecture.

  • Programming fluency with C++ and ideally Python.

  • Work experience with GPU computing (CUDA, OpenCL, OpenACC) and HPC (MPI, OpenMP) is a huge plus.

Benefits & conditions

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 224,000 USD - 356,500 USD.

About the company

Intelligent machines powered by AI computers that can learn, reason and interact with people are no longer science fiction. Today, a self-driving car powered by AI can meander through a country road at night and find its way. An AI-powered robot can learn motor skills through trial and error. This is truly an extraordinary time. The era of AI has begun. NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you’re creative, autonomous and love a challenge, we want to hear from you! Come, join our Deep Learning Architecture team and help build the real-time, cost-effective computing platform driving our success in this exciting and quickly growing field.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.juju.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:37 min

Simplifying parallel programming with the CUDA ecosystem

Paul Graham Paul Graham · LIVE

1:42 min

Navigating emerging hardware standardization in vendor programming ecosystems

Paul Graham Paul Graham · LIVE

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · WWC Europe 2026

2:08 min

History and scale of NVIDIA GPU computing

Paul Graham Paul Graham · LIVE

1:06 min

Compiling PyTorch environments for advanced time forecasting

Christoph Lohrmann Christoph Lohrmann +1 · WWC Europe 2026

Videos

See all

Related articles

See all