Senior Accelerated Computing Architect

NVIDIA Ltd.
Santa Clara, CA, United States
27 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
6 years minimum
Compensation
$184,000.0 - $287,500.0
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Data Analysis Unix C++ (Programming Language) Nvidia CUDA Computer Programming Computer Engineering Data Centers Data Structures Microprocessors Hardware Design
+14 more
Inter-Process Communication Python (Programming Language) Machine Learning OpenCL Performance Tuning Software Architecture Scientific Computating Software Engineering Systems Architecture Graphics Processing Unit (GPU) High Performance Computing Parallel Computation Gpu Programming Information Technology

Job description

We are now looking for a Senior Accelerated Computing Architect!

NVIDIA is developing software and system architectures for accelerated high performance computing, scientific computing, machine learning, AI, datacenter, and automotive computing. This position offers you the opportunity to make a meaningful impact in a fast-moving, technology focused company.

What you’ll be doing:

  • Performing in-depth analysis and optimization to ensure the best possible performance on current and/or next-generation NVIDIA GPUs.

  • Creating and optimizing core parallel algorithms, data structures, and reference codes to provide the best possible solutions for NVIDIA GPUs.

  • Understanding and analyzing the interplay of hardware and software architectures on core algorithms, programming models, and applications.

  • Actively collaborating with the hardware design, software engineering, product, and research teams to guide the direction of accelerated computing.

  • Diving into accelerated computing applications to facilitate software-hardware co-design.

  • Writing up and presenting your work by writing white papers, conference publications, official blog posts, patent applications, etc. as appropriate.

Requirements

  • An MS or Ph.D. in Computer Science, Computer Engineering or Electrical Engineering, or equivalent experience

  • 6+ years of relevant work experience

  • Strong mathematical fundamentals, including linear algebra and numerical methods.

  • A passion for performance optimization.

  • Hands-on experience with the massively parallel GPU programming model, e.g. CUDA or OpenCL. Familiarity with APIs for multi-node communication, like MPI or OpenSHMEM/NVSHMEM, is a plus.

  • Strong knowledge of C and C++ with solid understanding of software design, programming techniques, and algorithms. Familiarity with threading APIs for multicore CPUs and Unix-style Inter-process Communication (IPC) APIs is a plus.

  • Familiarity with Python is a plus.

  • Good communication and organization skills, with a logical approach to problem solving, good time management, and task prioritization skills.

Benefits & conditions

  • NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most hard-working and dedicated people in the world working for us. Are you creative and autonomous? Do you love the challenge of pushing an architecture to its limits? If so, we want to hear from you.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.

You will also be eligible for equity and benefits (https://www.nvidia.com/en-us/benefits/) .

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.juju.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:37 min

Accelerating compute with focused developer tools

Julia Koch Julia Koch +1 · WWC Europe 2026

2:03 min

Microsoft integrating native Unix coreutils into Windows environments

Chris Heilmann +2 · LIVE

1:42 min

Navigating emerging hardware standardization in vendor programming ecosystems

Paul Graham Paul Graham · LIVE

6:21 min

Previewing upcoming hardware acceleration capabilities for Python environments

Chris Heilmann +2 · LIVE

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · WWC Europe 2026

3:30 min

Transitioning from CUDA software architect to user

Stephen Jones · Coffee With Developers

Videos

See all

Related articles

See all