Senior Deep Learning Hardware Modeling Architect - LPU

NVIDIA Corporation
United States
2 days ago
Apply on nvidia.wd5.myworkdayjobs.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$152,000.0 - $241,500.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence C++ (Programming Language) Profiling Software Tools Large Language Models Deep Learning Parallel Computation

Job description

NVIDIA seeks a Senior DL Hardware Modeling Architect to join our group of pioneers who are pushing the performance boundaries of AI inference. Our team focuses on ambitious hardware-software co-design to optimize AI inference speed and efficiency. This role gives an outstanding opportunity to model critical components of our world-class LLM inference solutions. If you are passionate about LLM inference hardware, want to contribute at the deepest levels to the state of the art, and have a proven record of C++ hardware modeling experience, this role may be perfect for you!

What you’ll be doing:

  • Drive architectural specifications to closure across multiple stakeholders.
  • Develop written specifications for key component-level and system-level designs.
  • Embody specifications in an executable model used by many customers across NVIDIA.
  • Ensure high performance using good C++ software practices, solid algorithms and data structures, and parallelism.
  • Resolve performance and correctness issues across chip and hardware subsystems by working across teams.

Requirements

  • A BS or higher degree in a relevant field (CS, EE, Math) or equivalent experience, with 5+ years of relevant experience.
  • Expert programming and software skills in C++.
  • Experience in RTL design and/or architecture with the ability to understand common chip design concepts.
  • An automation-centered mindset.
  • A desire to improve work efficiency using AI.

Ways to stand out from the crowd:

  • Experience with systems-level/architectural/chip modeling, profiling, and analysis.
  • Strong combined RTL and C++ skillset.

Benefits & conditions

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on nvidia.wd5.myworkdayjobs.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

47 sec

Profiling native execution calls with async-profiler

Gonzalo Ortiz Jaureguizar Gonzalo Ortiz Jaureguizar · World Congress 2026 Europe

1:25 min

Distinguishing artificial intelligence from deep learning

Sam Witteveen · Coffee With Developers

2:50 min

Electronic diagnostic software tools and factory production flashing

Denis Grahovac · World Congress 2021

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

2:17 min

Comparing code profiling with surface level monitoring

Jérôme Vieilledent · LIVE

5:01 min

Leveraging large language models for code optimization and development

Stephan Gillich Stephan Gillich +3 · World Congress 2024

Videos

See all

Related articles

See all