Software Engineering Manager

NVIDIA Ltd.
Santa Clara, United States
3 days ago
Apply on startup.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$224,000.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence JIRA Basic Linear Algebra Subprograms C++ (Programming Language) Compilers Software Quality Nvidia CUDA Continuous Integration Data Centers DevOps Python (Programming Language) Object-Oriented Software Development
+11 more
OpenMP Performance Tuning Software Engineering System Software AI Infrastructure Graphics Processing Unit (GPU) Performance Testing Deep Learning Parallel Computation Information Technology Data Analytics

Job description

We are looking for a Software Engineering Manager to lead a team responsible for platform expansion & readiness, enabling CUDA Math Libraries on new and specialized platforms. Around the world, leading commercial and academic organizations are revolutionizing AI, data analytics, and scientific and engineering simulations using data centers powered by GPUs. NVIDIA’s math libraries are core to the world’s AI infrastructure and must deliver functionality and performance on every target.

In this role, you will lead and build a team that provides a single accountable organization for platform integration, functional qualification, compliance, and performance readiness across CUDA Math Libraries, working closely with the core library and devops engineering teams to ensure consistent, high-quality support on every expanding platform. Ideal candidates will be hands-on engineers who also have experience leading software product engineering teams in accelerated computing domains. If this sounds exciting, we would love to meet you!

What You’ll Be Doing:

  • Lead, mentor, and develop your team.
  • Own end-to-end platform readiness across Math Libraries for specialized platforms including integration, functional qualification, and compliance.
  • Collaborate with CI/CD, build, and test infrastructure teams to establish platform-specific qualification and validation pipelines.
  • Perform defect triage and isolation of platform-specific versus library-specific problems, fixing bugs to ensure functional correctness, and referring to a library specialist as needed.
  • Identify performance targets for new platforms, establish performance testing and fix performance regressions.
  • Define and deliver to a technical roadmap for platform readiness that scales with the number and complexity of supported platforms.
  • Work closely within a team of product, engineering, and program managers for dependency coordination across library teams, CUDA, compilers, QA, release processing, and platform organizations.

Requirements

  • PhD or MSc degree in Computational Science and Engineering, Computer Science, Applied Mathematics, or related science or engineering field (or equivalent experience).
  • 8+ years of overall experience developing high-performance numerical software.
  • 3+ years leading and mentoring software engineering teams.
  • Hands-on experience with object-oriented programming, large system software architecture development, parallel computing, testing, maintenance, and performance optimization of HPC software using C++ and Python.
  • Strong understanding of fundamental numerical methods and computations in science, engineering, and/or deep learning.
  • Strong communication, collaboration, and documentation habits.
  • Experience with, and motivation to adopt and advance, software development practices such as CI/CD systems and project management tools such as JIRA.

Ways to Stand Out from the Crowd:

  • Experience with CUDA, GPU-accelerated computing, and parallel programming (e.g. MPI, OpenMP, OpenACC, pthreads).
  • Familiarity with math libraries (BLAS, LAPACK, FFT, sparse solvers).
  • Proven track record using Agentic AI to boost your efficiency and code quality.
  • Experience with cross-platform software development and platform bring-up across multiple architectures.
  • Experience delivering software for safety-critical or embedded environments (e.g., DriveOS, ISO 26262).

Benefits & conditions

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 224,000 USD - 356,500 USD for Level 3, and 272,000 USD - 431,250 USD for Level 4.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on startup.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:37 min

Simplifying parallel programming with the CUDA ecosystem

Paul Graham Paul Graham · LIVE

1:22 min

Analyzing differences between mobile and traditional backend DevOps

Mete Baydar Mete Baydar · World Congress 2025

3:05 min

Integrating an assistant application with Jira software

Felix Augenstein · LIVE

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · World Congress 2026 Europe

3:27 min

Defining DevOps through its historical origins and foundational texts

Sonal Patil · LIVE

2:11 min

Updating the delivery architecture with Jira and Tekton pipelines

Lian Li · World Congress 2022

Videos

See all

Related articles

See all