Manager, Systems Software Engineering - NV Cloud Functions

NVIDIA Ltd.
Santa Clara, CA, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$224,000.0 - $356,500.0
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) C++ (Programming Language) Cloud Computing Computer Programming Distributed Systems Reliability Engineering Software Engineering AI Infrastructure Rust (Programming Language) Graphics Processing Unit (GPU) Containerization Kubernetes
+6 more
Information Technology Codebase Hardware Infrastructure Api Design Golang Microservices

Job description

  • Lead a world-class team of software engineers responsible for the development, reliability, and optimization of the NVIDIA Cloud Functions software stack.
  • Successfully implement and coordinate project planning, scheduling, and execution to ensure flawless delivery of platform features and improvements.
  • Collaborate closely with cross-functional teams, including product, security, site reliability engineering, GPU infrastructure, and architecture, to drive innovation and achieve project milestones.
  • Determine and prioritize project scope, ensuring high-quality technical solutions and seamless integration across the NVCF platform and its supported deployment environments.
  • Foster an inclusive and collaborative environment where team members can thrive and produce their best work.

Requirements

We are looking for an experienced and creative software engineering manager to lead an outstanding team working on NVIDIA Cloud Functions (NVCF). NVCF is a platform for deploying, managing, and running GPU- accelerated workloads at scale. This role is an opportunity to be at the forefront of cloud and AI infrastructure, shaping how developers and enterprises build and operate accelerated applications. If you have a passion for distributed systems, cloud infrastructure, and accelerated computing, and a proven record of managing ambitious projects, we want to hear from you!, * BS, MS, or equivalent experience in Computer Science, Electrical Engineering, or a related field.

  • Previous 3+ years of experience in managing software engineering teams.
  • 7+ overall years of industry experience (or equivalent) in distributed systems software engineering.
  • Strong leadership and interpersonal skills, with the ability to encourage and guide a diverse team.
  • Established background in decentralized architectures, cloud-native technologies, and the development of reliable, scalable software platforms.
  • Excellent verbal and written communication skills, capable of achieving objectives under tight deadlines.

Ways to stand out from the crowd:

  • In-depth knowledge of cloud computing, distributed systems, and GPU-accelerated infrastructure.
  • Extensive experience with Kubernetes, containers, workload orchestration, and public or private cloud architectures.
  • Proven programming skills in Golang, Java, Rust, C++, or similar languages, with experience working in large codebases.
  • Familiarity with API development, microservices, networking, observability, and site reliability engineering.
  • Experience building secure, highly available, multi-tenant platforms or operating production services at global scale.

Benefits & conditions

4.24.2 out of 5 stars 2788 San Tomas Expressway, Santa Clara, CA 95051 $224,000 - $356,500 a year - Full-time, Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 224,000 USD - 356,500 USD.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

3:34 min

Augmenting codebases with semantic architectures

Zaak Chalal Zaak Chalal · WWC Europe 2026

3:10 min

Understanding the core concepts of API design

Alen Pokos · LIVE

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · WWC Europe 2026

1:33 min

Case study on adopting Kubernetes and Golang effectively

Andrew Holway · LIVE

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · WWC Europe 2026

Videos

See all

Related articles

See all