Senior Solution Engineer, Networking

NVIDIA Ltd.
Santa Clara, CA, United States
2 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Compensation
$76,230.0 - $94,380.0
Working hours
Regular working hours
Job source

Tech stack

C (Programming Language) Artificial Intelligence BIOS Software Bug Management Network Operating System (NOS) Nvidia CUDA Computer Programming Computer Networks Software Debugging Linux Embedded Software Ethernet
+16 more
Network Interface Controllers Firmware Data-Flow Analysis InfiniBand Networking Hardware Python (Programming Language) Queue Management Systems Software Engineering System Software Virtual Switching AI Infrastructure Network Routers Graphics Processing Unit (GPU) Computer Network Operations Computer Network Technologies State Machines

Job description

  • Assist various network and AI cluster support teams in reproducing, resolving, and root causing sophisticated customer issues
  • Work with R&D teams to develop bug fixes, workarounds, and solutions for critical customers using NVIDIA’s network technologies
  • Become an authority in NVIDIA network technologies used in AI clusters such as Infiniband, NVLink, and Spectrum-X
  • Analyze network performance metrics and make tuning recommendations for high-performance, lossless networks
  • Develop support and analysis tools to help analyze and root cause field issues
  • Daily use of ground breaking AI tools for software development, log and trace analysis, and source code debugging
  • Occasional work on weekends or holidays to support customers

Requirements

The NVIDIA Enterprise Experience (NVEX) Solutions Engineering team is looking for a senior Computer or Software Engineer. This person will establish expertise in ground-breaking network technology used in AI clusters. Our software engineers connect customer support teams and R&D. They focus on solving tough problems from the front lines. They provide top support for high-speed interconnect technologies like InfiniBand, NVLink, and Spectrum-X that link GPUs and AI compute infrastructure. Candidates must have a software development background in the networking industry either for a network hardware manufacturer or software integrator. It is essential to have a proven grasp of in-field, production network operations and have experience in root-causing customer-found issues down to the source code level, primarily C and Python. Breadth of experience is key. We want to see experience in multiple areas such as network operating systems (NOS), Linux network drivers and internals, network hardware, NIC software, Smart NICs, DPUs, embedded firmware, Software Defined Networking, and infrastructure management technologies. IPC, race conditions, finite state machines, event processing loops, queue management, network traffic and flow analysis, and software design gaps will be common areas of focus. The individual will get to work across many NVIDIA teams and often interact with both internal and external customers, so superb interpersonal and communication skills are essential. Candidates will need to understand, root cause, and resolve complex issues, and provide detailed explanations of what you find., * Minimum of a BS in Computer, Electrical, or Software Engineering (or equivalent experience)

  • 8+ years of experience in C programming in Linux and embedded systems
  • Proficiency in Python
  • 8+ years of experience developing software for one or more of the following:
  • Linux NIC drivers, switch ASICs and SDKs, embedded network device firmware, Linux based network equipment (routers, switches, gateways, etc), network operating systems, virtual routers, SDN stacks, virtual switching, DPDK, SRIOV stacks
  • At least 3 years of experience directly supporting end-customers, partners, or integrators for network equipment and infrastructures
  • Strong system software (firmware, BIOS, kernel, driver, operating system) expertise
  • Professional-level communication skills, including adjusting communication to the technical level of the audience, and staying calm and focused in negative situations.
  • Passion for learning innovative tech and motivation to work hard on ground-breaking products

Ways to stand out from the crowd:

  • Background with AI infrastructure and HPC networking
  • Experience programming switch and NIC ASICs and SDKs
  • Experience with Infiniband or other non-Ethernet network technologies
  • Experience developing or supporting DPUs or SmartNICs
  • Knowledge of HPC performance test tools and NVIDIA AI stacks (NCCL, MPI, DOCA, CUDA)

Benefits & conditions

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 168,000 USD - 270,250 USD for Level 4, and 200,000 USD - 322,000 USD for Level 5.

You will also be eligible for equity and benefits .

About the company

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology-and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.jofdav.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

41 sec

Massive client data loss and bio-digital storage

Chris Heilmann +1 · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

4:52 min

Connecting namespaces with local virtual ethernet pairs

Oliver Seitz Oliver Seitz · WWC 2025

3:05 min

Acquiring Mellanox to build cohesive AI factories

Michael Kagan Michael Kagan +1 · WWC Europe 2026

1:12 min

Addressing the competitive landscape of specialized hardware demands

Hazal Mestci +1 · Coffee With Developers

2:39 min

Experiencing core Linux capabilities for DevOps administration

Michael Cade · LIVE

Videos

See all

Related articles

See all