Senior Solution Architect, AI Infrastructure

NVIDIA Corporation
Washington, DC, United States
10 days ago
Apply on nvidia.wd5.myworkdayjobs.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
6 years minimum
Compensation
$184,000.0 - $287,500.0
Working hours
Regular working hours

Tech stack

Big Data Cloud Computing Computer Clusters Customer Data Management Software Debugging Ethernet Network Interface Controllers InfiniBand Message Passing Interface Network Configuration and Change Management Network Architecture Network Planning and Design
+8 more
SAS (Software) AI Infrastructure Graphics Processing Unit (GPU) Computer Network Technologies Kubernetes Information Technology Slurm Hardware Infrastructure

Job description

NVIDIA is looking for a Senior AI Infrastructure Solutions Architect to join its Public Sector Continuous Bring Up team This person will be passionate about crafting an entire market and serving Public Sector customers during the AI transformation. The ideal candidate will have a strong technical background in accelerated computing technology and artificial intelligence. They will apply these skills to support government programs within the Public Sector market. They will collaborate closely with product, engineering, and customers to accelerate NVIDIA technology in the design to deployment of large-scale GPU infrastructure. A successful candidate will demonstrate skill in transcending boundaries, working effectively with product teams, account managers, field organization, and customers to drive success., * Working with NVIDIA Cloud Partners and OEMs in Public Sector on large data center GPU server and networking system deployments. Guide customer discussions on network design, compute/storage, and support bring up of server/network/cluster deployments. You will need to visit customer data center during bring up phase.

  • Become the primary technical driver for customers during the build, deployment, construction, integration, and production of GPU infrastructure and applications throughout the entire customer lifecycle.
  • Work as the customer’s trusted advisor conducting regular technical customer meetings for product roadmap, cluster issue debugging, feature discussions and introduction to new technology solutions.
  • Partner with other SAs, Account Managers, Engineering, Product, and business leaders to align on strategies, assess technical needs, and secure business opportunities for NVIDIA.
  • Analyze and debug compute/network configuration and performance issues to deliver performant clusters.
  • Prepare and deliver technical content to customers including presentations, workshops, reference architectures, tutorials, publications.
  • Lead communication with customers and NVIDIA Management.
  • Provide constructive feedback to engineering and product regarding product requirements, customer experience, documentation, and tools.

Requirements

The position requires solving complex multidisciplinary problems. Responsible for leading the resolution of technical issues across multiple engineering teams and coordinating the solutions with the customer.

  • BS/MS/PhD in Electrical Engineering, Computer Science or equivalent experience.
  • 6+ years supporting Solution Engineering (or similar Sales Engineering, Solution Architecture) including experience working directly with partners and customers.
  • Experience with high performance Networking or CPU/GPU application acceleration.
  • Background with schedulers such as SLURM, LSF, UGE, etc.
  • Experience with benchmarking tools such as HPL, NCCL tests, MLPerf as well as Kubernetes experience.
  • An ability to travel to customer sites up to 20% of the time.

Ways to stand out from the crowd:

  • Familiarity with NVIDIA GPUs, NVIDIA Networking technologies (e.g. NICs, RoCE, InfiniBand), and systems technology such as NCCL, DCGM, UFM, Mission Control, and Base Command Manager. Experience building and/or integration artificial intelligence solutions.
  • Experience with bring up and deployment of large GPU clusters, including deploying and optimizing high-speed networks (InfiniBand/Ethernet), with a clear understanding of how network architecture impacts GPU cluster performance.
  • Experience with MPI (Message Passing Interface).
  • Experience working with enterprise developers and strong customer-facing skills.
  • Active Security Clearance is highly desirable.

Benefits & conditions

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.

About the company

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you’re creative, independent, and focused on serving the mission of the U.S. Federal Government, we want to hear from you!

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on nvidia.wd5.myworkdayjobs.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:22 min

Infrastructure barriers and compliance risks in research

Jeremy Murray Jeremy Murray · World Congress 2026 Europe

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

4:52 min

Connecting namespaces with local virtual ethernet pairs

Oliver Seitz Oliver Seitz · World Congress 2025

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · World Congress 2026 Europe

1:51 min

Managing GPU quotas and multi-tenancy with Kueue

Jeremy Murray Jeremy Murray · World Congress 2026 Europe

1:12 min

Addressing the competitive landscape of specialized hardware demands

Hazal Mestci +1 · Coffee With Developers

Videos

See all

Related articles

See all