Cloud Platform Software Engineer - Platform APIs

NVIDIA Ltd.
Seattle, WA, United States
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Compensation
$184,000.0 - $287,500.0
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Microsoft Azure Cloud Computing Cloud Engineering Computer Clusters Information Systems Computer Programming Computer Engineering Distributed Systems Python (Programming Language)
+19 more
Open Source Technology Cloud Services Software Engineering Software Systems AI Infrastructure Cloud Platform System Delivery Pipeline Multi-Cloud Core Api Containerization Integration Tests Kubernetes Infrastructure Automation Frameworks Information Technology Free and Open-Source Software Front End Software Development Terraform Code Restructuring Golang

Job description

Are you passionate about Kubernetes and AI and want to help build the best platform for ML/AI infrastructure? Do you thrive when your work directly empowers teams to push the boundaries of what’s possible? We’re the Platform API team within NVIDIA’s DGX Cloud organization - a collaborative group of cloud platform engineers, architects, and SREs who are passionate about building and nurturing the declarative, Kubernetes-native control plane that powers GPU-accelerated infrastructure across multiple cloud providers. Together, we’re empowering the world’s leading AI teams to train and deploy at datacenter scale.

We design and extend Kube-like APIs, and we craft Go-based reconciliation controllers that thoughtfully turn high-level intent into production-ready AI infrastructure. We take pride in owning our code end-to-end, and we care deeply about the full lifecycle of multi-cloud GPU clusters, from customer onboarding and provisioning through upgrades and decommissioning. We partner closely with our runtime, cloud architecture, observability, and storage teams to solve sophisticated distributed systems challenges together. As a team, we’re strengthening NVIDIA’s approach to Cloud Native development.

What you will be doing:

  • Develop software systems to support large scale deployments of cloud infrastructure
  • Design and develop APIs to support Infrastructure as Code (IaC) automation and deployment workflows.
  • Responsible for contributing to multiple source code projects to fulfill NVIDIA requirements with software services
  • Work and collaborate with engineering managers, architects, designers, and frontend engineers to deliver high quality software
  • Automate the validation of software solutions with unit and integration tests
  • Participate in the ownership and health of CI/CD pipelines from dev to production environments
  • Collaborate with other specialists for feedback on proposed designs and product direction
  • Openly share successes and failures in a no blame environment

Requirements

Do you have experience in Tooling?, Do you have a Bachelor’s degree?, * BS in Computer Science, Information Systems, Computer Engineering or equivalent experience

  • 8+ years of proven experience in large scale software development
  • Experience building and shipping services on Kubernetes
  • Background with using and chipping in to open-source projects
  • Collaborated with teams to write software to support cloud services at scale
  • Programming experience in a relevant language, e.g. Golang, Python
  • Communicate design and quality strategy in written, visual, and oral formats
  • Experience with a wide range of modern infrastructure tools and technologies

Ways to stand out from the crowd:

  • Experience with Kubernetes Cluster API, Terraform, Tinkerbell, and other infrastructure tooling
  • Practical experience with Azure, GCP, or AWS
  • Capable of refactoring software to run in systems such as Kubernetes
  • Ability to discuss and work with CSI, CNI, and CRI and/or familiarity with the CNCF and the tooling across the ecosystem
  • Upstream contribution in open source projects

Benefits & conditions

4.24.2 out of 5 stars 4545 Roosevelt Way NE, Seattle, WA 98105 Hybrid work $184,000 - $287,500 a year - Full-time, Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD.

About the company

NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High-Performance Computing and Visualization. The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our products and services. Our work opens up new universes to explore, enables amazing creativity and discovery, and powers what were once science fiction inventions from artificial intelligence to autonomous cars. NVIDIA is looking for great people like you to help us accelerate the next wave of artificial intelligence. NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people on the planet working for us. If you’re a creative, curious, and driven technical leader, we want to hear from you!

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · WWC Europe 2026

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

2:42 min

Core API types for Angular signals

Daniela Bonvini · LIVE

1:34 min

Essential commands for running and testing Terraform configurations

Hennie Francis · LIVE

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

1:33 min

Case study on adopting Kubernetes and Golang effectively

Andrew Holway · LIVE

Videos

See all

Related articles

See all