Systems Software Engineer - AI and Cloud

NVIDIA Ltd.
Santa Clara, CA, United States
5 days ago
Apply on www.jofdav.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$62,400.0 - $112,320.0
Working hours
Regular working hours
Job source

Tech stack

JavaScript (Programming Language) Application Programming Interfaces (APIs) Artificial Intelligence C++ (Programming Language) Cloud Computing Cloud Engineering Computer Programming Computer Engineering Data Structures Software Debugging Python (Programming Language) Open Source Technology
+11 more
Software Architecture Software Engineering High Performance Computing Large Language Models Generative AI Solid Principles Kubernetes Information Technology TensorRT Api Design Microservices

Job description

  • Evaluate cloud-native, full-stack applications using microservices architecture to power AI use cases, bringing to bear NVIDIA frameworks, SDKs, and microservices.
  • Design and implement agentic workflows with advanced techniques like Retrieval-Augmented Generation (RAG) and the latest AI models.
  • Evaluate user experiences and analyze the technical performance of AI solutions, compiling findings into comprehensive reports. Offer practical suggestions for product improvement to senior executives and engineering management.
  • Engage with various teams across NVIDIA such as product, marketing, hardware, software engineering, and QA to improve NVIDIA’s product offerings.
  • Develop developer-focused content, including detailed tutorials and code samples, to demonstrate the latest features in NVIDIA’s tools and libraries.
  • Write technical whitepapers and product briefs, and run technical demos of our products at prominent industry conferences.

Requirements

  • A Bachelor’s or Master’s in Software Engineering, Computer Science, Computer Engineering, Electrical Engineering or a related degree (or equivalent experience)
  • 3+ years of experience.
  • Proficiency in Python and JavaScript for programming and debugging, with a strong foundation in data structures, algorithms, and software design principles.
  • Basic familiarity with C++ programming and its application in high-performance computing environments.
  • Experience in crafting cloud-native systems optimized for Kubernetes deployment, using inference frameworks such as vLLM and NVIDIA Triton Inference Server.
  • A solid understanding of API design principles for building scalable, production-ready inference systems.

Ways to stand out from the crowd:

  • Advanced knowledge of LLMs, modern AI software architecture, and cloud APIs.
  • Contributions to public-facing technical content and open-source projects.
  • Expertise in deploying LLM inference frameworks like Triton Inference Server, vLLM, or TensorRT, including on Kubernetes or edge devices to improve performance.

Benefits & conditions

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 124,000 USD - 195,500 USD for Level 2, and 152,000 USD - 241,500 USD for Level 3.

About the company

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology-and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.

Join NVIDIA, where we are pushing the boundaries of what’s possible in AI and cloud computing. As a versatile System Software Engineer - AI and Cloud, you will be part of a team of dedicated professionals that thrives on innovation and collaboration. Located in the heart of Silicon Valley, you will have the opportunity to work on groundbreaking projects that craft the future of technology. This role offers an outstanding chance to engage with advanced AI models and cloud-native architectures, making significant contributions to NVIDIA’s versatile products and technologies.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.jofdav.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · World Congress 2026 Europe

2:28 min

Understanding Kubernetes architecture and core cluster components

Marc Nimmerrichter · World Congress 2022

3:10 min

Understanding the core concepts of API design

Alen Pokos · LIVE

4:52 min

Essential phases in building and refining language models

Anshul Jindal Anshul Jindal +1 · World Congress 2025

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

4:04 min

Overview of Kubernetes operators and custom resource definitions

Philipp Krenn · World Congress 2022

Videos

See all

Related articles

See all