> Markdown version of [/jobs/ext/2198714-systems-software-engineer-ai-and-cloud](https://www.wearedevelopers.com/jobs/ext/2198714-systems-software-engineer-ai-and-cloud). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Systems Software Engineer - AI and Cloud - **Company:** NVIDIA Ltd. - **Location:** Santa Clara, CA, United States - **Experience:** Experienced - **Salary:** $62,400.0 - $112,320.0 - **Contract:** Permanent contract - **Skills:** JavaScript (Programming Language), Application Programming Interfaces (APIs), Artificial Intelligence, C++ (Programming Language), Cloud Computing, Cloud Engineering, Computer Programming, Computer Engineering, Data Structures, Software Debugging, Python (Programming Language), Open Source Technology, Software Architecture, Software Engineering, High Performance Computing, Large Language Models, Generative AI, Solid Principles, Kubernetes, Information Technology, TensorRT, Api Design, Microservices - **Published:** August 23, 2026 - **Apply:** https://www.jofdav.com/jobs/59370693-systems-software-engineer-ai-and-cloud ## About the Role * A Bachelor's or Master's in Software Engineering, Computer Science, Computer Engineering, Electrical Engineering or a related degree (or equivalent experience) * 3+ years of experience. * Proficiency in Python and JavaScript for programming and debugging, with a strong foundation in data structures, algorithms, and software design principles. * Basic familiarity with C++ programming and its application in high-performance computing environments. * Experience in crafting cloud-native systems optimized for Kubernetes deployment, using inference frameworks such as vLLM and NVIDIA Triton Inference Server. * A solid understanding of API design principles for building scalable, production-ready inference systems. Ways to stand out from the crowd: * Advanced knowledge of LLMs, modern AI software architecture, and cloud APIs. * Contributions to public-facing technical content and open-source projects. * Expertise in deploying LLM inference frameworks like Triton Inference Server, vLLM, or TensorRT, including on Kubernetes or edge devices to improve performance. ## Description * Evaluate cloud-native, full-stack applications using microservices architecture to power AI use cases, bringing to bear NVIDIA frameworks, SDKs, and microservices. * Design and implement agentic workflows with advanced techniques like Retrieval-Augmented Generation (RAG) and the latest AI models. * Evaluate user experiences and analyze the technical performance of AI solutions, compiling findings into comprehensive reports. Offer practical suggestions for product improvement to senior executives and engineering management. * Engage with various teams across NVIDIA such as product, marketing, hardware, software engineering, and QA to improve NVIDIA's product offerings. * Develop developer-focused content, including detailed tutorials and code samples, to demonstrate the latest features in NVIDIA's tools and libraries. * Write technical whitepapers and product briefs, and run technical demos of our products at prominent industry conferences. ## Related Videos - [Tour de Force: Open-Source LLM Inference Optimization from Simple to Sophisticated](https://www.wearedevelopers.com/videos/100099-tour-de-force-open-source-llm-inference-optimization-from-simple-to-sophisticated) - [Your Next AI Needs 10,000 GPUs. Now What?](https://www.wearedevelopers.com/videos/1590-your-next-ai-needs-10-000-gpus-now-what) - [Understanding Kubernetes in a visual way](https://www.wearedevelopers.com/videos/100085-understanding-kubernetes-in-a-visual-way) - [API Design - Getting Started](https://www.wearedevelopers.com/videos/33-api-design-getting-started) - [Efficient deployment and inference of GPU-accelerated LLMs​](https://www.wearedevelopers.com/videos/929-efficient-deployment-and-inference-of-gpu-accelerated-llms) - [Trends, Challenges and Best Practices for AI at the Edge](https://www.wearedevelopers.com/videos/630-trends-challenges-and-best-practices-for-ai-at-the-edge) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers)