> Markdown version of [/jobs/ext/2875307-lead-software-engineer-ml-network-stack-annapurna-labs-in-cupertino](https://www.wearedevelopers.com/jobs/ext/2875307-lead-software-engineer-ml-network-stack-annapurna-labs-in-cupertino). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Lead Software Engineer, ML Network Stack - Annapurna Labs in Cupertino - **Company:** Energy Jobline - **Location:** Cupertino, CA, United States - **Experience:** Expert - **Salary:** $193,300.0 - $261,500.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Elastic Compute Cloud, C++ (Programming Language), Code Review, Protocol Stack, Nvidia CUDA, Software Design Patterns, Linux, Remote Direct Memory Access, Software Engineering, Information Technology, Build Process, Machine Learning Operations, Software Coding, Software Version Control - **Published:** September 13, 2026 - **Apply:** https://www.energyjobline.com/job/lead-software-engineer-ml-network-stack-annapurna-labs-cupertino-31640579 ## About the Role nBASIC QUALIFICATIONS- 5+ years of leading design or architecture (design patterns, reliability and scaling) of new and existing systems experience\n \n- 5+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience\n \n- Experience as a mentor, tech lead or leading an engineering team, or experience in development in the last 3 years\n \n- 5+ years of experience with programming : C or C++\n \nPREFERRED QUALIFICATIONS- Bachelor's degree in computer science or equivalent\n \n- Experience working with ML Communication Libraries or RDMA Networking\n ## Description Job DescriptionWe are seeking an experienced engineer and technical leader to join our team that owns the network stack for EC2 distributed AI/ML systems. The team develops support for a variety of frameworks and communication libraries including NCCL, NVSHMEM, NIXL, NCCL GIN, and CUDA kernels. Solid knowledge of Linux, networking, and performant coding is important. Experience with embedded systems is valued, and experience with high-speed networking or HPC/RDMA interconnects is highly valued. ## Related Videos - [Coffee with Developers - Stephen Jones - NVIDIA](https://www.wearedevelopers.com/videos/1303-coffee-with-developers-stephen-jones-nvidia) - [LLMOps-driven fine-tuning, evaluation, and inference with NVIDIA NIM & NeMo Microservices](https://www.wearedevelopers.com/videos/1582-llmops-driven-fine-tuning-evaluation-and-inference-with-nvidia-nim-nemo-microservices) - [Are Code Reviews Worth It? Insights from 16 Years of Review Data](https://www.wearedevelopers.com/videos/1135-are-code-reviews-worth-it-insights-from-16-years-of-review-data) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Developer Experience, Platform Engineering and AI powered Apps](https://www.wearedevelopers.com/videos/990-developer-experience-platform-engineering-and-ai-powered-apps) - [The weekly developer show: Boosting Python with CUDA, CSS Updates & Navigating New Tech Stacks](https://www.wearedevelopers.com/videos/1293-the-weekly-developer-show-boosting-python-with-cuda-css-updates-navigating-new-tech-stacks) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)