> Markdown version of [/jobs/ext/1387032-software-development-engineer-i-ai-ml-network-infrastructure-annapurna-labs](https://www.wearedevelopers.com/jobs/ext/1387032-software-development-engineer-i-ai-ml-network-infrastructure-annapurna-labs). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Software Development Engineer I - AI/ML Network Infrastructure, Annapurna Labs - **Company:** Amazon.com, Inc. - **Location:** Cupertino, CA, United States - **Experience:** Starter - **Salary:** $127,100.0 - $185,000.0 - **Contract:** Internship / Graduate position - **Skills:** Artificial Intelligence, Amazon Web Services, Amazon Elastic Compute Cloud, C++ (Programming Language), Computer Clusters, Profiling, Protocol Stack, Nvidia CUDA, Computer Engineering, Continuous Integration, Distributed Systems, Memory Management, Fault Tolerance, Python (Programming Language), Linux Kernel, Linux System Administration, Message Passing Interface, Network Architecture, Network Programming, Network Service, Open Source Technology, Remote Direct Memory Access, Software Engineering, Toolchain, Multithreading, Grafana, Gpu Programming, Linux Development, Information Technology, Machine Learning Operations - **Published:** July 22, 2026 - **Apply:** https://www.amazon.jobs/en/jobs/10475660/software-development-engineer-i-ai-ml-network-infrastructure-annapurna-labs ## About the Role Bachelor's or Master's degree in Computer Science, Computer Engineering, or related field (recent graduates welcome) - Strong proficiency in C/C++ - Solid coursework or project experience in: 1/ Operating Systems (Linux internals, kernel concepts, memory management) 2/ Parallel Computer Architecture (multi-threading, SIMD, GPU programming, cache coherence) 3/ Distributed Systems (consensus, message passing, fault tolerance, scalability) - Familiarity with Linux development environments and toolchains Preferred Qualifications - Internship experience in ML communications, HPC networking, or RDMA/high-speed interconnects - Exposure to network programming (sockets, MPI, collective communication patterns) - Experience with performance profiling and optimization - Familiarity with GPU programming (CUDA) or hardware-software co-design - Contributions to open-source projects in systems, networking, or HPC ## Description We're looking for a talented early-career engineer to join our team that owns the network stack for EC2 distributed AI/ML systems. You'll work on software that enables the world's largest AI models to train across massive GPU clusters, developing support for communication libraries and frameworks like NCCL, NVSHMEM, and NIXL., Write high-performance C/C++ code for network communication libraries running on custom AWS hardware - Build and maintain infrastructure that monitors functionality and performance of large-scale AI/ML workloads - Develop automation using Python and AWS tools (CI/CD, Grafana, Athena) to test, benchmark, and deliver software to customers - Design mechanisms to detect functional and performance regressions before they reach production - Work across many instance types, software stacks, and Linux environments ## Related Videos - [The Gashlycrumb Tinies of AI Networking You Must Know (or Languish!)](https://www.wearedevelopers.com/videos/2067-the-gashlycrumb-tinies-of-ai-networking-you-must-know-or-languish) - [Profiling Symfony & PHP apps with Blackfire](https://www.wearedevelopers.com/videos/265-profiling-symfony-php-apps-with-blackfire) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Speeding up Web Apps performance with WebAssembly and Emscripten](https://www.wearedevelopers.com/videos/1985-speeding-up-web-apps-performance-with-webassembly-and-emscripten) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [The Power of Developer Communities](https://www.wearedevelopers.com/videos/1109-the-power-of-developer-communities) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Dev Digest 162: AI careers, MCP, AWS best practices & floppy sweaters](https://www.wearedevelopers.com/magazine/571-dev-digest-162-ai-careers-mcp-aws-best-practices-floppy-sweaters) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Why Attend a Developer Event in 2026?](https://www.wearedevelopers.com/magazine/688-why-attend-a-developer-event-in-2026)