> Markdown version of [/jobs/ext/1743083-senior-software-architect-ai-systems](https://www.wearedevelopers.com/jobs/ext/1743083-senior-software-architect-ai-systems). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Software Architect, AI Systems... - **Company:** NVIDIA Ltd. - **Location:** Santa Clara, CA, United States - **Experience:** Expert - **Salary:** $224,000.0 - **Contract:** Permanent contract - **Skills:** Amazon S3, C++ (Programming Language), Profiling, Computer Engineering, Software Debugging, Distributed Computing Environment, InfiniBand, Storage Area Network (SAN), Remote Direct Memory Access, Software Engineering, System Programming, Reinforcement Learning, Large Language Models, Information Technology, Machine Learning Operations, TensorRT, Nvme - **Published:** July 12, 2026 - **Apply:** https://www.juju.com/job/00000000gfrx8b ## About the Role + 12+ years in systems software and/or networking with demonstrated ownership of complex projects. + MS, PhD or equivalent experience in Computer Science, Computer Engineering, Electrical Engineering, or a related field. + Solid understanding of high-performance networking: InfiniBand, RoCE, RDMA, NVLink, GPUDirect. + Strong C/C++/Rust systems programming with comfort in performance profiling and low-level debugging. + Understanding of ML systems concepts-transformer architectures, KV cache mechanics, model parallelism, or distributed training and inference patterns. Ways to stand out from the crowd: + Knowledge of ML inference frameworks (vLLM, SGLang, TensorRT-LLM) and their communication requirements. + Knowledge of storage networking (NVMe-oF, GPUDirect Storage, S3). + Background of Reinforcement Learning systems. ## Related Videos - [Tour de Force: Open-Source LLM Inference Optimization from Simple to Sophisticated](https://www.wearedevelopers.com/videos/100099-tour-de-force-open-source-llm-inference-optimization-from-simple-to-sophisticated) - [Coffee with Developers - Stephen Jones - NVIDIA](https://www.wearedevelopers.com/videos/1303-coffee-with-developers-stephen-jones-nvidia) - [Profiling Symfony & PHP apps with Blackfire](https://www.wearedevelopers.com/videos/265-profiling-symfony-php-apps-with-blackfire) - [Efficient deployment and inference of GPU-accelerated LLMs​](https://www.wearedevelopers.com/videos/929-efficient-deployment-and-inference-of-gpu-accelerated-llms) - [Trends, Challenges and Best Practices for AI at the Edge](https://www.wearedevelopers.com/videos/630-trends-challenges-and-best-practices-for-ai-at-the-edge) - [Your Next AI Needs 10,000 GPUs. Now What?](https://www.wearedevelopers.com/videos/1590-your-next-ai-needs-10-000-gpus-now-what) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models)