> Markdown version of [/jobs/ext/2281164-senior-solutions-architect-cluster-design-and-architecture-networking](https://www.wearedevelopers.com/jobs/ext/2281164-senior-solutions-architect-cluster-design-and-architecture-networking). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Solutions Architect, Cluster Design and Architecture - Networking - **Company:** NVIDIA Corporation - **Location:** Santa Clara, CA, United States - **Experience:** Expert - **Salary:** $184,000.0 - $287,500.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Computing Platforms, Computer Clusters, Computer Engineering, Network Congestion, Distributed Computing Environment, Distributed Systems, Network Topologies, InfiniBand, Network Architecture, Network Planning and Design, Software Systems, Supercomputing, AI Infrastructure, Graphics Processing Unit (GPU), Computer Network Technologies, Information Technology - **Published:** August 28, 2026 - **Apply:** https://nvidia.wd5.myworkdayjobs.com/NVIDIAExternalCareerSite/job/US-CA-Santa-Clara/Senior-Solutions-Architect--Cluster-Design-and-Architecture---Networking_JR2008988 ## About the Role * BS, MS, or PhD in Computer Science, Electrical Engineering, Computer Engineering, Physics, or related field (or equivalent experience) * 8+ years of experience in network architecture, network design, network validation and troubleshooting * Proven expertise in designing large-scale distributed systems, AI clusters, or HPC infrastructure * Ability to translate complex engineering concepts into customer-ready documentation, diagrams, and reference material Ways to stand out from the crowd: * Experience leading large-scale AI Factory or HPC cluster bring-ups or builds * Hands-on experience with NVIDIA networking products including, but not limited to, Infiniband, Spectrum-X, BlueField, etc. * Knowledge of NCCL, MPI, and collective communication patterns in distributed training as it pertains to networking patterns and design * Background in network performance optimization, congestion control, and validation at scale * External customer facing skill-set and background ## Description NVIDIA is building the world's most groundbreaking and innovative accelerated computing platforms for AI and HPC. Because of our work, scientists, researchers, and engineers can push the boundaries of what's possible. We pioneered a supercharged form of computing that powers everything from breakthrough AI research to the world's fastest supercomputers. We are seeing a highly motivated Senior Solutions Architect to join the Cluster Design and Architecture team with a focus on networking technologies. As AI workloads scale to unprecedented levels, the network is the backbone that makes large compute clusters possible. In this role, you will be at the forefront of assisting with designs and architectures for next-generation networking solutions that connect thousands of GPUs and enable the world's most advanced AI supercomputers and enterprise AI infrastructure in the field. As a Solutions Architect, you will act as a key technical expert connecting NVIDIA's new networking technology builds. These include Infiniband, Spectrum-X, NVLink, and all software solutions. You will work directly between engineering and field teams to support customers with fast paced requirements. You will work on end-to-end cluster design, network topology and architecture optimization, performance modeling and validation. Your expertise will directly influence how the world's leading AI companies, cloud providers, hyperscalers, research institutions, and enterprises build their infrastructure. What you'll be doing: * Partner with internal engineering efforts in GPU cluster building and networking and convey architecture and guidelines information both direct to customer and with field teams supporting customers * Guide field teams and their customers in cluster design, weighing design principles but also complex, situational limitations to make the most performant and supportable GPU clusters possible * Work closely with field teams supporting customers to ensure successful first deployments with new products, including new network architectures and topologies * Feedback customer/field perspectives on networking development and workflows back to engineering teams building internal clusters and/or composing customer facing documentation on guidelines and service flows * Perform hands-on work to assist field teams debugging issues relating to network build, configuration, and performance, bringing to bear internal engineering expertise and known bugs ## Related Videos - [The Gashlycrumb Tinies of AI Networking You Must Know (or Languish!)](https://www.wearedevelopers.com/videos/2067-the-gashlycrumb-tinies-of-ai-networking-you-must-know-or-languish) - [Rethinking Intelligence: AI, Accessibility, and the Future of Inclusive Work - Artur Ortega](https://www.wearedevelopers.com/videos/1377-rethinking-intelligence-ai-accessibility-and-the-future-of-inclusive-work-artur-ortega) - [Quantum Tech: Preparing for the Next Leap](https://www.wearedevelopers.com/videos/1693-quantum-tech-preparing-for-the-next-leap) - [AI Factories at Scale](https://www.wearedevelopers.com/videos/1139-ai-factories-at-scale) - [A Deep Dive on How To Leverage the NVIDIA GB200 for Ultra-Fast Training and Inference on Kubernetes](https://www.wearedevelopers.com/videos/1625-a-deep-dive-on-how-to-leverage-the-nvidia-gb200-for-ultra-fast-training-and-inference-on-kubernetes) - [Your Next AI Needs 10,000 GPUs. Now What?](https://www.wearedevelopers.com/videos/1590-your-next-ai-needs-10-000-gpus-now-what) ## Related Articles - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Top 6 Hackathons for Developers in 2023](https://www.wearedevelopers.com/magazine/263-top-6-hackathons-for-developers-in-2023)