> Markdown version of [/jobs/ext/2717801-senior-solution-architect-ai-gpu-cloud](https://www.wearedevelopers.com/jobs/ext/2717801-senior-solution-architect-ai-gpu-cloud). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Solution Architect - AI / GPU Cloud - **Company:** GMI Cloud - **Location:** Mountain View, CA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Cloud Computing, Computer Clusters, Data Centers, Distributed Computing Environment, InfiniBand, Node.Js, Performance Tuning, AI Infrastructure, High Performance Computing, Kubernetes, Storage Technologies, Slurm, Machine Learning Operations - **Published:** September 4, 2026 - **Apply:** https://www.disabledperson.com/jobs/74866474-senior-solution-architect-ai-gpu-cloud ## About the Role Technical Background * 5-10+ years in cloud infrastructure, GPU cloud, HPC, AI/ML infrastructure, or data center engineering. * Strong understanding of: * Distributed training & inference architectures * Kubernetes, Slurm, or other cluster/orchestration systems * NVIDIA GPU stack (H100/H200/B200/GB200 or similar) * InfiniBand / high-speed networking * Storage architectures for AI workloads Customer-Facing Skills * Experience working directly with enterprise or hyperscaler technical teams. * Ability to simplify complex infra concepts for both technical and non-technical audiences. * Strong communication, solution-design, and project coordination skills. Soft Skills * Self-starter, ownership mindset, excellent follow-through. * Comfortable working in a fast-moving, high-growth environment. * Strong problem-solving and "architect + advisor" mentality. Preferred Qualifications (Nice to Have) * Hands-on with large-scale GPU deployments (multi-node, multi-cluster). * Exposure to hyperscaler capacity planning or AI infrastructure procurement teams. * Experience with multi-region or global GPU deployments (US + APAC/Taiwan). ## Description As a Solution Architect, you will be the primary technical interface for our enterprise and hyperscaler accounts. You will design GPU-cloud and AI infrastructure solutions, lead PoCs and benchmarks, guide customers through deployment, and partner closely with internal engineering, infra, and operations teams to ensure successful delivery. This role is ideal for someone who understands large-scale AI/ML/HPC workloads, enjoys working directly with customers, and wants to shape the future of AI infrastructure., Customer Engagement & Technical Leadership * Serve as the primary technical point-of-contact for enterprise and hyperscaler customers. * Deeply understand customer AI/ML/HPC workloads, scaling requirements, and deployment models. * Architect GPU clusters, storage, networking, and orchestration solutions tailored to customer needs. Solution Design & PoC Execution * Lead Proof-of-Concepts, benchmarks, and workshops demonstrating performance, reliability, and scalability. * Produce technical proposals, architecture diagrams, capacity plans, and cost/performance recommendations * Translate complex technical issues into clear actions for both engineering and business stakeholders. Deployment & Enablement * Guide customers through onboarding, cluster setup, performance tuning, and scaling. * Partner with internal Infra, DC Ops, and Engineering teams to ensure smooth delivery and implementation. * Identify optimization opportunities in customer workloads (GPU utilization, networking, scheduling, cost). Ongoing Support & Relationship Building * Act as a trusted advisor on GPU/AI infrastructure best practices, roadmap, and long-term planning. * Maintain regular technical check-ins, capacity reviews, and performance reviews with customers. * Gather customer feedback and collaborate with product/engineering to improve our platform. ## Related Videos - [Running Secure Life Science Research at Scale using Hybrid GPU HPC and Kubernetes 🧬](https://www.wearedevelopers.com/videos/100355-running-secure-life-science-research-at-scale-using-hybrid-gpu-hpc-and-kubernetes) - [The Gashlycrumb Tinies of AI Networking You Must Know (or Languish!)](https://www.wearedevelopers.com/videos/2067-the-gashlycrumb-tinies-of-ai-networking-you-must-know-or-languish) - [Stop using Node.js like in 2020! What changed and what you can do today with Node.js](https://www.wearedevelopers.com/videos/100011-stop-using-node-js-like-in-2020-what-changed-and-what-you-can-do-today-with-node-js) - [AI Factories at Scale](https://www.wearedevelopers.com/videos/1139-ai-factories-at-scale) - [Stop Using Node.js Like It’s 2020! - Alfonso Graziano](https://www.wearedevelopers.com/videos/1863-stop-using-node-js-like-it-s-2020-alfonso-graziano) - [Your Next AI Needs 10,000 GPUs. Now What?](https://www.wearedevelopers.com/videos/1590-your-next-ai-needs-10-000-gpus-now-what) ## Related Articles - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Stephan Gillich - Bringing AI Everywhere](https://www.wearedevelopers.com/magazine/489-stephan-gillich-bringing-ai-everywhere) - [What Industries Outside of AI Are Hiring The Most AI Experts?](https://www.wearedevelopers.com/magazine/98-what-industries-outside-of-ai-are-hiring-the-most-ai-experts) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering)