> Markdown version of [/jobs/ext/610266-senior-architect-ai-solutions-engineering](https://www.wearedevelopers.com/jobs/ext/610266-senior-architect-ai-solutions-engineering). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Architect, AI Solutions Engineering - **Company:** NVIDIA Ltd. - **Location:** Santa Clara, CA, United States - **Experience:** Expert - **Salary:** $224,000.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Microsoft Windows, Application Programming Interfaces (APIs), Artificial Intelligence, Cascading, Cloud Computing, Computer Programming, Databases, Linux, Distributed Systems, Elasticsearch, Apache Hadoop, Python (Programming Language), Machine Learning, MongoDB, MySQL, NoSQL, OpenStack, Perforce, Performance Tuning, Software Product Management, Cloud Services, Standard Sql, Service-Oriented Architecture, Shell Script, Software Engineering, Software Systems, Subsystems, System Programming, Virtual Machines, Ceph (Software), Graphics Processing Unit (GPU), Large Language Models, Deep Learning, Multi-Cloud, Generative AI, Git, Kubernetes, Cassandra, Apache Kafka, Multiaccess Edge Computing, Lxc, Puppet, Restful APIs, Docker - **Published:** June 23, 2026 - **Apply:** https://www.juju.com/job/00000000gagkih ## About the Role + BS EE/CS or equivalent experience with 12+ years of systems software development with at least 1 year of experience in developing/exploring AI. + Development with Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), Fine-Tuning LLMs, AI Agentic workflows, LangChain, LangGraphs, and Cascading models. + Experience in deploying in hybrid, multi-cloud architecture and edge computing. + Extensive experience architecting and shipping large-scale distributed software systems. + Ability to identify gaps and bottlenecks, and develop solutions to optimize performance. + Strong programming and software development skills in JAVA, Python, Shell-script along with good understanding of distributed systems and REST APIs. + Experience in working with SQL/NoSQL database systems such as MySQL, Cassandra, MongoDB or Elasticsearch. + Excellent knowledge and working experience with Docker containers and Virtual Machines. + Good background of Cloud technologies like: OpenStack, Docker, Kubernetes, Chef/Puppet, Hadoop/Ceph/SwiftStack, LXC, Git, Perforce, JFrog, Kafka. + Ability to work across organizational boundaries optimally to improve alignment and productivity between teams in a multi-national, multi-time-zone corporate environment. Ways to stand out from the crowd: + MS or PhD in EE/CS + Depth in AI, Machine Learning and Deep Learning algorithms and techniques. + Strong collaborative and interpersonal skills, with a consistent record of guiding and influencing others in dynamic environments. + Experience developing large-scale software systems using service-oriented architecture under real-time performance requirements. + Background in designing high-performance, scalable software systems with a strong focus on hardware cost optimization. ## Description NVIDIA is seeking an AI Solutions Architect to join its Infrastructure Planning and Process Team! This role will focus on the extensive scale-up of key AI solutions for NVIDIA's internal cloud infrastructure. IPP (Infrastructure, Planning and Process) is a global organization within NVIDIA, working closely with various teams such as Graphics Processors, Mobile Processors, Deep Learning, Artificial Intelligence, and Driverless Cars to meet their infrastructure needs. The cloud services support nearly half a million automated jobs daily on five thousand servers, enhancing the productivity of thousands of NVIDIA software developers worldwide. The cloud hosts a diverse mix of machines and devices with various operating systems (Windows/Linux/Android) and hardware platforms, including NVIDIA GPUs and Tegra processors. As an AI Solutions Architect, you will manage the tools NVIDIAns use to deliver solutions quickly, and identify any gaps in these tools. You will also understand overall movement of data in the entire platform, identifying bottlenecks, defining solutions, developing key pieces, writing APIs, and owning deployment. You will collaborate with internal and external development teams to discover opportunities and solve complex problems. Your role will also involve guiding engineers in solving complex problems, developing acceptance tests, and reviewing their work and test results. Exceptional technical leadership, communication, organizational, and analytical skills are required, along with a passion for solving large and complex problems, e.g. Peta Bytes of fast storage, Million cores, 100,000 builds and 100,000 tests. What you'll be doing: + Serve as an Architect developing internal AI systems used by thousands of NVIDIANs globally. + Identify gaps and issues and resolve ones are better suited for AI solutions versus conventional approaches. + Further divide the AI category into 'buy versus build' options by researching available tools in the market. + Align with teams across Nvidia to establish overall AI system goals and break them down into specific objectives for each sub-system. + Drive, motivate, convince, and mentor sub-system leads to achieve improvements with agility and speed. + Identify performance bottlenecks and optimize the speed and cost efficiency of AI development and testing systems. + Drive the planning of software/hardware capacity, covering both internal and public cloud, addressing the balance between time and utilization. + Introduce technologies enabling massively parallel systems to improve turnaround time by an order of magnitude. + Collaborate with AI product vendors to gain deep insights of the AI industry, and share them with leaders and developers internally. ## Related Videos - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Your Next AI Needs 10,000 GPUs. Now What?](https://www.wearedevelopers.com/videos/1590-your-next-ai-needs-10-000-gpus-now-what) - [Tomorrow's cloud data platforms - fully managed database-as-a-service (DBaaS)](https://www.wearedevelopers.com/videos/254-tomorrow-s-cloud-data-platforms-fully-managed-database-as-a-service-dbaas) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) - [AI That Fits Your Business, Not the Other Way Around](https://www.wearedevelopers.com/videos/100148-ai-that-fits-your-business-not-the-other-way-around) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models)