> Markdown version of [/jobs/ext/2497535-senior-solutions-architect-higher-education-and-research-open-models-and-llm](https://www.wearedevelopers.com/jobs/ext/2497535-senior-solutions-architect-higher-education-and-research-open-models-and-llm). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Solutions Architect, Higher Education and Research - Open Models and LLM - **Company:** Nvidia - **Location:** Reading, UK - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Nvidia CUDA, Microprocessors, Python (Programming Language), Graphics Processing Unit (GPU), Large Language Models, Deep Learning, Machine Learning Operations, TensorRT, Nim (Programming Language) - **Published:** August 29, 2026 - **Apply:** https://uk.indeed.com/viewjob?jk=7bc5a467c81e3b0c ## About the Role * A graduate degree from a leading university in a STEM related discipline. * 5+ years of hands-on experience running the large language model lifecycle across multi-node systems: training, fine-tuning, inference, serving, and/or agentic workflows. * Strong collaboration and communication skills, with the ability to build relationships with academic and research stakeholders, and to communicate complex ideas clearly to both expert and non-expert audiences. * Action oriented, analytical, self-motivated, and a passionate learner, with excellent organization skills to work in a heavily multi-tasked environment. * Fluent in English, both oral and written, and comfortable working in Python., * A PhD from a leading university in a STEM related discipline, with 3+ years of research on large language models or foundation models and their applications. * A track record of thought leadership: high-impact publications, talks at academic conferences and workshops, as well as good understanding of scientific policy engagement, grant processes, or national/European research program structures. * Experience with NVIDIA's AI software stack, powered by CUDA and CUDA-X libraries, e.g. NVIDIA AI Enterprise, NeMo Framework, Megatron Bridge, NIM, TensorRT-LLM, Dynamo, NeMo Agent Toolkit, and Triton Inference Server, as well as the Nemotron open-model methodology. ## Description * Partner directly with leading research labs, researchers, and university faculty as a trusted technical advisor: develop a keen understanding of their scientific goals and drive joint research projects at supercomputing scale. * Identify and accelerate high-impact workloads by integrating NVIDIA's frameworks, libraries, and core software stack into research projects, and help researchers amplify their impact through publications, conference presentations, and technical content. * Advocate for accelerated computing and Deep Learning, and deliver hands-on trainings, workshops, lectures, and demonstrations across NVIDIA's platforms, and mentor power users to become NVIDIA champions. * Track emerging research trends and turn gaps between researcher needs and NVIDIA's o erings into prototypical solutions and direct feedback to NVIDIA Engineering. * Maintain deep expertise in your domain while staying versatile across NVIDIA's full platform: GPUs, CPUs, networking, and software ## Related Videos - [Tour de Force: Open-Source LLM Inference Optimization from Simple to Sophisticated](https://www.wearedevelopers.com/videos/100099-tour-de-force-open-source-llm-inference-optimization-from-simple-to-sophisticated) - [Coffee with Developers - Stephen Jones - NVIDIA](https://www.wearedevelopers.com/videos/1303-coffee-with-developers-stephen-jones-nvidia) - [Nemotron: NVIDIA's open model strategy for developers](https://www.wearedevelopers.com/videos/100064-nemotron-nvidia-s-open-model-strategy-for-developers) - [Getting Started with Machine Learning](https://www.wearedevelopers.com/videos/260-getting-started-with-machine-learning) - [Efficient deployment and inference of GPU-accelerated LLMs​](https://www.wearedevelopers.com/videos/929-efficient-deployment-and-inference-of-gpu-accelerated-llms) - [Trends, Challenges and Best Practices for AI at the Edge](https://www.wearedevelopers.com/videos/630-trends-challenges-and-best-practices-for-ai-at-the-edge) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [The Best Large Language Models on The Market](https://www.wearedevelopers.com/magazine/319-the-best-large-language-models-on-the-market) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production)