> Markdown version of [/jobs/ext/2412374-senior-solutions-engineer](https://www.wearedevelopers.com/jobs/ext/2412374-senior-solutions-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Solutions Engineer - **Company:** L.J.B & Co. Construction Recruitment - **Location:** London, UK (Remote available) - **Experience:** Expert - **Salary:** £260,000.0 - £312,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Systems Engineering, Computer Clusters, Data Centers, InfiniBand, Performance Tuning, AI Infrastructure, High Performance Computing, Kubernetes, Bare Metal, Slurm, Hardware Infrastructure - **Published:** August 18, 2026 - **Apply:** https://find.jobs/jobs-near-me/apply/ats-redirect/?id=2932102359-2 ## About the Role * 5+ years in Solution Architecture, Systems Engineering, Technical Pre-Sales, HPC or AI infrastructure. * Strong hands-on knowledge of NVIDIA GPU infrastructure. * Experience with HGX / DGX / Blackwell / B300 / GB300 / GB200. * Strong knowledge of NVLink / NVSwitch. * Expert knowledge of InfiniBand and/or RoCE/RoCEv2. * Experience with GPU clusters, HPC or AI infrastructure. * Knowledge of Kubernetes and/or Slurm. * Experience producing HLDs, LLDs, architecture diagrams and BOMs. * Strong customer-facing and technical presentation skills. * Understanding of high-density data centre power and cooling requirements. ## Description We are looking for a Senior Solution Engineer GPU & AI Infrastructure to support the design of cutting-edge AI and high-performance computing infrastructure. This is a senior Solution Architecture / Technical Pre-Sales role focused on NVIDIA GPU infrastructure, AI workloads, HPC environments and high-speed networking. The engagement will initially be 2 days per week, with the potential to increase to 5 days per week as requirements develop. Extremely Flexible Working We can offer a highly flexible working arrangement around your existing commitments. * Fully remote * Flexible working hours * Evening and weekend availability can be accommodated * Work can be structured around your existing role or commitments * Initially 2 days per week * Potential to increase to 5 days per week * Flexible approach to when the work is completed, provided agreed deliverables and deadlines are met Key Responsibilities * Design large-scale NVIDIA GPU clusters for AI and HPC workloads. * Develop HLDs, LLDs, rack designs and detailed BOMs. * Design NVLink / NVSwitch architectures. * Design high-speed InfiniBand and RoCE/RoCEv2 networking. * Develop infrastructure solutions using NVIDIA Blackwell, B300, GB300 and GB200 platforms. * Design both bare-metal and Kubernetes-based GPU environments. * Work with Slurm, Kubernetes, NVIDIA GPU Operator, NCCL and GPUDirect. * Design high-performance storage solutions for AI workloads. * Lead technical discussions with CTOs, AI leaders and infrastructure teams. * Support RFPs, RFIs, technical proposals and customer presentations. * Lead technical workshops and Proof of Concept deployments. * Support GPU cluster benchmarking and performance optimisation. ## Related Videos - [Running Secure Life Science Research at Scale using Hybrid GPU HPC and Kubernetes 🧬](https://www.wearedevelopers.com/videos/100355-running-secure-life-science-research-at-scale-using-hybrid-gpu-hpc-and-kubernetes) - [The Gashlycrumb Tinies of AI Networking You Must Know (or Languish!)](https://www.wearedevelopers.com/videos/2067-the-gashlycrumb-tinies-of-ai-networking-you-must-know-or-languish) - [Understanding Kubernetes in a visual way](https://www.wearedevelopers.com/videos/100085-understanding-kubernetes-in-a-visual-way) - [AI Factories at Scale](https://www.wearedevelopers.com/videos/1139-ai-factories-at-scale) - [Instant KAI Sandboxes with vCluster: Multi-Tenant, Multi-Scheduler GPU Sharing](https://www.wearedevelopers.com/videos/100333-instant-kai-sandboxes-with-vcluster-multi-tenant-multi-scheduler-gpu-sharing) - [Your Next AI Needs 10,000 GPUs. Now What?](https://www.wearedevelopers.com/videos/1590-your-next-ai-needs-10-000-gpus-now-what) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)