Senior Solutions Engineer

L.J.B & Co. Construction Recruitment
London, UK
18 days ago
Apply on find.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
£260,000.0 - £312,000.0
Working hours
Shift work
Job source

Tech stack

Artificial Intelligence Systems Engineering Computer Clusters Data Centers InfiniBand Performance Tuning AI Infrastructure High Performance Computing Kubernetes Bare Metal Slurm Hardware Infrastructure

Job description

We are looking for a Senior Solution Engineer GPU & AI Infrastructure to support the design of cutting-edge AI and high-performance computing infrastructure.

This is a senior Solution Architecture / Technical Pre-Sales role focused on NVIDIA GPU infrastructure, AI workloads, HPC environments and high-speed networking.

The engagement will initially be 2 days per week, with the potential to increase to 5 days per week as requirements develop.

Extremely Flexible Working

We can offer a highly flexible working arrangement around your existing commitments.

  • Fully remote
  • Flexible working hours
  • Evening and weekend availability can be accommodated
  • Work can be structured around your existing role or commitments
  • Initially 2 days per week
  • Potential to increase to 5 days per week
  • Flexible approach to when the work is completed, provided agreed deliverables and deadlines are met

Key Responsibilities

  • Design large-scale NVIDIA GPU clusters for AI and HPC workloads.
  • Develop HLDs, LLDs, rack designs and detailed BOMs.
  • Design NVLink / NVSwitch architectures.
  • Design high-speed InfiniBand and RoCE/RoCEv2 networking.
  • Develop infrastructure solutions using NVIDIA Blackwell, B300, GB300 and GB200 platforms.
  • Design both bare-metal and Kubernetes-based GPU environments.
  • Work with Slurm, Kubernetes, NVIDIA GPU Operator, NCCL and GPUDirect.
  • Design high-performance storage solutions for AI workloads.
  • Lead technical discussions with CTOs, AI leaders and infrastructure teams.
  • Support RFPs, RFIs, technical proposals and customer presentations.
  • Lead technical workshops and Proof of Concept deployments.
  • Support GPU cluster benchmarking and performance optimisation.

Requirements

  • 5+ years in Solution Architecture, Systems Engineering, Technical Pre-Sales, HPC or AI infrastructure.
  • Strong hands-on knowledge of NVIDIA GPU infrastructure.
  • Experience with HGX / DGX / Blackwell / B300 / GB300 / GB200.
  • Strong knowledge of NVLink / NVSwitch.
  • Expert knowledge of InfiniBand and/or RoCE/RoCEv2.
  • Experience with GPU clusters, HPC or AI infrastructure.
  • Knowledge of Kubernetes and/or Slurm.
  • Experience producing HLDs, LLDs, architecture diagrams and BOMs.
  • Strong customer-facing and technical presentation skills.
  • Understanding of high-density data centre power and cooling requirements.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on find.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:57 min

Routing cross-rack traffic seamlessly with NCCL

Kevin Klues Kevin Klues · World Congress 2025

2:22 min

Infrastructure barriers and compliance risks in research

Jeremy Murray Jeremy Murray · World Congress 2026 Europe

2:28 min

Understanding Kubernetes architecture and core cluster components

Marc Nimmerrichter · World Congress 2022

1:12 min

Addressing the competitive landscape of specialized hardware demands

Hazal Mestci +1 · Coffee With Developers

1:51 min

Managing GPU quotas and multi-tenancy with Kueue

Jeremy Murray Jeremy Murray · World Congress 2026 Europe

2:33 min

Architecting CUDA and the AI software stack

Michael Kagan Michael Kagan +1 · World Congress 2026 Europe

Videos

See all

Related articles

See all