> Markdown version of [/jobs/ext/2857065-founding-gpu-engineer](https://www.wearedevelopers.com/jobs/ext/2857065-founding-gpu-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Founding GPU Engineer - **Company:** Opportunitydemand - **Location:** London, UK - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Grid System, Systems Engineering, C++ (Programming Language), Profiling, Nvidia CUDA, Data Centers, InfiniBand, Python (Programming Language), PCI Express, Remote Direct Memory Access, Network Switches, Graphics Processing Unit (GPU), Kubernetes, Low Latency, Slurm - **Published:** September 12, 2026 - **Apply:** https://www.apply4u.co.uk/jobs/founding-gpu-engineer/46632600 ## About the Role training/inference pipelines.Benchmark against CPU/GPU baselines and drive continuous performance improvements.Contribute to internal libraries, documentation, and best practices for GPU performance engineering.4+ years of experience writing production CUDA code, or equivalent strong project/industry experience.Deep understanding of GPU architecture (SMs, warps, memory hierarchy, occupancy).Proficiency in C++ and CUDA; experience with Python for tooling/orchestration.Experience with performance profiling tools (Nsight Systems/Compute).Familiarity with multi-GPU/multi-node scaling (NCCL, MPI, RDMA/InfiniBand).Strong grasp of memory optimisation, kernel fusion, and parallel algorithm design.Comfortable working across the stack from low-level kernels to system-level infrastructure.Nice to HaveExperience with Triton, cuDNN, cuBLAS, or custom ML inference/training frameworks.Exposure to data center power/thermal management or demand-response systems.Background in HPC, quantitative finance, or large-scale distributed systems.Familiarity with Kubernetes/Slurm for GPU cluster orchestration.Interest or experience in energy markets, grid systems, or sustainability-focused compute.Competitive salary and an equity sign-on bonus.Biannual bonus scheme.Fully expensed tech to match your needs.Breakfast and dinner allowance for office based employees. #J-18808-Ljbffr ## Description The OpportunityDemand for high-performance compute capacity across the markets we operate in significantly outpaces what we can currently build, meaning speed to power and reliability are critical to how we scale. This puts CUDA/GPU performance engineering at the center of how Fuse scales its compute infrastructure.ResponsibilitiesDesign, implement, and optimise CUDA kernels for high-throughput, latency-sensitive workloads.Profile and tune GPU performance across compute, memory bandwidth, and interconnect (NVLink/PCIe) bottlenecks.Build tooling to correlate GPU cluster power draw and utilisation with real-time energy pricing and grid signals.Optimise multi-GPU and multi-node scaling using NCCL, MPI, or similar communication libraries.Work with data center infrastructure teams on power capping, dynamic voltage/frequency scaling, and workload scheduling strategies that reduce energy cost and carbon intensity.Collaborate with ML/systems engineers to integrate custom kernels into ## Related Videos - [Running Secure Life Science Research at Scale using Hybrid GPU HPC and Kubernetes 🧬](https://www.wearedevelopers.com/videos/100355-running-secure-life-science-research-at-scale-using-hybrid-gpu-hpc-and-kubernetes) - [The Gashlycrumb Tinies of AI Networking You Must Know (or Languish!)](https://www.wearedevelopers.com/videos/2067-the-gashlycrumb-tinies-of-ai-networking-you-must-know-or-languish) - [Profiling Symfony & PHP apps with Blackfire](https://www.wearedevelopers.com/videos/265-profiling-symfony-php-apps-with-blackfire) - [Accelerating Python on GPUs](https://www.wearedevelopers.com/videos/859-accelerating-python-on-gpus) - [Enhancing Workload Security in Kubernetes](https://www.wearedevelopers.com/videos/356-enhancing-workload-security-in-kubernetes) - [Coffee with Developers - Stephen Jones - NVIDIA](https://www.wearedevelopers.com/videos/1303-coffee-with-developers-stephen-jones-nvidia) ## Related Articles - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [What’s the latest in NVIDIA CUDA Python](https://www.wearedevelopers.com/magazine/568-what-s-the-latest-in-nvidia-cuda-python) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers)