> Markdown version of [/jobs/ext/2892812-founding-gpu-engineer](https://www.wearedevelopers.com/jobs/ext/2892812-founding-gpu-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Founding GPU Engineer - **Company:** Descriptionfuse Energy - **Location:** London, UK - **Contract:** Permanent contract - **Skills:** Grid System, Artificial Intelligence, Systems Engineering, C++ (Programming Language), Profiling, Nvidia CUDA, Data Centers, Distributed Systems, InfiniBand, Python (Programming Language), Node.Js, PCI Express, Remote Direct Memory Access, Network Switches, Graphics Processing Unit (GPU), Kubernetes, Low Latency, Slurm - **Published:** September 14, 2026 - **Apply:** https://www.apply4u.co.uk/jobs/founding-gpu-engineer/47212282 ## About the Role infrastructure teams on power capping, dynamic voltage/frequency scaling, and workload scheduling strategies that reduce energy cost and carbon intensityCollaborate with ML/systems engineers to integrate custom kernels into training/inference pipelinesBenchmark against CPU/GPU baselines and drive continuous performance improvementsContribute to internal libraries, documentation, and best practices for GPU performance engineeringRequirements4+ years writing production CUDA code, or equivalent strong project/industry experienceDeep understanding of GPU architecture (SMs, warps, memory hierarchy, occupancy)Proficiency in C++ and CUDA; experience with Python for tooling/orchestrationExperience with performance profiling tools (Nsight Systems/Compute)Familiarity with multi-GPU/multi-node scaling (NCCL, MPI, RDMA/InfiniBand)Strong grasp of memory optimisation, kernel fusion and parallel algorithm designComfortable working across the stack, from low-level kernels to system-level infrastructureBonus: Triton, cuDNN, cuBLAS or custom ML inference/training frameworks; data centre power/thermal management or demand-response systems; HPC, quantitative finance or large-scale distributed systems; Kubernetes/Slurm for GPU cluster orchestration; interest in energy markets, grid systems or sustainability-focused computeBenefitsCompetitive salary and eligibility for equityBiannual bonus schemeFully expensed tech to match your needsPrivate health insuranceBreakfast and dinner allowance for office-based employeesAs we hire globally, benefits vary by location.Job SummaryID: 41EF45DDDFDepartment:Type: full time ## Description high-performance compute infrastructure at the intersection of energy and AI, optimising how power-dense GPU workloads are scheduled, cooled and balanced against grid conditions in real time. We're looking for a Founding GPU Engineer to develop and optimise GPU-accelerated software for data centre systems: low-level performance engineering for large-scale compute clusters, tying GPU workload behaviour to energy availability and grid demand. This puts CUDA/GPU performance engineering at the centre of how Fuse scales its compute infrastructure.ResponsibilitiesDesign, implement, and optimise CUDA kernels for high-throughput, latency-sensitive workloadsProfile and tune GPU performance across compute, memory bandwidth, and interconnect (NVLink/PCIe) bottlenecksBuild tooling to correlate GPU cluster power draw and utilisation with real-time energy pricing and grid signalsOptimise multi-GPU and multi-node scaling using NCCL, MPI, or similar communication librariesWork with data center ## Related Videos - [Running Secure Life Science Research at Scale using Hybrid GPU HPC and Kubernetes 🧬](https://www.wearedevelopers.com/videos/100355-running-secure-life-science-research-at-scale-using-hybrid-gpu-hpc-and-kubernetes) - [Accelerating Python on GPUs](https://www.wearedevelopers.com/videos/859-accelerating-python-on-gpus) - [Profiling Symfony & PHP apps with Blackfire](https://www.wearedevelopers.com/videos/265-profiling-symfony-php-apps-with-blackfire) - [Stop using Node.js like in 2020! What changed and what you can do today with Node.js](https://www.wearedevelopers.com/videos/100011-stop-using-node-js-like-in-2020-what-changed-and-what-you-can-do-today-with-node-js) - [Accelerating Python on GPUs](https://www.wearedevelopers.com/videos/1521-accelerating-python-on-gpus) - [Coffee with Developers - Stephen Jones - NVIDIA](https://www.wearedevelopers.com/videos/1303-coffee-with-developers-stephen-jones-nvidia) ## Related Articles - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [What’s the latest in NVIDIA CUDA Python](https://www.wearedevelopers.com/magazine/568-what-s-the-latest-in-nvidia-cuda-python) - [The Fastest-Growing Tech Sectors to Look Out for in 2025](https://www.wearedevelopers.com/magazine/373-the-fastest-growing-tech-sectors-to-look-out-for-in-2025) - [Dev Digest 157: CUDA in Python, Gemini Code Assist and Back-dooring LLMs](https://www.wearedevelopers.com/magazine/557-dev-digest-157-cuda-in-python-gemini-code-assist-and-back-dooring-llms)