> Markdown version of [/jobs/ext/2727504-software-engineer-ai-compute-libraries-performance](https://www.wearedevelopers.com/jobs/ext/2727504-software-engineer-ai-compute-libraries-performance). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Software Engineer - AI Compute Libraries & Performance - **Company:** Graphcore - **Location:** Bristol, UK - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Basic Linear Algebra Subprograms, C++ (Programming Language), Profiling, Software Debugging, Linux, Python (Programming Language), Regression Testing, Pytorch, Hardware Acceleration - **Published:** September 5, 2026 - **Apply:** https://startup.jobs/senior-software-engineer-ai-compute-libraries-performance-graphcore-9561750 ## About the Role * Excellent programming and scripting skills using C++ and Python * Understanding of processor architectures and profiling on Linux * Possess excellent written and oral communication skills, good work ethics, high sense of team-work * Love to produce quality work and be a team player Desirable * Strong command of algorithmic performance - vectorisation, memory hierarchy, threading, lock-free patterns * Hands-on with at least one BLAS/DNN stack and able to read/extend kernels * Comfort with CPU micro-optimisations and numerical stability/trade-offs across FP32/FP16/BF16/FP8 * Experience integrating native code into PyTorch or similar (custom ops, extensions, dispatch keys) * ABI/API stability and packaging for Linux system (manylinux, wheels) ## Description As a Senior Software Engineer, you will create high-performance AI compute libraries for Graphcore's next-generation hardware. Your work will sit close to the hardware, where every design choice matters. You will own kernels for linear algebra and tensor operations, including GEMM, convolutions, reductions and fused operations. You will improve performance, correctness and numerical reliability across critical AI workloads. This role is for engineers who enjoy hard performance problems and care deeply about quality. You will help shape software that enables customers to get more from AI hardware., * Design and implement kernels for linear algebra and tensor ops (GEMM, batched GEMM, convolutions, reductions, elementwise and fused operations) in C++ * Own performance and correctness - add microbenchmarks, regression tests, numerics validation * Profile and optimise across for next generation of AI hardware - threading, cache locality, memory layout, and kernel launch efficiency. * Debug issues, resolve bugs and generally improve the quality and functionality of the product * Actively engage in and support Agile ways of working within the team * Mentor colleagues within the team, sharing knowledge and providing guidance where appropriate ## Related Videos - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Accelerating Python on GPUs](https://www.wearedevelopers.com/videos/859-accelerating-python-on-gpus) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Profiling Symfony & PHP apps with Blackfire](https://www.wearedevelopers.com/videos/265-profiling-symfony-php-apps-with-blackfire) - [Accelerating Python on GPUs](https://www.wearedevelopers.com/videos/1112-accelerating-python-on-gpus) - [Geometric deep learning for drug discovery](https://www.wearedevelopers.com/videos/264-geometric-deep-learning-for-drug-discovery) ## Related Articles - [Dev Digest 157: CUDA in Python, Gemini Code Assist and Back-dooring LLMs](https://www.wearedevelopers.com/magazine/557-dev-digest-157-cuda-in-python-gemini-code-assist-and-back-dooring-llms) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [What’s the latest in NVIDIA CUDA Python](https://www.wearedevelopers.com/magazine/568-what-s-the-latest-in-nvidia-cuda-python) - [ Dev Digest 213: Petrol Prices, Agentic Workflows, AI Skills and CODE100!](https://www.wearedevelopers.com/magazine/718-dev-digest-213-petrol-prices-agentic-workflows-ai-skills-and-code100) - [Dev Digest 112 - The True Crime of AI Development](https://www.wearedevelopers.com/magazine/421-dev-digest-112-the-true-crime-of-ai-development)