> Markdown version of [/jobs/ext/3675195-ai-performance-engineer](https://www.wearedevelopers.com/jobs/ext/3675195-ai-performance-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # AI Performance Engineer - **Company:** Graphcore Limited - **Location:** Austin, TX, United States - **Experience:** Expert - **Contract:** Temporary to permanent - **Skills:** Artificial Intelligence, Systems Engineering, Big Data, C++ (Programming Language), Profiling, Communication Softwares, Distributed Systems, Python (Programming Language), Machine Learning, Performance Tuning, Remote Direct Memory Access, Tensorflow, Systems Architecture, High Performance Computing - **Published:** October 10, 2026 - **Apply:** https://job-boards.greenhouse.io/graphcore/jobs/8870170002 ## About the Role · Strong experience profiling and optimizing AI, machine learning or high-performance computing workloads. · Experience with distributed systems and communication libraries such as MPI, NCCL, UCX or libfabric. · Strong C++ and Python skills, including reliable tools or performance-sensitive software. · Deep understanding of compute, memory and communication behavior in large-scale systems. · Ability to own complex technical work and coordinate improvements across teams. · Familiarity with MLPerf, accelerated architectures, ML frameworks or high-performance interconnects. While we have outlined a set of requirements, we value transferable skills and diverse experiences. We also welcome engineers returning to the profession after a career break, including through returnship routes. ## Description As a Senior AI Performance Engineer, you will analyze and optimize AI training and inference workloads across large-scale distributed systems. Your work will connect compute, memory, communication and software behavior. You will own complex performance investigations from problem definition through validation. You will turn profiling, benchmarking and modeling results into improvements engineers can build from. You will design benchmarks, investigate system bottlenecks and optimize distributed communication software across technologies such as MPI, NCCL, UCX and RDMA. Your work will help Graphcore deliver efficient, reliable AI systems for large-scale deployment. This role gives you rare scope across hardware, software, networking and system architecture. It is based in Austin, Texas. The team and culture The System Engineering Performance team architects, evaluates and optimizes high-performance infrastructure for large-scale data center deployments. The team works across the computing stack to understand real system behavior. Work moves through evidence, ownership and clear technical judgment. You will define investigations, coordinate across teams and validate impact with reliable performance data. Decisions are shaped through benchmarks, models, simulations and practical engineering tradeoffs. The team values engineers who think big, act fast, take responsibility, speak up and lead beyond their own area.