> Markdown version of [/jobs/ext/2309395-on-device-ml-performance-engineer-graphics-games-and-machine-learning](https://www.wearedevelopers.com/jobs/ext/2309395-on-device-ml-performance-engineer-graphics-games-and-machine-learning). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # On-device ML Performance Engineer, Graphics, Games and Machine Learning - **Company:** Apple Inc. - **Location:** Seattle, WA, United States - **Salary:** $142,300.0 - **Contract:** Permanent contract - **Skills:** Apple Mac Systems, Apple Xcode, Application Frameworks, C++ (Programming Language), Compilers, Software Debugging, Linux, Python (Programming Language), Machine Learning, Open Source Technology, Tensorflow, Shell Script, Scripting, Pytorch, Large Language Models, Swift (Programming Language), Information Technology, ONNX (Open Neural Network Exchange) Format, Objective C++ - **Published:** August 30, 2026 - **Apply:** https://www.jobmonkeyjobs.com/career/27977308/On-Device-Ml-Performance-Engineer-Graphics-Games-Machine-Learning-Washington-Seattle-7413 ## About the Role Experience with ML inference, quantization, performance and accuracy Familiarity and experience with the most popular ML architectures (e.g. LLM's, Diffusion models, CNN's) A passion to explore and learn about the latest advances in ML model design and architecture, particularly as related to model implementation on HW and on-device inference Familiarity with Operating Systems, embedded systems, and CPU/GPU HW architectures Highly proficient in Python/C++ and shell scripting Familiarity with Linux or macOS Exceptional clarity in verbal and written communication, including the ability to present and lead discussions in larger groups Preferred Qualifications Masters or PhDs in Computer Science or relevant disciplines. Experience with Apple's CoreML, MPS Graph, Metal Performance Shader's or MLX frameworks Experience with any on-device ML stack, such as TFLite, ONNX, ExecuTorch, etc. Experience with any ML authoring framework (PyTorch, TensorFlow, JAX, etc.). Experience with Apple's App development framework such as Xcode, Swift, Objective-C Experience with any compiler stack (MLIR/LLVM/TVM etc.) ## Description As an engineer in this role, you will be primarily focused on analyzing and optimizing the performance of the latest ML models on the latest iPhones and Mac's. You will work with models created by the most popular ML frameworks (PyTorch, MLX, etc) and will analyze the inference of those models on device to ensure the stack achieves full machine performance on Apple Silicon. The role also includes scripting, coding, and generation of utilities and debug tools to extract, analyze, and report performance and power related metrics for Apple HW. The ideal candidate will have a passion for ML model architectures and ML inference, deep knowledge of GPU and CPU, computer architecture and memory, compilers and HW drivers. Responsibilities Driving the on-device performance analysis of Apple SoC's and ML SW stack across a wide range of Apple internal or open-source ML models Optimize model conversion, compilation and on-device inference for Apple SoC's, achieving objectives such as performance, memory, and energy efficiency Developing tools and scripts to generate and analyze ML performance data Work across multiple teams and organizations to support the design and delivery of best in class on-device ML hardware and software stack Generate and present ML performance data to internal and external teams and stakeholders ## Related Videos - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Edge AI on iOS: Beyond the Cloud, Designing the Next Generation of Intelligent On-Device Apps](https://www.wearedevelopers.com/videos/100225-edge-ai-on-ios-beyond-the-cloud-designing-the-next-generation-of-intelligent-on-device-apps) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [JavaScript? No. Java Scripts! - Scripting with Java](https://www.wearedevelopers.com/videos/2094-javascript-no-java-scripts-scripting-with-java) - [Serverless deployment of (large) NLP models ](https://www.wearedevelopers.com/videos/158-serverless-deployment-of-large-nlp-models) - [Harnessing Apple Intelligence: Live Coding with Swift for iOS](https://www.wearedevelopers.com/videos/1515-harnessing-apple-intelligence-live-coding-with-swift-for-ios) ## Related Articles - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [ Dev Digest 213: Petrol Prices, Agentic Workflows, AI Skills and CODE100!](https://www.wearedevelopers.com/magazine/718-dev-digest-213-petrol-prices-agentic-workflows-ai-skills-and-code100) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Dev Digest 118 - not a total recall](https://www.wearedevelopers.com/magazine/452-dev-digest-118-not-a-total-recall) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)