> Markdown version of [/jobs/ext/2472782-sr-ai-inference-platform-engineer](https://www.wearedevelopers.com/jobs/ext/2472782-sr-ai-inference-platform-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Sr. AI Inference Platform Engineer - **Company:** Apple Inc. - **Location:** Seattle, WA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Data Analysis, C++ (Programming Language), Continuous Integration, Distributed Systems, Python (Programming Language), Prometheus, Software Engineering, AI Infrastructure, Data Logging, Large Language Models, Grafana, Kubernetes, Information Technology, Data Analytics, TensorRT, Splunk, Data Pipelines, Golang, Programming Languages - **Published:** August 13, 2026 - **Apply:** https://www.seattlejobs.com/job.asp?id=3351754374&tx=IT9184THV&pt=1&aff=0B19D771-A501-4A5E-8338-2A822B784D54&utm_source=Job%20Feed&utm_medium=textkernel&utm_campaign=DE&utm_term=0B19D771-A501-4A5E-8338-2A822B784D54 ## About the Role * BS or MS in Computer Science or related technical field. * Solid understanding of AI/ML inference architecture and the performance characteristics of serving systems. * 7 or more years of experience with performance and infrastructure engineering in distributed systems. * 7 years of experience coding in Python, Go, C++, or other programming languages. * Experience with automation engineering, tooling, and data pipelines to support engineering workflows. * Strong knowledge of GPU/accelerator architecture as it relates to AI workloads. * Practical statistical knowledge applicable to performance analysis and forecasting. * Excellent communication skills and ability to turn data into clear guidance for infrastructure teams and capacity planners., * Experience with performance benchmarking and methodologies for AI/ML inference systems. * Familiarity with capacity planning and forecasting/projection models for large-scale infrastructure. * Experience with GPU profiling and observability tools (e.g., Nsight, other vendor-specific profilers). * Experience with data visualization and reporting tools/frameworks for surfacing performance trends to stakeholders. * Familiarity with ML serving frameworks and runtimes (e.g., Triton, TensorRT-LLM, vLLM, or similar). * Experience with CI/CD and workflow orchestration tools for building automated performance analysis pipelines. * Knowledge of cluster schedulers and orchestration platforms (e.g., Kubernetes). * Experience with metrics and logging tools (e.g., Prometheus, Grafana, Splunk). ## Description We are looking for senior engineer to build tooling, automation, and analysis capabilities that strengthen our AI inference platform. This role will focus on developing sophisticated performance benchmarking systems, capacity projection models, and data analysis pipelines that directly inform our AI infrastructure teams and capacity planners. You'll work at the intersection of AI systems performance, distributed infrastructure, and software engineering to help the team make data-driven decisions about scaling and optimizing our inference platform. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [AI Factories at Scale](https://www.wearedevelopers.com/videos/1139-ai-factories-at-scale) - [All your telemetry data from any source in one place](https://www.wearedevelopers.com/videos/57-all-your-telemetry-data-from-any-source-in-one-place) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Stephan Gillich - Bringing AI Everywhere](https://www.wearedevelopers.com/magazine/489-stephan-gillich-bringing-ai-everywhere) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production)