> Markdown version of [/jobs/ext/1346356-staff-sr-machine-learning-engineer-ai-search-knowledge-platforms](https://www.wearedevelopers.com/jobs/ext/1346356-staff-sr-machine-learning-engineer-ai-search-knowledge-platforms). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff / Sr. Machine Learning Engineer, AI, Search & Knowledge Platforms - **Company:** Apple Inc. - **Location:** United States - **Experience:** Expert - **Salary:** $175,000.0 - $308,500.0 - **Contract:** Permanent contract - **Skills:** HTML, Java (Programming Language), Artificial Intelligence, Amazon Web Services, Big Data, Data Deduplication, Data Structures, Data Systems, Distributed Systems, Graph Database, Information Extraction, Python (Programming Language), Machine Learning, Scala (Programming Language), Software Engineering, Feature Engineering, Large Language Models, Apache Spark, Siri, Web Filtering, Containerization, Kubernetes, Information Technology, Cassandra, Machine Learning Operations, TensorRT, Docker, Golang, Microservices - **Published:** July 19, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=dcd6dfcbb10c88eb ## About the Role Experience with training and fine-tuning large language models Experience with optimizing ML training and serving performance, including GPU tuning, batch size optimization, and multi-node scheduling Familiarity with Nvidia TensorRT-LLM, vLLM, Nvidia Triton Server, or similar inference frameworks. Experience with NLP, information extraction, or web data systems. Excellent interpersonal skills, able to work independently as well as in a team Minimum Qualifications Bachelor's degree or higher in Computer Science or related technical field 3+ years of experience in software engineering or ML engineering Experience with Golang, Java, Scala, or Python Background in computer science: algorithms, data structures, and distributed systems Experience working in a cloud-native environment such as AWS Experience working with large-scale data processing pipelines (Spark, Cassandra, etc.) Experience with micro-service architecture in a containerized environment (Docker, Kubernetes, etc.) Experience with machine learning workflows, including feature engineering, training, evaluation, deployment and serving ## Description Join a dynamic team within Apple's Information Intelligence Infrastructure organization that designs, builds, and operates large-scale systems powering search and AI experiences for billions of users. We develop distributed, data-intensive infrastructure that processes web data at global scale, enabling extraction, enrichment, and knowledge graph construction across diverse content such as HTML, PDF, and other unstructured formats.","responsibilities":"Build and optimize large-scale extraction and enrichment pipelines that transform raw web data into structured knowledge powering Siri, Spotlight, Safari, and other Apple experiences. Design pipelines that leverage large language models, advanced NLP, and entity linking frameworks to identify, deduplicate, and contextualize information from the open web. Optimize extraction infrastructure for cost, throughput, and reliability including model serving and batch ML workloads. Develop ML classifiers for data quality, deduplication, and content filtering across billions of records. Improve extraction quality to produce high-fidelity training corpora for Apple's foundation models, directly improving reasoning and grounding capabilities. ## Related Videos - [Harnessing Apple Intelligence: Live Coding with Swift for iOS](https://www.wearedevelopers.com/videos/1515-harnessing-apple-intelligence-live-coding-with-swift-for-ios) - [The Resilience of the World Wide Web](https://www.wearedevelopers.com/videos/1281-the-resilience-of-the-world-wide-web) - [Lessons from Steve Jobs - Learnings from the Past for the Future](https://www.wearedevelopers.com/videos/1021-lessons-from-steve-jobs-learnings-from-the-past-for-the-future) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [NoLoJS - Avoiding JavaScript Cruft with HTML and CSS - Aaron T. Grogg](https://www.wearedevelopers.com/videos/1806-nolojs-avoiding-javascript-cruft-with-html-and-css-aaron-t-grogg) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it)