AIML - Machine Learning Researcher - Multimodal Agent

Apple Inc.
Seattle, WA, United States
3 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Java (Programming Language) Artificial Intelligence Apple Products Computer Vision C++ (Programming Language) Information Retrieval Python (Programming Language) Machine Learning Natural Language Processing Cloud Services Large Language Models Deep Learning
+5 more
Siri Data Analytics Search Engines Artificial Intelligence Markup Language (AIML) Golang

Job description

The AIML Multimodal Foundation Model Team is pioneering next-generation intelligent agent technologies that combine multimodal reasoning, tool-use, and visual understanding. Our innovative features redefine how hundreds of millions of people utilize their computers and mobile devices for search and information retrieval. Our universal search engine powers search capabilities across a range of Apple products, including Siri, Spotlight, Safari, Messages, and Lookup. Additionally, we develop cutting-edge generative AI technologies based on multimodal large language models to enable innovative features in both Apple’s devices and cloud-based services. As a member of this team, you will design new architectures for multimodal agents, explore advanced training paradigms, and build robust agentic capabilities such as planning, grounding, tool-use, and autonomous task execution. You will collaborate closely with researchers and engineers to bring cutting-edge agent research into production, transforming Apple devices into intelligent partners that help users get things done., As a member of our fast-paced group, you’ll have the unique and rewarding opportunity to shape upcoming products from Apple. We are looking for people with excellent applied machine learning, computer vision, multimodal LLM, and agent training experience and solid engineering skills.

This role will have the following responsibilities:

  • Developing state-of-the-art multimodal foundation models for Apple Intelligence.
  • Developing various agent capabilities for multimodal LLMs, including computer use agents, visual tool use, thinking with images, and multimodal web search.
  • Developing, fine-tuning, and evaluating domain specific foundation models for various tasks and applications in Apple’s AI powered products
  • Conducting applied research to transfer the pioneering research in generative AI to production ready technologies
  • Understanding product requirements, translate them into modeling tasks and engineering tasks

Requirements

  • PhD, MS or equivalent experience

  • Experience in machine learning, deep learning, computer vision, or natural language processing

  • Proficiency in one of following languages: Python, Go, Java, C+ Preferred Qualifications

  • Excellent data analytical skills

  • Good interpersonal skills and team player

  • PhD preferred

About the company

Apple

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on worksourcewa.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:16 min

Evaluating the enduring financial and technological legacy of Apple

Marco Landi · WWC 2024

40 sec

Navigating Apple's evolving on-device AI and machine learning stack

Precious Osaro Precious Osaro · WWC Europe 2026

1:25 min

Distinguishing artificial intelligence from deep learning

Sam Witteveen · Coffee With Developers

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

1:34 min

Introduction to the Apple Intelligence developer ecosystem

MIlan Todorović MIlan Todorović · WWC 2025

2:17 min

Distinguishing between AI, machine learning, and deep learning

Mary Grygleski Mary Grygleski · LIVE

Videos

See all

Related articles

See all