Machine Learning Engineer

Apple Inc.
Cupertino, United States
1 day ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
1 year minimum
Working hours
Regular working hours
Job source

Tech stack

Code Review Software Debugging Machine Learning Large Language Models Generative AI Information Technology

Job description

We are looking for talented machine learning engineers who are excited to tackle some of the most meaningful and technically challenging problems in building and deploying foundation model-based products for our customers.

As a Machine Learning Engineer focused on foundation model evaluation, you will play a critical role in assessing the capabilities of the models that power Apple Intelligence features.

You will work closely with machine learning researchers to translate evaluation insights into actionable improvements that advance future model performance., As a foundation model evaluation Machine Learning Engineer, you will be entrusted with ensuring that foundation model performance can be measured quickly and reliably, in order to support crucial model shipping decisions.

You will design, implement, and maintain crucial evaluation infrastructure.

You will collaborate extensively with ML researchers on both model hillclimbing and developing novel methodologies for measuring model performance.

Your responsibilities will span a number of high-impact parts of the Apple product and foundation model lifecycle.

Requirements

5+ years of hands on ML engineering experiences, with at least 1+ years working directly on large language models or generative AI.

Bachelor’s, Master’s, or PhD in Computer Science, Machine Learning, or a related technical field - or equivalent practical experience.

Strong software engineering fundamentals: debugging, testing, code reviews, and production reliability / scalability.

Hands-on experience with LLM training and / or evaluation workflows, including any of the following: pre-training, post-training, online evaluation, offline evaluation, automated evaluation, human evaluation.

Preferred Qualifications

Hands on experience with evaluating large language models at scale or designing large language model benchmarks.

Strong communication skills, able to clearly and concisely convey important information.

Self-motivated and curious. Strive to continually learn on the job.

High level of creative and critical thinking skills with an innate drive to improve how things work. Have a high tolerance for ambiguity and the ability to identify the most important problems to solve.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

5:01 min

Leveraging large language models for code optimization and development

Stephan Gillich Stephan Gillich +3 · World Congress 2024

3:39 min

Addressing code review surrender and process exploitation

Laura Tacho Laura Tacho · World Congress 2026 Europe

2:37 min

Tracing the evolution from early AI to generative AI

Mike Mike · World Congress 2025

2:36 min

Applying supervised machine learning for practical rule extraction

Katja Träumner

1:12 min

Training and fine-tuning models natively using MLX

MIlan Todorović MIlan Todorović · World Congress 2025

56 sec

The hidden costs of delayed peer code reviews

Tim Gilboy Tim Gilboy

Videos

See all

Related articles

See all