Visual Intelligence and Machine Learning Research Scientist

Apple Inc.
San Jose, CA, United States
1 day ago
Apply on www.themuse.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$150,400.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Apple Mac Systems Computing Platforms Systems Engineering Computer Vision Python (Programming Language) Machine Learning NumPy Object Detection OpenCV Tensorflow Sensor Fusion
+6 more
Signal Processing Pytorch Deep Learning Scikit Learn Information Technology Machine Learning Operations

Job description

We are looking for a Visual Intelligence and Machine Learning Scientist who bridges computational neuroscience and modern AI. You will build models that reason about human visual attention, behavioral states, comfort, and perceptual salience from video and multimodal sensor streams. You will have a track record of original research contributions, a drive to deploy algorithms on real-world data, and the versatility to take an idea from exploratory research to polished user experience., Apply large and on-device vision foundation models, multimodal transformers, and generative AI to build visual intelligence systems that predict saliency, understand scene context, and estimate perceptual state from video and sensor streams.

Design and own end-to-end AI pipelines: from data curation and model training to fine-tuning and adapting pretrained vision-language and multimodal foundation models for on-device, real-time inference.

Build multimodal AI systems that fuse visual signals with IMU, depth, and temporal context, improving robustness and perceptual fidelity across diverse real-world conditions.

Use generative AI and computational modeling to develop learned representations and priors that extend model generalization beyond labeled data, across devices, environments, and user populations.

Partner with app teams, UX designers, and systems engineers to translate algorithmic outputs directly into user-facing features on iOS, macOS, and spatial computing platforms.

Define and track perceptual and behavioral KPIs at the object, scene, and user level to rigorously validate algorithm quality and drive continuous improvement.

Stay at the frontier of visual intelligence, multimodal AI, and computational neuroscience, rapidly prototyping novel ideas and carrying the most promising from research to shipped product.

Requirements

PhD in Computer Science, Biomedical Engineering, Computational Neuroscience, Applied Mathematics, or a related field, with a focus on computer vision or machine learning.

Experience with vision foundation models, vision transformers (ViT), and adapter-based fine-tuning for perceptual or behavioral downstream tasks.

Familiarity with generative AI and neural rendering approaches (diffusion models, implicit neural representations) as components of computational perception pipelines., Background in computational modeling of perceptual or behavioral phenomena: modeling user state, comfort, visual salience, or behavioral response from video or sensor data.

Experience developing or deploying models for spatial computing platforms (AR/VR headsets, iOS, macOS).

Track record of shipping algorithms in resource-constrained, real-time on-device inference environments.

Ability to be versatile across a multi-faceted role, moving fluidly between deep research, engineering rigor, and direct collaboration with product and design teams to bring science to life in user-facing features., MS with 3+ years of research experience, in Computer Science, Biomedical Engineering, Computational Neuroscience, Applied Mathematics, or a related field, with a focus on computer vision or machine learning.

Demonstrated research contributions to visual intelligence or perceptual modeling, including peer-reviewed publications or equivalent industry impact.

Deep expertise in visual perception: saliency modeling, object detection and localization, optical flow, depth estimation, and multimodal learning including sensor fusion with IMU and temporal signals.

Strong hands-on experience training and deploying deep learning models using modern frameworks (PyTorch, TensorFlow, JAX) and Python scientific libraries (NumPy, OpenCV, scikit-learn).

Strong mathematical foundations in linear algebra, probability, optimization, and signal processing, with experience in large-scale dataset curation and evaluation methodology.

Excellent communication and collaboration skills; ability to thrive in a fast-paced environment alongside scientists, engineers, designers, and domain experts from the behavioral and cognitive sciences.

Benefits & conditions

At Apple, base pay is one part of our total compensation package and is determined within a range. This provides the opportunity to progress as you grow and develop within a role. The base pay range for this role is between $150,400 and $277,600, and your base pay will depend on your skills, qualifications, experience, and location.

Apple employees also have the opportunity to become an Apple shareholder through participation in Apple’s discretionary employee stock programs. Apple employees are eligible for discretionary restricted stock unit awards, and can purchase Apple stock at a discount if voluntarily participating in Apple’s Employee Stock Purchase Plan. You’ll also receive benefits including: Comprehensive medical and dental coverage, retirement benefits, a range of discounted products and free services, and for formal education related to advancing your career at Apple, reimbursement for certain educational expenses - including tuition. Additionally, this role might be eligible for discretionary bonuses or commission payments as well as relocation. Learn more about Apple Benefits

About the company

Apple is where individual imaginations gather together, committing to the values that lead to great work. Every new product we build, every service we deliver, is the result of us making each other’s ideas stronger. Here, you’ll do more than join something - you’ll add something.

Our team sits at the intersection of computational neuroscience, visual intelligence, and AI, translating cutting-edge research directly into experiences that users see and feel every day. We prototype algorithms from first principles all the way through to shipping features, working hand-in-hand with app teams and designers so that the science we build has immediate, tangible impact on how people interact with Apple products. If you want your research to matter beyond a paper, this is where that happens.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.themuse.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Introduction to the Apple Intelligence developer ecosystem

MIlan Todorović MIlan Todorović · World Congress 2025

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

4:53 min

Achieving real-time tracking performance with OpenCV and segmentation

Thomas Endres Thomas Endres +2 · World Congress 2021

2:34 min

Maximizing execution memory effectively via python numpy broadcasting

Jodie Burchell · LIVE

40 sec

Navigating Apple's evolving on-device AI and machine learning stack

Precious Osaro Precious Osaro · World Congress 2026 Europe

1:06 min

Compiling PyTorch environments for advanced time forecasting

Christoph Lohrmann Christoph Lohrmann +1 · World Congress 2026 Europe

Videos

See all

Related articles

See all