> Markdown version of [/jobs/ext/281627-machine-learning-research-engineer-siml-ise](https://www.wearedevelopers.com/jobs/ext/281627-machine-learning-research-engineer-siml-ise). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Machine Learning Research Engineer, SIML - ISE - **Company:** Apple Inc. - **Location:** Cupertino, CA, United States - **Salary:** $147,400.0 - $272,100.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Artificial Neural Networks, Computer Vision, Python (Programming Language), Machine Learning, Language Modeling, Pytorch, Large Language Models, Deep Learning, Information Technology, Multiaccess Edge Computing - **Published:** May 29, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=ed7eb340ef4dd43e ## About the Role Do you have experience in Neural networks?, Do you have a Master's degree?, This role requires experience in vision-language models, and ability to fine-tune/adapt/distill multi-modal LLMs. You will be part of a fast-paced, impact-driven Applied Research organization working on cutting-edge machine learning that is at the heart of the most loved features on Apple platforms, including Apple Intelligence, Camera, Photos, Visual Intelligence, and more!, Proven track record of research contributions demonstrated through publications in top-tier conferences and journals. Background in multi-modal reasoning, VLM, and MLLM research with impactful software projects. Solid understanding of natural language processing (NLP) and computer vision fundamentals. Minimum Qualifications Master's or Ph.D. in Computer Science, Artificial Intelligence, Machine Learning, or related field - or relevant industry experience Proficiency in Python and deep learning frameworks such as PyTorch, or equivalent Practical experience with training and evaluating neural networks Familiarity with multimodal learning, vision-language models, or large language models Strong problem-solving skills and ability to work in a collaborative, product-focused environment Ability to communicate technical results clearly and concisely ## Description As a Machine Learning Research Engineer, you will help design and develop models and algorithms for multimodal perception and reasoning leveraging Vision-Language Models (VLMs) and Multimodal Large Language Models (MLLMs). You will collaborate with experienced researchers and engineers to explore new techniques, evaluate performance, and translate product needs into impactful ML solutions. Your work will contribute directly to user-facing features across billions of devices. Your primary responsibilities include: * Contribute to the development and adaptation of AI/ML models for multimodal perception and reasoning * Innovate robust algorithms that integrate visual and language data for comprehensive understanding * Collaborate closely with cross-functional teams to translate product requirements into effective ML solutions. * Conduct hands-on experimentation, model training, and performance analysis * Communicate research outcomes effectively to technical and non-technical stakeholders, providing actionable insights. ## Related Videos - [Getting Started with Machine Learning](https://www.wearedevelopers.com/videos/260-getting-started-with-machine-learning) - [Harnessing Apple Intelligence: Live Coding with Swift for iOS](https://www.wearedevelopers.com/videos/1515-harnessing-apple-intelligence-live-coding-with-swift-for-ios) - [Focoos AI: Building the Future of Computer Vision](https://www.wearedevelopers.com/videos/1659-focoos-ai-building-the-future-of-computer-vision) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Geometric deep learning for drug discovery](https://www.wearedevelopers.com/videos/264-geometric-deep-learning-for-drug-discovery) - [Computer Vision from the Edge to the Cloud done easy](https://www.wearedevelopers.com/videos/263-computer-vision-from-the-edge-to-the-cloud-done-easy) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [Stephan Gillich - Bringing AI Everywhere](https://www.wearedevelopers.com/magazine/489-stephan-gillich-bringing-ai-everywhere) - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this)