System Performance Engineer - AI/ML, Platform Architecture

Apple Inc.
Santa Clara, CA, United States
20 days ago
Apply on www.techcareers.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours

Tech stack

Microsoft Windows Artificial Intelligence Apple Products Apple Mac Systems Computing Platforms Profiling Nvidia CUDA Computer Engineering Software Debugging Linux Python (Programming Language) Tensorflow
+4 more
Graphics Processing Unit (GPU) Pytorch Information Technology TensorRT

Job description

As a System Performance Engineer - AI/ML on the Platform Architecture team, you will play a critical role in ensuring the performance excellence of Apple products by benchmarking AI experiences.

Requirements

  • Bachelor’s Degree in Computer Engineering, Electrical Engineering, Computer Science, or a related degree.
  • Experience in performance measurement, analysis, debug, and optimization using software profiling tools (e.g. VTune, Nsight, uProf, WPA, HWInfo).
  • Experience curating, running, and interpreting AI/ML workloads or benchmarks, and experience with the frameworks and hardware-specific runtimes used (e.g., PyTorch, TensorFlow, TensorRT/CUDA, OpenVINO)., * 3+ years of industry experience in performance analysis.
  • Understanding of the hardware that accelerates AI/ML - GPUs, NPUs, and the memory and bandwidth factors that drive performance - and why they cause performance differences.
  • Understanding of OS fundamentals and system-level performance, including performance, power, and thermal profiling tools.
  • Good communication and presentation skills, with the ability to distill complex analysis for technical and executive audiences.
  • Excellent interpersonal skills; ability to work cross-functionally from silicon to SW teams.
  • Passion and desire to learn new things, from deep technical topics to user workflows.
  • Proficiency in Python for scripting and automation.
  • Knowledge of macOS, Windows, and Linux operating systems.

About the company

Imagine what you can do here. At Apple, new ideas have a way of becoming extraordinary products very quickly. Bring passion and dedication to your job, and there’s no telling what we can accomplish together.

Do you love to spend your free time tinkering with new gadgets and exploring cutting-edge technologies? Do you obsess over every detail of product experience? Are you constantly looking out for the latest technology developments and tracking the shifting landscape?

We explore real user workflows and everyday use cases, measure them rigorously across hardware and software, and turn what we learn into insights that shape current and future products. As AI becomes central to the product experience, we’re looking for an engineer to work on our AI workload strategy end to end - from curating what we measure, to running it across devices, to presenting what it means.

Our team is collaborative, creative, and passionate about the value we add to future product designs. Join this team, and you’ll collaborate with engineers across Apple to deep dive into hardware and software technologies, uncover potential optimization opportunities to enhance and improve customer experiences across Apple devices. Come join us!

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.techcareers.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

40 sec

Navigating Apple's evolving on-device AI and machine learning stack

Precious Osaro Precious Osaro · World Congress 2026 Europe

4:52 min

Essential phases in building and refining language models

Anshul Jindal Anshul Jindal +1 · World Congress 2025

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

1:31 min

Profiling computing workloads across environments using Performance Studio

Andrew Wafaa Andrew Wafaa · World Congress 2024

2:32 min

Core libraries driving inference engines and multi-GPU networking

Adolf Hohl Adolf Hohl · World Congress 2024

Videos

See all

Related articles

See all