Data Center - MLB Reliability Engineer

Apple Inc.
Austin, TX, United States
6 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

Data Analysis Systems Engineering Data Centers Failure Mode Effects Analysis Python (Programming Language) MATLAB Minitab Reliability Engineering Scripting JMP (Statistical Software) Markov

Job description

At Apple, we don’t just follow industry standards-we define them. We value creative problem-solving and the ability to adapt to new technical challenges. In this role, you will collaborate with diverse hardware teams to ensure our data center infrastructure is not only durable, but built to exceed expectations. From early concepts to field optimization, you will drive innovative reliability strategies and continuous improvement to shape the future of the critical systems powering our global services.

Requirements

The position is on Apple’s innovative Datacenter MLB Reliability Engineering team, wherein we are seeking a highly analytical individual to develop the reliability strategy for our next-generation Data Center motherboards. In this role, you will bridge the gap between component-level and package level physics of failures and system-level availability. You will drive SoC and board integration reliability through rigorous stress testing and physics-of-failure analysis, utilizing advanced statistical modeling to ensure these critical modules meet the uptime and availability requirements of a high-demand data center environment, * BS in Materials Science, Electrical, Mechanical Engineering or an equivalent field desired with 5+ years of experience.

  • Proficiency in statistical life data analysis
  • Strong knowledge of Semiconductor package integration, SLI reliability, passive components, PCB reliability, bonding materials, and warpage control.
  • Ability to apply FMEA (Failure Modes and Effects Analysis) methodologies.
  • Excellent written and verbal communication skills with the ability to explain complex statistical concepts to non-experts.
  • Ability to manage multiple projects simultaneously in a fast-paced environment., * MS or PhD in Reliability Engineering, Systems Engineering, Electrical Engineering, Materials Science, or an equivalent field..
  • Background in reliability for large-die packages, heterogeneous integration, and high-power server motherboards.
  • Experience applying reliability modeling (Markov, RBD, Bayesian) and RAS metrics to Data Center architectures.
  • Proficiency with reliability software (e.g., Weibull++, BlockSim, JMP, Minitab) and scripting languages for statistical modeling like Python, R, or MATLAB.
  • A record of initiating innovation and continuous improvement in reliability methodologies.
  • Dynamic and “can-do” attitude with a desire to work with a great team and product.
  • Exceptional problem-solving abilities and strong attention to detail.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.austinjobsite.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

7:08 min

Engineering practices for extreme platform reliability

Justin Kitagawa · Coffee With Developers

2:21 min

Applying diffusion models for image upscaling and refinement

Han Xiao · WWC 2022

2:40 min

Motivations for transitioning legacy MATLAB repositories to Python

Michael Niebisch Michael Niebisch · WWC 2024

40 sec

Hardware durability labs and robot testing methods

Chris Heilmann +1 · LIVE

1:32 min

Using rule-based heuristics to simulate realistic typing errors

Artur Naumenko Artur Naumenko · WWC Europe 2026

2:14 min

Custom domain controllers and vehicle communication networks

Denis Grahovac · WWC 2021

Videos

See all

Related articles

See all