Data Center - MLB Reliability Engineer

Apple Inc.

Austin (TX)

On-site

USD 140,000 - 210,000

Full time

6 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Apple Inc. is seeking an analytical Reliability Engineer to shape the reliability strategy for next-generation data center motherboards. You will bridge component to system-level availability, drive physics-of-failure analyses, and implement rigorous testing plans to ensure uptime for high-demand services.

The role requires a BS in engineering with 5+ years of experience, plus proficiency in statistical life data analysis and reliability modeling. Collaboration across hardware teams is essential.

Qualifications

  • BS in Materials Science, Electrical, Mechanical Engineering or equivalent field.
  • 5+ years of reliability engineering experience in data center or high-power electronics.
  • Proficiency in statistical life data analysis and FMEA methodologies.
  • Experience with semiconductor package integration, SLI, PCB reliability, bonding materials, and warpage control.
  • Strong communication skills to explain statistics to non-experts.

Responsibilities

  • Lead board-level reliability strategy for HPC motherboards and server systems.
  • Apply FMEA to identify and mitigate failure modes early in design.
  • Analyze semiconductor packages, SLIs, and PCB materials using physics-of-failure.
  • Develop reliability models (RBDs, Markov Chains, Bayesian) for system availability.
  • Translate system availability requirements into component specs and targets.
  • Design and execute mechanical, environmental, and ALT test plans.
  • Perform RCA and statistical analysis to improve design maturity.
  • Collaborate with cross-functional teams to meet data center availability targets.

Skills

Statistical life data analysis
FMEA methodologies
Group cross-functional collaboration
Communication of complex statistics
Problem solving

Education

BS in Materials Science, Electrical, Mechanical Engineering or equivalent
MS or PhD in Reliability/Systems/Electrical/Materials (preferred)

Tools

Weibull++
BlockSim
JMP
Minitab
Python
R
MATLAB

Job description

At Apple, we don't just follow industry standards-we define them. We value creative problem-solving and the ability to adapt to new technical challenges. In this role, you will collaborate with diverse hardware teams to ensure our data center infrastructure is not only durable, but built to exceed expectations. From early concepts to field optimization, you will drive innovative reliability strategies and continuous improvement to shape the future of the critical systems powering our global services.

Description

The position is on Apple's innovative Datacenter MLB Reliability Engineering team, wherein we are seeking a highly analytical individual to develop the reliability strategy for our next-generation Data Center motherboards. In this role, you will bridge the gap between component-level and package level physics of failures and system-level availability. You will drive SoC and board integration reliability through rigorous stress testing and physics-of-failure analysis, utilizing advanced statistical modeling to ensure these critical modules meet the uptime and availability requirements of a high-demand data center environment

Responsibilities
  • Lead the board-level reliability strategy for high-performance computing (HPC) motherboards and server systems.
  • Utilize Design and Process FMEA to identify, categorize, and eliminate failure modes early in the design cycle.
  • Apply physics-of-failure principles to analyze semiconductor packages, Second Level Interconnects (SLI), and PCB materials.
  • Develop advanced reliability models such as Reliability Block Diagrams (RBDs), Markov Chains, and Bayesian analysis to quantify board-level risk and its contribution to overall system availability.
  • Translate Reliability, Availability, and Serviceability (RAS) metrics for large-scale data center deployments into actionable board-level design targets and validation criteria.
  • Design and execute comprehensive test plans including Mechanical Stress, Environmental Testing, and Accelerated Life Testing (ALT).
  • Drive rigorous Root Cause Analysis (RCA) and perform statistical analysis to assess design maturity.
  • Work with cross-functional teams to translate system-level availability requirements into component specifications.
Minimum Qualifications
  • BS in Materials Science, Electrical, Mechanical Engineering or an equivalent field desired with 5+ years of experience.
  • Proficiency in statistical life data analysis
  • Strong knowledge of Semiconductor package integration, SLI reliability, passive components, PCB reliability, bonding materials, and warpage control.
  • Ability to apply FMEA (Failure Modes and Effects Analysis) methodologies.
  • Excellent written and verbal communication skills with the ability to explain complex statistical concepts to non-experts.
  • Ability to manage multiple projects simultaneously in a fast-paced environment.
Preferred Qualifications
  • MS or PhD in Reliability Engineering, Systems Engineering, Electrical Engineering, Materials Science, or an equivalent field..
  • Background in reliability for large-die packages, heterogeneous integration, and high-power server motherboards.
  • Experience applying reliability modeling (Markov, RBD, Bayesian) and RAS metrics to Data Center architectures.
  • Proficiency with reliability software (e.g., Weibull++, BlockSim, JMP, Minitab) and scripting languages for statistical modeling like Python, R, or MATLAB.
  • A record of initiating innovation and continuous improvement in reliability methodologies.
  • Dynamic and "can-do" attitude with a desire to work with a great team and product.
  • Exceptional problem-solving abilities and strong attention to detail.

Apple is an equal opportunity employer that is committed to inclusion and diversity. We seek to promote equal opportunity for all applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, Veteran status, or other legally protected characteristics. Learn more about your EEO rights as an applicant

At Apple, we believe accessibility is a fundamental human right. You'll find that idea reflected in everything here - in our culture, our benefits and our digital tools. By welcoming as many perspectives as possible, we help you build a career where you feel like you belong.

Learn about accessibility in Apple's workplace

Learn about reasonable accommodations for job applicants

Apple accepts applications to this posting on an ongoing basis.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Center - MLB Reliability Engineer
Data Center - MLB Reliability Engineer

Socket.dev • Austin (TX)

On-site
USD 140,000 - 180,000
Data Center Reliability Engineer - HPC Board & SoC
Data Center Reliability Engineer - HPC Board & SoC

Apple Inc. • Austin (TX)

On-site
USD 140,000 - 210,000
Hardware Reliability Engineer - Apple Vision Products
Hardware Reliability Engineer - Apple Vision Products

Apple Inc. • Cupertino (CA)

On-site
USD 147,000 - 273,000
Comprehensive medical and dental coverage
Employee Stock Purchase Plan
Tuition reimbursement
+1
Hardware Reliability Engineer - Mac System Reliability
Hardware Reliability Engineer - Mac System Reliability

Socket.dev • Cupertino (CA)

On-site
USD 110,000 - 170,000
SoC Silicon Reliability Engineer
SoC Silicon Reliability Engineer

Apple Inc. • Cupertino (CA)

On-site
USD 184,700 - 324,800
Medical and dental coverage
Retirement benefits
Employee stock purchase plan
+2
Data Center Hardware Engineering - Hardware System Integration Engineer
Data Center Hardware Engineering - Hardware System Integration Engineer

Apple Inc. • Austin (TX)

On-site
USD 140,000 - 210,000
Hardware Reliability Engineer - iPad System Reliability
Hardware Reliability Engineer - iPad System Reliability

Socket.dev • Cupertino (CA)

On-site
USD 150,400 - 277,600
Medical and dental coverage
Employee stock programs
Education reimbursement
+1
Hardware Reliability Engineer - iPad System Reliability
Hardware Reliability Engineer - iPad System Reliability

Apple Inc. • Cupertino (CA)

On-site
USD 150,000 - 278,000
Medical + dental coverage
Employee stock programs
Relocation assistance
+1
Hardware Reliability Engineer - Watch System Reliability
Hardware Reliability Engineer - Watch System Reliability

Apple Inc. • San Diego (CA)

On-site
USD 142,000 - 263,000
Data Center Reliability Engineer: SoC & Board Analysis
Data Center Reliability Engineer: SoC & Board Analysis

Socket.dev • Austin (TX)

On-site
USD 140,000 - 180,000