SW ML Optimization Engineer

Apple Inc.

Cupertino, Northern (CA, KY)

Hybrid

USD 129,000 - 225,000

Full time

17 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Apple Inc. in Cupertino seeks an engineer for the Platform Architecture group to analyze AI/ML workloads on Apple Silicon and identify performance bottlenecks across GPU, ANE, and CPU.

You will build microbenchmarks, develop optimized software implementations, and contribute to forward-looking workloads while collaborating with silicon, software, and tools teams to advance performance and efficiency.

Qualifications

  • Bachelor's degree in Computer Science, Computer Engineering, Mathematics, Electrical Engineering, or related field (or equivalent practical experience).
  • Experience with GPU or parallel programming (Metal, OpenCL, CUDA or similar).
  • Experience with profiling/performance analysis tools (Xcode Instruments, VTune, Nsight Compute or equivalent).
  • Development experience in Python, C or C++.

Responsibilities

  • Profile and analyze AI/ML workloads across Apple Silicon compute engines (GPU, ANE, and CPU) to help identify performance bottlenecks, working alongside senior engineers.
  • Build targeted microbenchmarks to characterize strengths, weaknesses, and usage patterns of different devices and workloads.
  • Develop and tune optimized implementations of software workloads on Apple Silicon (GPU and CPU), applying the latest instruction sets and frameworks.
  • Contribute to the development of forward-looking AI/ML workloads that represent where the industry is headed.
  • Partner with system and tools teams to prototype performance experiments and help build performance analysis instruments, libraries, and frameworks.
  • Grow toward a system-wide understanding of the stack—from high-level APIs down to the hardware.

Skills

GPU programming
Python
C/C++
Metal
OpenCL/CUDA
Xcode Instruments
Profiling tools

Education

Bachelor's degree in CS/CE/Math/EE

Tools

VTune
Nsight Compute

Job description

Cupertino, California, United States Hardware

At Apple, our Platform Architecture group is responsible for connecting our hardware and software into one unified system. You’ll collaborate with engineers across Apple to design how all of our technologies work in unison, drive development of our renowned system-on-a-chip architecture and develop forward-looking prototype systems and software.

Description

Our team is driving performance enhancements in application and system software and developing novel algorithms to deliver integrated, highly optimized solutions based on Apple Silicon. In this role, you will analyze existing and new workloads to identify performance bottlenecks in the hardware and/or software. Working with your colleagues, you will address performance limitations and provide recommendations for Apple hardware and software improvements. In addition to working directly with developers, you will identify patterns of performance challenges on Apple silicon, emerging new usage models, and provide feedback to the silicon and software teams for potential improvements.

Responsibilities
  • Profile and analyze AI/ML workloads across Apple Silicon compute engines (GPU, ANE, and CPU) to help identify performance bottlenecks, working alongside senior engineers.
  • Build targeted microbenchmarks to characterize the strengths, weaknesses, and usage patterns of different devices and workloads.
  • Develop and tune optimized implementations of software workloads on Apple Silicon (GPU and CPU), learning to apply the latest instruction sets and frameworks.
  • Contribute to the development of forward-looking AI/ML workloads that represent where the industry is headed.
  • Partner with system and tools teams to prototype performance experiments and help build performance analysis instruments, libraries, and frameworks.
  • Grow toward a system-wide understanding of the stack—from high-level APIs down to the hardware.
Minimum Qualifications
  • Bachelor's degree in Computer Science, Computer Engineering, Mathematics, Electrical Engineering, or a related quantitative field (or equivalent practical experience).
  • Experience with GPU or parallel programming—e.g., Metal, OpenCL, CUDA, or similar—through coursework, personal projects, internships, or research.
  • Experience with profiling/performance analysis tools (e.g., Xcode Instruments, VTune, Nsight Compute, or equivalent) and basic performance analysis concepts.
  • Development experience in Python, C or C++.
Preferred Qualifications
  • Solid foundation in mathematics, algorithms, and/or computer architecture fundamentals.
  • Experience writing or tuning compute kernels (e.g., GEMM, attention, or other numerically intensive routines).
  • Exposure to ML frameworks such as PyTorch, and to AI/ML, graphics, or HPC workloads and benchmarks.
  • Coursework or projects involving parallel computing, numerical methods, signal processing, or performance optimization.
  • Interest in (or exposure to) the deeper stack - drivers, firmware, compilers, or low-level libraries.
  • Interest in Apple Silicon and its frameworks (Metal, MLX, Core ML).
  • Curiosity about hardware/software co-design and a demonstrated drive to learn independently.
  • Strong communication skills and the ability to collaborate effectively across teams.

At Apple, base pay is one part of our total compensation package and is determined within a range. This provides the opportunity to progress as you grow and develop within a role. The base pay range for this role is between $129,300 and $225,300, and your base pay will depend on your skills, qualifications, experience, and location.

Apple employees also have the opportunity to become an Apple shareholder through participation in Apple’s discretionary employee stock programs. Apple employees are eligible for discretionary restricted stock unit awards, and can purchase Apple stock at a discount if voluntarily participating in Apple’s Employee Stock Purchase Plan. You’ll also receive benefits including: Comprehensive medical and dental coverage, retirement benefits, a range of discounted products and free services, and for formal education related to advancing your career at Apple, reimbursement for certain educational expenses — including tuition. Additionally, this role might be eligible for discretionary bonuses or commission payments as well as relocation. Learn more about Apple Benefits

Note: Apple benefit, compensation and employee stock programs are subject to eligibility requirements and other terms of the applicable plan or program.

Apple is an equal opportunity employer that is committed to inclusion and diversity. We seek to promote equal opportunity for all applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, Veteran status, or other legally protected characteristics. Learn more about your EEO rights as an applicant

At Apple, we believe accessibility is a fundamental human right. You’ll find that idea reflected in everything here — in our culture, our benefits and our digital tools. By welcoming as many perspectives as possible, we help you build a career where you feel like you belong.

Learn about accessibility in Apple’s workplace

Learn about reasonable accommodations for job applicants

Apple accepts applications to this posting on an ongoing basis.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

SW Optimization Engineer AI/ML
SW Optimization Engineer AI/ML

Apple Inc. • Cupertino (CA), Northern (KY)

Hybrid
USD 185,000 - 325,000
Machine Learning Engineer, Platform Architecture
Machine Learning Engineer, Platform Architecture

Apple Inc. • Cupertino (CA)

On-site
USD 150,000 - 278,000
Employee stock programs
Relocation assistance
Education reimbursement
System Performance Engineer - AI/ML, Platform Architecture
System Performance Engineer - AI/ML, Platform Architecture

Apple Inc. • Santa Clara (CA), Northern (KY)

Hybrid
USD 150,000 - 225,000
Medical and dental coverage
Retirement benefits
Discounted products and services
+1
Applied Machine Learning Engineer, Platform Architecture
Applied Machine Learning Engineer, Platform Architecture

Apple Inc. • Cupertino (CA), Northern (KY)

On-site
USD 185,000 - 278,000
ML Software Engineer
ML Software Engineer

Apple Inc. • Seattle (WA)

On-site
USD 175,000 - 263,300
Employee stock purchase plan
discretionary stock awards
Medical and dental coverage
+3
GPU ML Engineer
GPU ML Engineer

Apple Inc. • Cupertino (CA)

On-site
USD 150,400 - 277,600
Medical & dental
Retirement benefits
Employee stock programs
+2
Software Engineer - Model Performance - Special Projects
Software Engineer - Model Performance - Special Projects

Apple Inc. • Cupertino (CA)

On-site
USD 184,000 - 325,000
Employee stock programs
Education reimbursement
Medical and dental coverage
SW ML Optimization Engineer
SW ML Optimization Engineer

Socket.dev • Cupertino (CA)

On-site
USD 180,000 - 240,000
Systems Performance Architect
Systems Performance Architect

Apple Inc. • Cupertino (CA)

On-site
USD 150,000 - 278,000
Stock programs
RSUs
Employee Stock Purchase Plan
+3
Pre-silicon Compute Framework Manager
Pre-silicon Compute Framework Manager

Apple Inc. • Cupertino (CA), Northern (KY)

On-site
USD 206,000 - 356,000
Medical coverage
Retirement benefits
Employee stock plan