Machine Learning Engineer, Proactive

Apple Inc.

Cupertino, Northern (CA, KY)

Hybrid

USD 150,000 - 225,000

Full time

25 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Medical & dental coverage
Retirement benefits
Stock purchase program
Tuition reimbursement

Job summary

Apple Inc. in Cupertino, CA, is seeking a Machine Learning Engineer to build the next generation of intelligent search and AI experiences.

You will design, train, fine-tune, and deploy large language models and semantic retrieval systems for on-device use. Collaborate with engineers, researchers, product managers, and designers to translate research into production, balancing model quality, latency, memory, and power constraints while preserving user privacy.

Qualifications

  • Bachelor degree in Computer Science, Machine Learning, Artificial Intelligence, or a related field.
  • Background in machine learning, deep learning, natural language processing, information retrieval, search, recommender systems, or generative AI.
  • Experience training, fine-tuning, or deploying transformer-based models and large language models.
  • Experience with modern deep learning architectures and techniques, including transformers, embeddings, representation learning, and neural ranking.
  • Programming skills in Python and/or C/C++, with experience building production-quality software using PyTorch, JAX, or TensorFlow.
  • Ability to work onsite in Cupertino, California, in accordance with Apple's applicable work policies.

Responsibilities

  • Build semantic retrieval, embedding, reranking, and retrieval-augmented generation systems, along with models for query understanding, intent prediction, personalization, retrieval, and ranking.
  • Analyze search relevance and user behavior to design evaluation methodologies, offline benchmarks, and online metrics that measure retrieval quality, ranking, personalization, and language model performance.
  • Build scalable experimentation and evaluation pipelines for LLMs and search models, including model quality, robustness, latency, efficiency, and end-to-end product metrics.
  • Design, train, fine-tune, distill, and optimize transformer-based language models and foundation models for efficient on-device deployment.
  • Develop LLM fine-tuning and post-training approaches, including supervised fine-tuning, instruction tuning, preference optimization, parameter-efficient fine-tuning, and task-specific adaptation.
  • Research and prototype approaches for on-device generative AI, including knowledge distillation, model compression, quantization, pruning, and low-latency inference.
  • Develop techniques to transfer capabilities from large foundation models into compact on-device models while balancing model quality, latency, memory footprint, power consumption, and compute constraints.
  • Partner with engineers, researchers, product managers, and designers to bring AI capabilities from research into production, driving technical strategy across projects and exploring new applications of foundation models, multimodal AI, agentic retrieval, and personalized intelligence.

Skills

Machine learning
Natural language processing
Information retrieval
Recommender systems
Generative AI
Transformer models
Embeddings
Neural ranking
Python
On-device ML

Education

Bachelor's degree in CS/ML/AI
Master's or PhD in CS/ML/AI

Tools

PyTorch
JAX
TensorFlow
On-device inference frameworks

Job description

Cupertino, California, United States Machine Learning and AI

At Apple, machine learning powers experiences that anticipate what people need before they ask. We're looking for a Machine Learning Engineer to help build the next generation of intelligent search and AI experiences technology that understands user intent, context, and personal information while preserving privacy. In this role, you'll design, train, fine-tune, optimize, and deploy large language models, semantic retrieval systems, and ranking models that power relevant, personalized, and context-aware experiences across Apple's ecosystem.

Description

You’ll design, train, fine-tune, and optimize transformer-based language models and foundation models for efficient on-device deployment, and build semantic retrieval, embedding, reranking, and retrieval-augmented generation systems that improve search quality and AI-powered experiences. You'll develop models for query understanding, intent prediction, personalization, retrieval, and ranking, while researching new approaches to LLM fine-tuning, knowledge distillation, model compression, quantization, and low-latency inference. You'll explore techniques for adapting large foundation models into smaller, highly capable models that can operate efficiently under on-device memory, compute, power, and latency constraints.You'll partner with engineers, researchers, product managers, and designers to bring new AI capabilities from research into production, driving technical strategy and leading projects from early exploration through large-scale deployment. This is an opportunity to explore new applications of foundation models, multimodal AI, agentic retrieval, and personalized intelligence, shaping the next generation of proactive and intelligent user experiences.

Responsibilities
  • Build semantic retrieval, embedding, reranking, and retrieval-augmented generation systems, along with models for query understanding, intent prediction, personalization, retrieval, and ranking.
  • Analyze search relevance and user behavior to design evaluation methodologies, offline benchmarks, and online metrics that measure retrieval quality, ranking, personalization, and language model performance.
  • Build scalable experimentation and evaluation pipelines for LLMs and search models, including model quality, robustness, latency, efficiency, and end-to-end product metrics.
  • Design, train, fine-tune, distill, and optimize transformer-based language models and foundation models for efficient on-device deployment.
  • Develop LLM fine-tuning and post-training approaches, including supervised fine-tuning, instruction tuning, preference optimization, parameter-efficient fine-tuning, and task-specific adaptation.
  • Research and prototype approaches for on-device generative AI, including knowledge distillation, model compression, quantization, pruning, and low-latency inference.
  • Develop techniques to transfer capabilities from large foundation models into compact on-device models while balancing model quality, latency, memory footprint, power consumption, and compute constraints.
  • Partner with engineers, researchers, product managers, and designers to bring AI capabilities from research into production, driving technical strategy across projects and exploring new applications of foundation models, multimodal AI, agentic retrieval, and personalized intelligence.
Minimum Qualifications
  • Bachelor degree in Computer Science, Machine Learning,
  • Artificial Intelligence, or a related field.
  • Background in machine learning, deep learning, natural
  • language processing, information retrieval, search,
  • recommender systems, or generative AI.
  • Experience training, fine-tuning, or deploying transformer-
  • based models and large language models.
  • Experience with modern deep learning architectures and
  • techniques, including transformers, embeddings,
  • representation learning, and neural ranking.
  • Programming skills in Python and/or C/C++, with experience
  • building production-quality software using modern machine
  • learning frameworks such as PyTorch, JAX, or TensorFlow.
  • Ability to work onsite in Cupertino, California, in accordance
  • with Apple's applicable work policies.
Preferred Qualifications
  • Master's or Ph.D. in Computer Science, Machine Learning, Artificial Intelligence, or a related field.
  • Experience optimizing machine learning models for resource-constrained environments, including knowledge distillation, model compression, quantization, and pruning.
  • Experience with on-device machine learning or edge AI, or mobile inference frameworks, including optimizing models for latency, memory, compute, and power constraints.
  • Experience distilling capabilities from large foundation models into small language models or task-specific models for efficient inference.
  • Experience building retrieval-augmented generation, vector search, embedding retrieval, neural reranking, or semantic search systems.
  • Experience with query understanding, query rewriting, intent classification, personalized retrieval, learning-to-rank, or recommendation models.
  • Experience evaluating language models, designing AI quality metrics, and building automated and human-in-the-loop evaluation pipelines.
  • Experience building large-scale production search, recommendation, personalization, or generative AI systems.
  • Familiarity with multimodal foundation models, tool use, agentic AI, or agentic retrieval systems.
  • Strong understanding of the tradeoffs among model quality, latency, memory, power consumption, privacy, and reliability for production on-device AI systems.
  • Ability to prototype new ideas, conduct rigorous experiments, solve ambiguous technical problems, and translate research advances into production-quality machine learning solutions.
At Apple, base pay is one part of our total compensation package and is determined within a range. This provides the opportunity to progress as you grow and develop within a role. The base pay range for this role is between $150,400 and $225,300, and your base pay will depend on your skills, qualifications, experience, and location. Apple employees also have the opportunity to become an Apple shareholder through participation in Apple’s discretionary employee stock programs. Apple employees are eligible for discretionary restricted stock unit awards, and can purchase Apple stock at a discount if voluntarily participating in Apple’s Employee Stock Purchase Plan. You’ll also receive benefits including: Comprehensive medical and dental coverage, retirement benefits, a range of discounted products and free services, and for formal education related to advancing your career at Apple, reimbursement for certain educational expenses — including tuition. Additionally, this role might be eligible for discretionary bonuses or commission payments as well as relocation. Learn more about Apple Benefits Note: Apple benefit, compensation and employee stock programs are subject to eligibility requirements and other terms of the applicable plan or program. Apple is an equal opportunity employer that is committed to inclusion and diversity. We seek to promote equal opportunity for all applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, Veteran status, or other legally protected characteristics. Learn more about your EEO rights as an applicant At Apple, we believe accessibility is a fundamental human right. You’ll find that idea reflected in everything here — in our culture, our benefits and our digital tools. By welcoming as many perspectives as possible, we help you build a career where you feel like you belong. Learn about accessibility in Apple’s workplace Learn about reasonable accommodations for job applicants Apple accepts applications to this posting on an ongoing basis.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Machine Learning Engineer, Proactive
Senior Machine Learning Engineer, Proactive

Apple Inc. • Cupertino (CA), Northern (KY)

Hybrid
USD 185,000 - 325,000
Medical and dental coverage
Employee stock purchase plan
Restricted stock unit awards
+1
Senior Machine Learning Engineer, Proactive
Senior Machine Learning Engineer, Proactive

Apple Inc. • Santa Clara (CA)

On-site
USD 185,000 - 325,000
Apple Benefits
Employee Stock Purchase Plan
Relocation assistance
+1
ASE - Machine Learning Engineer, MLPT
ASE - Machine Learning Engineer, MLPT

Apple Inc. • Santa Clara (CA), Northern (KY)

Hybrid
USD 185,000 - 325,000
Medical & dental coverage
Employee stock programs
Tuition reimbursement
Machine Learning Engineer, Natural Language Understanding, Proactive
Machine Learning Engineer, Natural Language Understanding, Proactive

Apple Inc. • Santa Clara (CA)

On-site
USD 150,000 - 278,000
Machine Learning/ Search Engineer - Services Special Projects
Machine Learning/ Search Engineer - Services Special Projects

Apple Inc. • Cupertino (CA), Northern (KY)

On-site
USD 185,000 - 325,000
Senior AI Engineer - Services Special Projects
Senior AI Engineer - Services Special Projects

Apple Inc. • San Francisco (CA)

On-site
USD 185,000 - 325,000
Medical and dental coverage
Retirement benefits
Employee stock purchase plan
+4
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Apple Inc. • Cupertino (CA)

On-site
USD 185,000 - 325,000
Machine Learning Engineer - Notifications & Personalization
Machine Learning Engineer - Notifications & Personalization

Apple Inc. • Cupertino (CA), Northern (KY)

Hybrid
USD 185,000 - 325,000
Apple benefits
Stock programs
Relocation assistance
AIML - Machine Learning Researcher, Data and ML Innovation
AIML - Machine Learning Researcher, Data and ML Innovation

Apple Inc. • Seattle (WA)

On-site
USD 150,000 - 278,000
AIML - Machine Learning Researcher, Foundation Models
AIML - Machine Learning Researcher, Foundation Models

Apple Inc. • Cupertino (CA)

On-site
USD 150,000 - 278,000
Medical coverage
Dental coverage
Stock programs
+1