Staff Machine Learning Engineer – AI/ML Compiler

Qualcomm

San Diego (CA)

On-site

USD 160,500 - 240,700

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Annual discretionary bonus program
Opportunity for RSU grants
Comprehensive benefits package

Job summary

A leading technology company is seeking an experienced Machine Learning Engineer to own the infrastructure for model compilations within the Qualcomm AI Hub. You will design and maintain the end-to-end compilation pipeline, ensuring efficient model execution across various backends. Candidates should have a relevant degree and substantial experience in compiler engineering, ML infrastructure, and proficiency in Python and C++. A competitive salary range of $160,500 to $240,700 is offered, alongside a comprehensive benefits package.

Qualifications

  • Bachelor's, Master's, or PhD in relevant field with appropriate work experience.
  • 3+ years of industry experience in ML infrastructure or compiler engineering.
  • Proficiency in Python and C++.

Responsibilities

  • Design, develop, and maintain the compilation pipeline for Qualcomm AI Hub Workbench.
  • Build automated compilation pipelines and CI/CD evaluation harnesses.
  • Collaborate with internal teams for model onboarding.

Skills

Compiler engineering
Python
C++
Machine Learning infrastructure
Graph IRs
Automation
Technical documentation

Education

Bachelor's degree in Computer Science, Engineering, Information Systems, or related field
Master's degree in Computer Science, Engineering, Information Systems or related field
PhD in Computer Science, Engineering, Information Systems, or related field

Tools

PyTorch
ONNX
CI/CD pipelines

Job description

Overview

Company
Qualcomm Technologies, Inc.

Job Area

Job Area
Engineering Group, Engineering Group > Machine Learning Engineering

About the Role

Qualcomm AI Hub is the platform for on-device AI — enabling developers to easily integrate, optimize, and deploy ML models on Qualcomm devices. Qualcomm AI Hub Workbench lets developers compile trained PyTorch or ONNX models into deployable artifacts targeting a variety of runtimes — LiteRT, ONNXRuntime, or Qualcomm AI Engine Direct SDK (QAIRT) — and profile and validate them on real Qualcomm devices hosted in the cloud.

Join the Qualcomm AI Hub Compiler team and own the infrastructure that powers these model compilations. You will work across the full compilation pipeline — from model ingestion and graph optimization to backend dispatch across CPU, GPU, and NPU — ensuring models compile correctly, execute efficiently, and scale across a growing catalog of on-device use cases spanning vision, audio, speech, and multi-modal models.

Responsibilities
  • Compiler Pipeline & Infrastructure
  • Design, develop, and maintain the end-to-end compilation pipeline powering Qualcomm AI Hub Workbench, from PyTorch and ONNX model ingestion through graph optimization to deployable artifacts targeting LiteRT, ONNXRuntime, or QAIRT on Snapdragon SoCs
  • Build and maintain ONNX-based compilation paths using ONNX IR: graph transformation passes, op validation, and opset compatibility handling
  • Build and maintain PyTorch compilation paths consuming torch.export output, including dynamic shapes, custom ops, and ATen IR decomposition
  • Contribute to ONNXRuntime QNN execution provider: graph optimizations, graph partitioning, and op validation and lowerings
  • Collaborate with QAIRT and QNN teams to ensure correct and efficient model execution across CPU, GPU, and NPU backends
  • Build tooling to analyze, profile, and debug compilation failures, accuracy regressions, and performance degradations; develop clear, actionable developer-facing diagnostics
  • Model Catalog, Automation & Collaboration
  • Own compilation and validation of models published on Qualcomm AI Hub, ensuring correct conversion and verified performance across supported runtime targets
  • Build and maintain automated compilation pipelines and CI/CD evaluation harnesses to scale model onboarding as the Qualcomm AI Hub model catalog grows
  • Partner with internal Business Units to onboard models through Qualcomm AI Hub compilation workflows, translating deployment constraints (target SoC, latency budgets, memory limits) into concrete compilation strategies
  • Author technical documentation, tutorials, and example notebooks for the Qualcomm AI Hub developer community
Minimum Qualifications
  • Bachelor's degree in Computer Science, Engineering, Information Systems, or related field and 4+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience.
  • OR Master's degree in Computer Science, Engineering, Information Systems, or related field and 3+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience.
  • OR PhD in Computer Science, Engineering, Information Systems, or related field and 2+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience.
Preferred Qualifications
  • 3+ years of industry experience in ML infrastructure, compiler engineering, or AI framework development
  • Proficient in Python and C++
  • Solid understanding of ML compiler concepts (graph IRs, operator fusion, shape inference, lowering passes, backend partitioning) and hands-on experience with one or more compiler stacks such as MLIR, ONNX, or TVM
  • Experience with PyTorch model export (torch.export, torch.compile, FX, ATen IR) and on-device deployment frameworks such as LiteRT, ExecuTorch, or ONNXRuntime
  • Familiarity with SoC-level constraints (memory bandwidth, compute precision, NPU/DSP execution) and hardware-specific runtimes such as QAIRT/QNN is a plus
  • Experience building automated CI/CD pipelines for model compilation and validation at scale
  • Strong written and verbal communication skills; proficiency with git and software engineering best practices
Level of Responsibility
  • Works independently on open-ended compiler and infrastructure challenges
  • Provides technical guidance and mentorship to team members
  • Decision-making has broad impact — affecting compilation correctness, runtime performance, and the developer experience across Qualcomm AI Hub
  • Communicates complex compiler and runtime concepts to varied audiences: SoC engineers, BU partners, and external ML developers
  • Has meaningful influence on the Qualcomm AI Hub compiler roadmap, model catalog strategy, and cross-team runtime integration priorities

Qualcomm is an equal opportunity employer. If you are an individual with a disability and need an accommodation during the application/hiring process, rest assured that Qualcomm is committed to providing an accessible process. You may e-mail disability-accomodations@qualcomm.com or call Qualcomm's toll-free number found here. Upon request, Qualcomm will provide reasonable accommodations to support individuals with disabilities to be able participate in the hiring process. Qualcomm is also committed to making our workplace accessible for individuals with disabilities. (Keep in mind that this email address is used to provide reasonable accommodations for individuals with disabilities. We will not respond here to requests for updates on applications or resume inquiries).

To all Staffing and Recruiting Agencies: Our Careers Site is only for individuals seeking a job at Qualcomm. Staffing and recruiting agencies and individuals being represented by an agency are not authorized to use this site or to submit profiles, applications or resumes, and any such submissions will be considered unsolicited. Qualcomm does not accept unsolicited resumes or applications from agencies. Please do not forward resumes to our jobs alias, Qualcomm employees or any other company location. Qualcomm is not responsible for any fees related to unsolicited resumes/applications.

EEO Employer: Qualcomm is an equal opportunity employer; all qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, Veteran status, or any other protected classification.

Qualcomm expects its employees to abide by all applicable policies and procedures, including but not limited to security and other requirements regarding protection of Company confidential information and other confidential and/or proprietary information, to the extent those requirements are permissible under applicable law.

Pay Range And Other Compensation & Benefits

$160,500.00 - $240,700.00

The above pay scale reflects the broad, minimum to maximum, pay scale for this job code for the location for which it has been posted. Salary is only one component of total compensation at Qualcomm. We also offer a competitive annual discretionary bonus program and opportunity for annual RSU grants. In addition, our benefits package supports your success at work, at home, and at play. Your recruiter can discuss details, and you can review more about our US benefits at the linked page.

If you would like more information about this role, please contact Qualcomm Careers.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Machine Learning Engineer – AI/ML Compiler
Staff Machine Learning Engineer – AI/ML Compiler

Qualcomm • Santa Clara (CA)

On-site
USD 160,500 - 240,700
Staff Machine Learning Engineer – Model Optimization & Quantization
Staff Machine Learning Engineer – Model Optimization & Quantization

Qualcomm • Santa Clara (CA)

On-site
USD 161,000 - 241,000
Staff/Sr. Staff Software Engineer, AI Software Tools Development
Staff/Sr. Staff Software Engineer, AI Software Tools Development

Qualcomm • San Diego (CA)

On-site
USD 158,000 - 238,000
Annual discretionary bonus
RSU grants
Comprehensive benefits package
Machine Learning Compiler
Machine Learning Compiler

Qualcomm • New York (NY)

On-site
USD 141,000 - 211,000
Software Engineer, Core AI Software
Software Engineer, Core AI Software

Qualcomm • San Diego (CA)

On-site
USD 122,000 - 185,000
Machine Learning Compiler Engineer
Machine Learning Compiler Engineer

Qualcomm • New York (NY)

On-site
USD 200,000 - 302,000
Annual discretionary bonus program
Opportunity for annual RSU grants
Competitive benefits package
Staff & Senior Staff Software Engineer, AI Software Platform (Onsite)
Staff & Senior Staff Software Engineer, AI Software Platform (Onsite)

Qualcomm • San Diego (CA)

On-site
USD 158,400 - 237,600
Annual discretionary bonus
Competitive benefits package
Opportunity for RSU grants
Staff Machine Learning Engineer – Model Optimization & Quantization
Staff Machine Learning Engineer – Model Optimization & Quantization

Socket.dev • Santa Clara (CA)

On-site
USD 161,000 - 241,000
Machine Learning Compiler
Machine Learning Compiler

Nutanix • New York (NY)

On-site
USD 141,000 - 211,000
Entry Level & Senior Software Engineer, AI Software Platform (Onsite)
Entry Level & Senior Software Engineer, AI Software Platform (Onsite)

Qualcomm • San Diego (CA)

On-site
USD 141,000 - 211,000
Yearly bonus
RSU grants
Comprehensive benefits