Senior/Staff AI Performance Engineer: Inference Optimization

Qualcomm

San Diego (CA)

On-site

USD 178,400 - 267,600

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Competitive annual discretionary bonus
RSU grants
Highly competitive benefits package

Job summary

Qualcomm in San Diego is looking for an AI Engineer specializing in machine learning. You will convert and optimize models, analyze performance, and collaborate across teams to advance AI technologies.

The ideal candidate should have extensive hands-on experience with PyTorch and ONNX, deep knowledge of transformer architectures, and a relevant master's or PhD degree. This position offers a competitive salary and benefits package.

Qualifications

  • 6+ years of engineering experience with relevant degrees.
  • Hands-on experience in production-grade environments.
  • Deep understanding of performance trade-offs.

Responsibilities

  • Convert, optimize, and deploy models for inference.
  • Analyze performance of AI workloads.
  • Collaborate with teams to drive solutions.

Skills

Building and optimizing language models
Python programming
Understanding transformer architectures
Machine learning compilers
Problem-solving skills

Education

MS in Computer Science or related field
PhD in Computer Science or related field

Tools

PyTorch
ONNX

Job description

Qualcomm in San Diego is looking for an AI Engineer specializing in machine learning. You will convert and optimize models, analyze performance, and collaborate across teams to advance AI technologies.

The ideal candidate should have extensive hands-on experience with PyTorch and ONNX, deep knowledge of transformer architectures, and a relevant master's or PhD degree. This position offers a competitive salary and benefits package.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Staff AI/ML Inference Engineer – Graph & Runtime Ops
Senior Staff AI/ML Inference Engineer – Graph & Runtime Ops

Qualcomm • San Diego (CA)

On-site
USD 158,000 - 238,000
AI Performance Engineer (Cloud AI Engineering), Sr | Staff | Sr. Staff
AI Performance Engineer (Cloud AI Engineering), Sr | Staff | Sr. Staff

Qualcomm • San Diego (CA)

On-site
USD 178,400 - 267,600
Competitive annual discretionary bonus
RSU grants
Highly competitive benefits package
Staff AI Software Engineer - Edge Inference
Staff AI Software Engineer - Edge Inference

Qualcomm • San Diego (CA)

On-site
USD 158,400 - 237,600
Competitive annual discretionary bonus
Annual RSU grants
Comprehensive benefits package
Staff AI Software Engineer: GenAI Inference on Snapdragon
Staff AI Software Engineer: GenAI Inference on Snapdragon

Qualcomm • Raleigh (NC)

On-site
USD 162,000 - 243,000
AI Model Optimization Architect
AI Model Optimization Architect

Qualcomm • Austin (TX)

On-site
USD 158,000 - 238,000
Staff/Sr. Staff Software Engineer, AI Software Tools (Onsite)
Staff/Sr. Staff Software Engineer, AI Software Tools (Onsite)

Qualcomm • San Diego (CA)

On-site
USD 158,000 - 238,000
Staff AI Model Optimization Architect for LLMs & Multimodal
Staff AI Model Optimization Architect for LLMs & Multimodal

Qualcomm • Austin (TX)

On-site
USD 158,000 - 238,000
AI Systems Hardware Engineer: Power, Perf & TCO Optimizer
AI Systems Hardware Engineer: Power, Perf & TCO Optimizer

Qualcomm • San Diego (CA)

On-site
USD 164,000 - 246,000
Competitive annual discretionary bonus
Opportunity for annual RSU grants
Highly competitive benefits package
Data Center AI Systems Engineer
Data Center AI Systems Engineer

Qualcomm • San Diego (CA)

On-site
USD 111,300 - 166,900
Senior AI Systems Performance Engineer: Drive SOTA Inference
Senior AI Systems Performance Engineer: Drive SOTA Inference

SambaNova • Palo Alto (CA)

On-site
USD 120,000 - 150,000
95% premium coverage for employee medical insurance
Health Savings Account with employer contribution
Flexible Spending Account options