Senior Cloud AI LLM Serving Engineer

Qualcomm

San Diego (CA)

On-site

USD 158,400 - 237,600

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive annual discretionary bonus program
Potential RSU grants
Comprehensive benefits package

Job summary

A leading technology firm in San Diego seeks an LLM Serving Engineer to develop scalable AI solutions. This role involves building LLM inference platforms and collaborating with teams to drive innovations in machine learning. Responsibilities include optimizing deep learning workloads and utilizing advanced techniques for efficient serving. Candidates should have strong experience with LLM packages, a solid foundation in computer science, and experience in Python development. Competitive salary and benefits are offered.

Qualifications

  • Hands-on experience with Triton-Inference Server and similar packages.
  • Strong experience in developing language models, especially using PyTorch.
  • Excellent understanding of algorithms and parallel programming.

Responsibilities

  • Build a scalable LLM inference platform using advanced techniques.
  • Contribute to development of LLM Serving packages.
  • Drive efficient serving with load balancing and routing.

Skills

Experience with LLM serving packages
Deep understanding of foundational LLMs
Experience in developing language models using PyTorch
Computer science fundamentals
Understanding of computer architecture and ML accelerators
Python development skills
Experience in optimizing deep learning workloads
Problem-solving skills
Excellent communication skills

Education

Bachelor’s degree in relevant field
Master’s degree in relevant field
PhD in relevant field

Job description

A leading technology firm in San Diego seeks an LLM Serving Engineer to develop scalable AI solutions. This role involves building LLM inference platforms and collaborating with teams to drive innovations in machine learning. Responsibilities include optimizing deep learning workloads and utilizing advanced techniques for efficient serving. Candidates should have strong experience with LLM packages, a solid foundation in computer science, and experience in Python development. Competitive salary and benefits are offered.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

LLM Serving Engineer, Cloud AI Platform Architect
LLM Serving Engineer, Cloud AI Platform Architect

Qualcomm • Austin (TX)

On-site
USD 158,000 - 238,000
Senior AI Engineer - LLM Platform Lead
Senior AI Engineer - LLM Platform Lead

ZS • South San Francisco (CA)

Hybrid
USD 120,000 - 150,000
Hybrid LLM Engineer: Build & Deploy Scalable AI Models
Hybrid LLM Engineer: Build & Deploy Scalable AI Models

Take2 Consulting, LLC • San Jose (CA)

Hybrid
USD 175,000 - 185,000
Lead AI Cloud Architect for Scalable LLM Platforms
Lead AI Cloud Architect for Scalable LLM Platforms

Zhone Technologies, Inc. • Plano (TX)

On-site
USD 130,000 - 160,000
Gen AI Engineer: LLMs, NLP, Cloud & DataMarts
Gen AI Engineer: LLMs, NLP, Cloud & DataMarts

TechDigital Group • New York (NY)

On-site
USD 80,000 - 150,000
Senior AI Engineer: Agentic LLMs & Scalable AI Systems
Senior AI Engineer: Agentic LLMs & Scalable AI Systems

Harnham • San Francisco (CA)

Hybrid
USD 175,000 - 210,000
Competitive base + equity
Work on cutting-edge AI
High-impact role with ownership
AI Engineer: LLM Architect for Scalable Systems
AI Engineer: LLM Architect for Scalable Systems

ZS • South San Francisco (CA)

Hybrid
USD 120,000 - 150,000
Health and well-being benefits
Financial planning
Professional development opportunities
Senior ML Serving Engineer for LLMs & Inference
Senior ML Serving Engineer for LLMs & Inference

Alldus • San Jose (CA)

On-site
USD 180,000 - 220,000
Staff Engineer, LLM Inference & Infra
Staff Engineer, LLM Inference & Infra

Prime Intellect • United States

Hybrid
USD 120,000 - 150,000
Competitive compensation
Flexible work arrangement
Full visa sponsorship
+2
Senior LLM Engineer (Java/Spring) — Cloud & NLP
Senior LLM Engineer (Java/Spring) — Cloud & NLP

TechDigital Group • Charlotte (NC)

On-site
USD 100,000 - 130,000