LLM Serving Engineer, Cloud AI Platform Architect

Qualcomm

Austin (TX)

On-site

USD 158,400 - 237,600

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Competitive salary
Annual discretionary bonus
RSU grants opportunity

Job summary

A leading technology company in Austin is seeking an LLM Serving Engineer to build scalable inference platforms. The role requires hands-on experience with LLM serving packages and strong Python skills. Candidates should be proactive learners familiar with developing language models using PyTorch. A collaborative mindset and excellent communication skills are essential for success in this fast-paced environment. The position offers a competitive compensation package along with growth opportunities and a dynamic work culture.

Qualifications

  • Hands-on experience in one or more LLM serving/Orchestration packages.
  • Strong experience in developing language models using PyTorch.
  • Strong Python development skills for large-scale projects.

Responsibilities

  • Build a scalable LLM inference platform using advanced techniques.
  • Contribute to LLM Serving packages development.
  • Work closely with customers to drive solutions.

Skills

LLM Serving/Orchestration packages
Deep understanding of foundational LLMs
Experience in developing language models using PyTorch
Strong computer science fundamentals
Understanding of computer architecture
Strong Python development skills
Experience in analyzing deep learning workloads
Proactive learning on inference optimization techniques
Communication and problem-solving skills

Education

Bachelor's degree in Computer Science or related field
Master's degree in Computer Science or related field
PhD in Computer Science or related field

Job description

A leading technology company in Austin is seeking an LLM Serving Engineer to build scalable inference platforms. The role requires hands-on experience with LLM serving packages and strong Python skills. Candidates should be proactive learners familiar with developing language models using PyTorch. A collaborative mindset and excellent communication skills are essential for success in this fast-paced environment. The position offers a competitive compensation package along with growth opportunities and a dynamic work culture.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Cloud AI LLM Serving Engineer
Senior Cloud AI LLM Serving Engineer

Qualcomm • San Diego (CA)

On-site
USD 158,400 - 237,600
Competitive annual discretionary bonus program
Potential RSU grants
Comprehensive benefits package
Staff Engineer, LLM Inference & Infra
Staff Engineer, LLM Inference & Infra

Prime Intellect • United States

Hybrid
USD 120,000 - 150,000
Competitive compensation
Flexible work arrangement
Full visa sponsorship
+2
Lead AI Engineer — LLM & Applied AI Solutions
Lead AI Engineer — LLM & Applied AI Solutions

AIS Info • Irving (TX)

On-site
USD 120,000 - 150,000
AI Infrastructure Engineer: Scale LLMs & Cloud
AI Infrastructure Engineer: Scale LLMs & Cloud

Newspaper WordPress • Austin (TX)

On-site
USD 100,000 - 140,000
Competitive salary exceeding six figures
STEM OPT extension
Potential H1-B sponsorship
+1
Staff ML Systems Engineer - LLM Serving & RL
Staff ML Systems Engineer - LLM Serving & RL

Prime Intellect • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 300,000
Remote option
Visa sponsorship
Relocation support
+2
LLM Inference Engineer (Mid, Senior, Staff)
LLM Inference Engineer (Mid, Senior, Staff)

Hippocratic AI Inc. • Menlo Park (CA)

On-site
USD 180,000 - 280,000
LLM Applications Engineer: AI-First Platform & Equity
LLM Applications Engineer: AI-First Platform & Equity

Uncountable Inc. • San Francisco (CA), New York (NY)

Hybrid
USD 130,000 - 175,000
Competitive Salary and Equity
Health and Dental Insurance
401K with Employer Contribution
Gen AI ML Engineer — Build & Deploy LLMs in Cloud
Gen AI ML Engineer — Build & Deploy LLMs in Cloud

MACHINE LEARNING TECHNOLOGIES LLC • Austin (TX)

On-site
USD 83,000 - 165,000
AI Engineer: LLM Apps, Retrieval & MLOps
AI Engineer: LLM Apps, Retrieval & MLOps

Ethereum Technologies LLC • Austin (TX)

On-site
USD 110,000 - 170,000
Senior AI Engineer - LLM Platform Lead
Senior AI Engineer - LLM Platform Lead

ZS • South San Francisco (CA)

Hybrid
USD 120,000 - 150,000
Health and well-being benefits
Financial planning support
Professional development programs