Senior ML Serving Engineer for LLMs & Inference

Alldus

San Jose (CA)

On-site

USD 180,000 - 220,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A tech company in AI/ML is seeking a Senior Software Engineer specializing in ML Serving to build robust infrastructure for ML models. The ideal candidate has 5+ years of experience in software engineering, with a focus on ML serving. Proficiency in Python and knowledge of various serving frameworks are essential. This full-time role is located in San Jose, California and offers a competitive salary.

Qualifications

  • 5+ years of software engineering experience with an ML serving focus.
  • Proven experience deploying large language models in production.
  • Strong programming skills in Python, Go, or C++.

Responsibilities

  • Design and build ML serving infrastructure.
  • Optimize inference pipelines for efficiency.
  • Integrate models into varied customer environments.

Skills

ML serving
Python
Cloud platforms (AWS, GCP, Azure)
Distributed systems
Container orchestration (Kubernetes, Docker)

Tools

TensorFlow Serving
TorchServe
Triton Inference Server
BentoML
Ray Serve

Job description

A tech company in AI/ML is seeking a Senior Software Engineer specializing in ML Serving to build robust infrastructure for ML models. The ideal candidate has 5+ years of experience in software engineering, with a focus on ML serving. Proficiency in Python and knowledge of various serving frameworks are essential. This full-time role is located in San Jose, California and offers a competitive salary.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Engineer - LLM Inference & Serving at Scale
Staff Engineer - LLM Inference & Serving at Scale

Prime Intellect • San Francisco (CA)

Hybrid
USD 150,000 - 300,000
Remote ML Engineering Manager: LLM Serving & Infra
Remote ML Engineering Manager: LLM Serving & Infra

Jobgether • United States

Remote
USD 176,000 - 252,000
Senior ML Infra Engineer: Build Scalable LLM Systems
Senior ML Infra Engineer: Build Scalable LLM Systems

ServiceNow • Mountain View (CA)

On-site
USD 130,000 - 180,000
Senior ML Inference & Serving Engineer (LLM Deploy)
Senior ML Inference & Serving Engineer (LLM Deploy)

Google DeepMind • Mountain View (CA)

On-site
USD 207,000 - 300,000
Senior ML Engineer: Build Scalable AI Solutions
Senior ML Engineer: Build Scalable AI Solutions

Arcade Software, Inc. • San Francisco (CA)

Hybrid
USD 180,000 - 300,000
Unlimited PTO
Rich health/401(k) plans
Meeting-light culture
+2
Senior Cloud AI LLM Serving Engineer
Senior Cloud AI LLM Serving Engineer

Qualcomm • San Diego (CA)

On-site
USD 158,000 - 238,000
Competitive annual discretionary bonus program
Potential RSU grants
Comprehensive benefits package
Senior ML Engineer – Enterprise AI & LLM Systems
Senior ML Engineer – Enterprise AI & LLM Systems

Factualiq • Mission (KS)

On-site
USD 110,000 - 130,000
Staff Engineer, LLM Inference & Infra
Staff Engineer, LLM Inference & Infra

Prime Intellect • United States

Hybrid
USD 120,000 - 150,000
Competitive compensation
Flexible work arrangement
Full visa sponsorship
+2
Senior ML Engineer: GenAI & Production ML Lead
Senior ML Engineer: GenAI & Production ML Lead

Adobe Inc. • San Jose (CA)

On-site
USD 183,000 - 266,000
Comprehensive benefits programs
Collaborative work culture
Opportunities for career advancement
Senior Model Serving Engineer - Low-Latency AI Platform
Senior Model Serving Engineer - Low-Latency AI Platform

Menlo Ventures • San Francisco (CA)

On-site
USD 192,000 - 260,000
Comprehensive benefits
Eligibility for annual performance bonus
Equity opportunities