Senior AI Model Serving Engineer — Low-Latency Inference

Menlo Ventures

San Francisco (CA)

On-site

USD 166,000 - 225,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Annual performance bonus
Equity options
Comprehensive benefits package

Job summary

A leading data and AI company in San Francisco is seeking a Senior Engineer to enhance their Model Serving platform. This role requires expertise in building large-scale distributed systems and collaboration across teams to optimize performance and reliability. Ideal candidates will have a strong foundation in algorithms and system design, along with a passion for mentoring others. The position offers a competitive salary and generous benefits.

Qualifications

  • 5+ years of experience building large-scale distributed systems.
  • Experience in model serving, inference systems, or related infrastructure.
  • Strong background in algorithms, data structures, and system design.

Responsibilities

  • Design core systems and APIs for Model Serving.
  • Drive architectural decisions for performance optimization.
  • Collaborate with cross-functional teams.

Skills

Building large-scale distributed systems
Model serving
System design
Collaborative communication
Customer-focused mindset
Mentoring engineers

Job description

A leading data and AI company in San Francisco is seeking a Senior Engineer to enhance their Model Serving platform. This role requires expertise in building large-scale distributed systems and collaboration across teams to optimize performance and reliability. Ideal candidates will have a strong foundation in algorithms and system design, along with a passion for mentoring others. The position offers a competitive salary and generous benefits.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Model Serving Engineer - Low-Latency AI Platform
Senior Model Serving Engineer - Low-Latency AI Platform

Menlo Ventures • San Francisco (CA)

On-site
USD 192,000 - 260,000
Comprehensive benefits
Eligibility for annual performance bonus
Equity opportunities
Senior Engineer, Model Serving & Inference
Senior Engineer, Model Serving & Inference

Databricks • San Francisco (CA)

On-site
USD 166,000 - 225,000
Senior ML Engineer - Low-Latency Inference & Systems
Senior ML Engineer - Low-Latency Inference & Systems

Inworld • Germany (OH)

Hybrid
USD 120,000 - 160,000
Lead AI Model-Serving Platform Engineer
Lead AI Model-Serving Platform Engineer

Sciforium • San Francisco (CA)

On-site
USD 180,000 - 240,000
Medical, dental, and vision insurance
401k plan
Daily lunch, snacks, and beverages
+2
Senior AI Inference Optimizations Engineer — Remote
Senior AI Inference Optimizations Engineer — Remote

DigitalOcean • Seattle (WA)

Remote
USD 167,000 - 209,000
Senior AI Systems Performance Engineer: Drive SOTA Inference
Senior AI Systems Performance Engineer: Drive SOTA Inference

SambaNova • Palo Alto (CA)

On-site
USD 120,000 - 150,000
95% premium coverage for employee medical insurance
Health Savings Account with employer contribution
Flexible Spending Account options
Senior Model Inference Engineer for Production-Scale AI
Senior Model Inference Engineer for Production-Scale AI

OpenAI • San Francisco (CA)

On-site
USD 325,000 - 490,000
Realtime ML Inference Engineer — Scalable Serving
Realtime ML Inference Engineer — Scalable Serving

Yobi AI • New York (NY)

Remote
USD 120,000 - 150,000
Senior AI Inference Infrastructure Engineer
Senior AI Inference Infrastructure Engineer

Modular • United States

Hybrid
USD 167,000 - 273,000
Staff AI Cloud Platform Engineer – Inference & Training
Staff AI Cloud Platform Engineer – Inference & Training

Cerebras • Sunnyvale (CA)

On-site
USD 120,000 - 150,000