Staff Engineer - Foundation Model Serving & Inference

Menlo Ventures

San Francisco (CA)

On-site

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading tech enterprise in San Francisco is seeking a Staff Engineer to shape their foundation model API product. You will design and build systems that ensure efficient performance on high-throughput, low-latency GPU workloads. The ideal candidate will have experience with operational sensitive systems but does not need prior AI experience. Collaboration with various teams is essential to improve product offerings and architectural decisions.

Qualifications

  • No prior ML or AI experience is necessary.
  • Strong engineering skills required.
  • Ability to work in a collaborative environment.

Responsibilities

  • Design and build systems for high-throughput, low-latency inference.
  • Influence architectural direction for AI model serving.
  • Collaborate across various teams to enhance product experience.

Skills

Experience with high scale operational sensitive systems
Interest in building LLM APIs
Experience in customer facing APIs

Job description

A leading tech enterprise in San Francisco is seeking a Staff Engineer to shape their foundation model API product. You will design and build systems that ensure efficient performance on high-throughput, low-latency GPU workloads. The ideal candidate will have experience with operational sensitive systems but does not need prior AI experience. Collaboration with various teams is essential to improve product offerings and architectural decisions.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Engineer: Foundation Model API & GPU Inference
Staff Engineer: Foundation Model API & GPU Inference

Databricks Inc. • San Francisco (CA)

On-site
USD 192,000 - 260,000
Comprehensive benefits
Diversity and inclusion initiatives
Staff Backend Engineer, Foundation Model Serving
Staff Backend Engineer, Foundation Model Serving

Menlo Ventures • San Francisco (CA)

On-site
USD 192,000 - 260,000
Comprehensive benefits
Annual performance bonus
Equity options
Staff Engineer: Foundation Model Serving & APIs
Staff Engineer: Foundation Model Serving & APIs

RoShay Services • San Francisco (CA)

On-site
USD 180,000 - 240,000
Staff Engineer – Foundation Model Serving & APIs
Staff Engineer – Foundation Model Serving & APIs

RoShay Services • San Francisco (CA)

On-site
USD 180,000 - 240,000
Staff Software Engineer - Foundation Model API & Infra
Staff Software Engineer - Foundation Model API & Infra

Neura Market • San Francisco (CA), Northern (KY)

Hybrid
USD 192,000 - 260,000
Staff Software Engineer, Foundational Model Serving
Staff Software Engineer, Foundational Model Serving

Cacheflow • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior Model Serving Engineer - Low-Latency AI Platform
Senior Model Serving Engineer - Low-Latency AI Platform

Menlo Ventures • San Francisco (CA)

On-site
USD 192,000 - 260,000
Comprehensive benefits
Eligibility for annual performance bonus
Equity opportunities
Staff ML Engineer: Foundation Models for AI & Search
Staff ML Engineer: Foundation Models for AI & Search

Apple Inc. • San Francisco (CA)

On-site
USD 181,100 - 318,400
Comprehensive medical and dental coverage
Employee stock purchase plan
Tuition reimbursement
Senior AI Model Serving Engineer — Low-Latency Inference
Senior AI Model Serving Engineer — Low-Latency Inference

Menlo Ventures • San Francisco (CA)

On-site
USD 166,000 - 225,000
Annual performance bonus
Equity options
Comprehensive benefits package
Staff Foundation Model Inference Engineer
Staff Foundation Model Inference Engineer

United States Digital Space LLC • San Francisco (CA)

On-site
USD 190,000 - 265,000