Staff Engineer: Foundation Model API & GPU Inference

Databricks Inc.

San Francisco (CA)

On-site

USD 192,000 - 260,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Comprehensive benefits
Diversity and inclusion initiatives

Job summary

A leading data and AI company is seeking a Staff Engineer to design and implement core systems for Foundation Model Serving. The ideal candidate will have over 10 years of experience in building large-scale distributed systems and will collaborate closely across teams to ensure operational excellence in GPU serving workloads. Competitive salary range of $192,000 to $260,000 is offered with comprehensive benefits.

Qualifications

  • 10+ years of experience building and operating large-scale distributed systems.
  • Experience leading high-scale operationally sensitive backend systems.
  • Strong foundation in algorithms, data structures, and system design.

Responsibilities

  • Design and implement core systems and APIs for Foundation Model Serving.
  • Drive architectural decisions and trade-offs for GPU serving workloads.
  • Collaborate cross-functionally to translate customer needs into reliable systems.

Skills

Building large-scale distributed systems
Leading backend systems
Algorithms and data structures
Mentoring engineers

Job description

A leading data and AI company is seeking a Staff Engineer to design and implement core systems for Foundation Model Serving. The ideal candidate will have over 10 years of experience in building large-scale distributed systems and will collaborate closely across teams to ensure operational excellence in GPU serving workloads. Competitive salary range of $192,000 to $260,000 is offered with comprehensive benefits.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff Engineer - Foundation Model Serving & Inference
Staff Engineer - Foundation Model Serving & Inference

Menlo Ventures • San Francisco (CA)

On-site
USD 120,000 - 160,000
Staff Backend Engineer, Foundation Model Serving
Staff Backend Engineer, Foundation Model Serving

Menlo Ventures • San Francisco (CA)

On-site
USD 192,000 - 260,000
Comprehensive benefits
Annual performance bonus
Equity options
Senior ML Training Systems Engineer - Distributed GPU Infra
Senior ML Training Systems Engineer - Distributed GPU Infra

Baseten • San Francisco (CA)

On-site
USD 150,000 - 200,000
Competitive compensation, including equity
100% coverage of medical, dental, and vision insurance
Generous PTO policy
+2
GPU Performance Engineer: Scale AI Inference
GPU Performance Engineer: Scale AI Inference

Anthropic • San Francisco (CA)

On-site
USD 315,000 - 560,000
Competitive salary
Equity opportunities
Flexible working hours
+1
Foundation Model AI Infra Architect - Production-Scale
Foundation Model AI Infra Architect - Production-Scale

Vinci4D.ai • Palo Alto (CA)

On-site
USD 180,000 - 220,000
Senior Distributed Systems Engineer — High-Perf GPU
Senior Distributed Systems Engineer — High-Perf GPU

Lever, Inc. • Sunnyvale (CA)

On-site
USD 180,000 - 250,000
Staff Engineer, Inference Runtime — High-Performance AI Serving
Staff Engineer, Inference Runtime — High-Performance AI Serving

Anthropic • Seattle (WA)

Hybrid
USD 405,000 - 485,000
Staff Software Engineer, Foundational Model Serving
Staff Software Engineer, Foundational Model Serving

Cacheflow • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior Inference Systems Engineer — Large-Scale GPUs
Senior Inference Systems Engineer — Large-Scale GPUs

RadixArk • Palo Alto (CA)

On-site
USD 190,000 - 260,000
Competitive compensation
Meaningful equity
Comprehensive benefits
+1
Senior Inference Performance Engineer — GPU & CUDA
Senior Inference Performance Engineer — GPU & CUDA

Inference • San Francisco (CA)

Hybrid
USD 220,000 - 320,000
Competitive compensation
Equity in a high-growth startup
Comprehensive benefits