Staff Software Engineer: Foundation Model Inference at Scale

Databricks

San Francisco (CA)

On-site

USD 190,000 - 265,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Databricks seeks highly autonomous backend/infrastructure engineers to power enterprise-scale model inference on the Mosaic AI platform. You will build and optimize LLM infrastructure, serving OpenAI, Anthropic, Gemini and self-hosted models, tackling reliability and latency at scale.

Collaborate with platform, infra, and ML teams to deliver end-to-end experiences, influence developer workflows, and drive observability.

Qualifications

  • 8+ years of experience in backend or infrastructure engineering.
  • Experience with distributed systems, scalable APIs, or cloud-native infrastructure.
  • Experience with real-time serving, ML infrastructure, or GPU orchestration.
  • Familiarity with service-oriented architecture, deployment pipelines, and system observability.

Responsibilities

  • Build LLM infrastructure powering large-scale inference workloads for customers through partner models (OpenAI, Anthropic, Gemini) and self-hosted models (Qwen, GPT-OSS, Llama).
  • Improve reliability, latency, and efficiency of distributed AI workloads.
  • Collaborate with platform, infra, and ML teams to deliver seamless end-to-end experiences.
  • Shape how developers and data scientists build and interact with AI on Databricks.

Skills

Backend engineering
Distributed systems
Cloud-native infrastructure
Real-time serving
GPU orchestration
Observability

Tools

SageMaker
Vertex AI
Azure ML
MLflow
PyTorch
Ray
vLLM
SGLang

Job description

Databricks seeks highly autonomous backend/infrastructure engineers to power enterprise-scale model inference on the Mosaic AI platform. You will build and optimize LLM infrastructure, serving OpenAI, Anthropic, Gemini and self-hosted models, tackling reliability and latency at scale.

Collaborate with platform, infra, and ML teams to deliver end-to-end experiences, influence developer workflows, and drive observability.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff Software Engineer, Foundational Model Serving
Staff Software Engineer, Foundational Model Serving

Cacheflow • San Francisco (CA)

On-site
USD 120,000 - 160,000
Staff Engineer, Foundation Models & AI Infrastructure
Staff Engineer, Foundation Models & AI Infrastructure

Databricks • San Francisco (CA)

On-site
USD 190,000 - 265,000
Staff Software Engineer- Foundation Model Inference San Francisco, California
Staff Software Engineer- Foundation Model Inference San Francisco, California

Databricks Inc. • San Francisco (CA)

On-site
USD 190,000 - 265,000
Staff Software Engineer- Foundation Model Inference
Staff Software Engineer- Foundation Model Inference

Databricks • San Francisco (CA)

On-site
USD 190,000 - 265,000
Staff Software Engineer- Foundation Model Inference
Staff Software Engineer- Foundation Model Inference

Menlo Ventures • San Francisco (CA)

On-site
USD 190,000 - 265,000
Senior AI Infra Engineer — Inference Platform
Senior AI Infra Engineer — Inference Platform

Databricks Inc. • San Francisco (CA)

On-site
USD 190,000 - 265,000
Staff Software Engineer, GenAI Inference Engine
Staff Software Engineer, GenAI Inference Engine

Databricks • United States

Remote
USD 191,000 - 233,000
Equity
Staff AI Infra Engineer, Foundation Model Inference
Staff AI Infra Engineer, Foundation Model Inference

Menlo Ventures • San Francisco (CA)

On-site
USD 190,000 - 265,000
Staff Backend Software Engineer
Staff Backend Software Engineer

Databricks • New York (NY)

On-site
USD 190,000 - 261,250
Staff Software Engineer, Foundational Model Serving
Staff Software Engineer, Foundational Model Serving

Databricks Inc. • San Francisco (CA)

On-site
USD 192,000 - 260,000
Comprehensive benefits
Diversity and inclusion initiatives