Cloud-Scale Backend Engineer for ML Inference

Praxis, Inc.

San Francisco (CA)

On-site

USD 170,000 - 250,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NomadicML in San Francisco is seeking a Backend / Infrastructure Engineer to join our platform that powers video intelligence. You’ll build scalable cloud ingestion, distributed GPU inference pipelines, and robust APIs/SDKs used by enterprises worldwide.

You’ll collaborate with ML researchers to productionize models, automate deployment, and improve reliability, speed, and developer experience across storage, scheduling, and orchestration.

Qualifications

  • Backend systems and cloud experience for large-scale inference.
  • Experience designing REST/gRPC APIs and developer-facing SDKs.
  • Knowledge of asynchronous job orchestration frameworks.
  • Hands-on with GPU inference scaling and distributed compute.
  • Ability to productionize research-grade systems.

Responsibilities

  • Build and scale the backbone powering NomadicML's video intelligence platform.
  • Collaborate with ML researchers to productionize models and automate deployment.
  • Expose capabilities via clean APIs/SDKs and improve observability.

Skills

Python
Go
TypeScript

Tools

AWS
GCP
Azure
Kubernetes
Docker
Ray
Dagster
Temporal

Job description

NomadicML in San Francisco is seeking a Backend / Infrastructure Engineer to join our platform that powers video intelligence. You’ll build scalable cloud ingestion, distributed GPU inference pipelines, and robust APIs/SDKs used by enterprises worldwide.

You’ll collaborate with ML researchers to productionize models, automate deployment, and improve reliability, speed, and developer experience across storage, scheduling, and orchestration.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Backend Software Engineer (ML Infra)
Backend Software Engineer (ML Infra)

Rockstar • San Francisco (CA)

On-site
USD 100,000 - 130,000
Senior ML Infra Engineer: Low-Latency Inference
Senior ML Infra Engineer: Low-Latency Inference

Doist • New York (NY)

Hybrid
USD 170,000 - 250,000
Healthcare
401k plan with matching
Hybrid work model
+1
Staff Backend Engineer — ML Systems & Scale
Staff Backend Engineer — ML Systems & Scale

additiveai • San Francisco (CA)

On-site
USD 140,000 - 210,000
Relocation assistance
Backend Engineer (ML Infra) — Scale AI Training & Inference
Backend Engineer (ML Infra) — Scale AI Training & Inference

Rockstar • San Francisco (CA)

On-site
USD 100,000 - 130,000
ML Infrastructure Engineer
ML Infrastructure Engineer

Acceler8 Talent • San Francisco (CA)

Hybrid
USD 233,000 - 275,000
Member of Technical Staff
Member of Technical Staff

kadence • San Francisco (CA)

On-site
USD 120,000 - 160,000
ML Infra Tech Lead: Scalable Training & Inference
ML Infra Tech Lead: Scalable Training & Inference

Reducto • San Francisco (CA)

On-site
USD 180,000 - 260,000
Unlimited PTO
Daily Lunch
Commuter Reimbursement
+3
Senior ML Systems Engineer — Inference & Scale
Senior ML Systems Engineer — Inference & Scale

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 170,000 - 240,000
Senior ML Inference Engineer: Production Systems
Senior ML Inference Engineer: Production Systems

MakerMaker • San Francisco (CA)

On-site
USD 180,000 - 240,000
Senior ML Infrastructure Engineer - Low-Latency & Scale
Senior ML Infrastructure Engineer - Low-Latency & Scale

Patreon • New York (NY)

Hybrid
USD 150,000 - 210,000
Healthcare
Equity plans
Flexible time off
+5