ML Infrastructure Engineer

Acceler8 Talent

San Francisco (CA)

Hybrid

USD 247,500 - 302,500

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A forward-thinking AI company is seeking a Senior ML Infrastructure / Backend Engineer to develop backend systems supporting real-time AI functionality. You will manage APIs, scale cloud infrastructure, and ensure system reliability while collaborating with various teams. The ideal candidate should possess strong backend development skills and experience with cloud platforms, shaping the future of AI-powered digital interactions. The salary is competitive, reaching up to $275k depending on qualifications, with a hybrid work model in Los Angeles or San Francisco.

Qualifications

  • Strong experience building and owning backend or distributed systems.
  • Hands-on experience designing APIs, preferably using Python.
  • Experience running ML-backed systems in production environments.

Responsibilities

  • Own backend services and APIs that expose ML-powered features.
  • Design and operate orchestration layers for ML workloads.
  • Scale infrastructure to support high user traffic.

Skills

Backend services development
API design and operation (Python preferred)
Cloud platforms (AWS, GCP)
Performance tuning
Debugging and operational ownership

Tools

FastAPI
Flask
gRPC
Ray Serve
Triton

Job description

Series C Startup | AI-Powered 3D & Avatar Platform | Hybrid (LA or SF)

We’re hiring a Senior ML Infrastructure / Backend Engineer to join a well-funded AI company building the visual and interaction layer for the next generation of AI-powered digital identities.

This team is developing production systems that bring AI characters out of chat boxes and into real-time, interactive 3D experiences.

You’ll own backend and infrastructure systems that serve ML-powered functionality at scale — supporting high-concurrency user traffic, low-latency inference, and rapid iteration as the platform grows.

You’ll work closely with ML researchers, platform engineers, and product teams to take models from experimentation to reliable, scalable production services.

What You’ll Do:

  • Own backend services and APIs that expose ML-powered features to real users
  • Design and operate orchestration layers for ML workloads (routing, batching, retries, concurrency)
  • Deploy and scale ML-backed services in cloud environments
  • Scale infrastructure to support thousands to hundreds of thousands of daily requests
  • Implement observability, monitoring, and alerting to ensure system reliability
  • Partner closely with ML teams to productionize generative and ML models
  • Improve end-to-end efficiency across inference, post-processing, and data pipelines

What We’re Looking For:

  • Strong, production-level experience building and owning backend or distributed systems
  • Hands-on experience designing and operating APIs (Python preferred — FastAPI, Flask, or gRPC)
  • Experience deploying and running ML-backed systems in production environments
  • Proven ability to scale systems under real user traffic with attention to latency and reliability
  • Experience with cloud platforms (AWS, GCP, or similar) and containerized deployments
  • Strong debugging, performance tuning, and operational ownership skills

Nice to have:

  • Experience with ML inference optimization (quantization, mixed precision, ONNX, TensorRT)
  • Familiarity with scalable inference frameworks (Ray Serve, Triton, TorchServe, SageMaker)
  • Exposure to generative models (diffusion or transformer-based systems)
  • Experience running GPU-backed or high-performance workloads in production

This Role Is:

  • Hybrid (Los Angeles or San Francisco office)
  • Salary range: upto $275k base, depending on experience

If you’re excited about building the infrastructure that powers real-time, embodied AI — and want ownership over systems that actually ship — we’d love to talk.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Backend Software Engineer (ML Infra)
Backend Software Engineer (ML Infra)

Rockstar • San Francisco (CA)

On-site
USD 100,000 - 130,000
ML Infrastructure Engineer
ML Infrastructure Engineer

Lattice, Inc. • San Francisco (CA)

Hybrid
USD 200,000 - 280,000
Competitive salary
Premium health, dental, and vision insurance
Unlimited PTO
+2
Senior AI/ML Engineer
Senior AI/ML Engineer

CB Smart Recruit • Los Angeles (CA)

On-site
USD 180,000 - 350,000
Competitive sign-on bonus
Comprehensive benefits package
Member of Technical Staff
Member of Technical Staff

kadence • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior ML Infrastructure Engineer
Senior ML Infrastructure Engineer

Rebar • New York (NY)

On-site
USD 120,000 - 160,000
Comprehensive medical, dental, and vision coverage
Free lunches and dinners
Backend Engineer
Backend Engineer

Space Executive • Berkeley (CA)

Remote
USD 125,000 - 225,000
Medical, dental, vision
401(k)
Unlimited PTO
+2
Staff Software Engineer (AI Infrastructure)
Staff Software Engineer (AI Infrastructure)

DeepRec.ai • Palo Alto (CA)

On-site
USD 180,000 - 320,000
ML Infrastructure Engineer
ML Infrastructure Engineer

Strativ Group • Menlo Park (CA)

On-site
USD 250,000 - 320,000
ML Engineer – AI-Powered Automation & Workflow Intelligence
ML Engineer – AI-Powered Automation & Workflow Intelligence

Blue-Signal-Search • San Francisco (CA)

On-site
USD 130,000 - 160,000
Competitive compensation package
Significant equity upside
Collaborative in-person work environment
ML Platform Engineer
ML Platform Engineer

synthesia • United States

On-site
USD 100,000 - 140,000