ML Infrastructure Engineer

Acceler8 Talent

San Francisco (CA)

Hybrid

USD 247,500 - 302,500

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

A forward-thinking AI company is seeking a Senior ML Infrastructure / Backend Engineer to develop backend systems supporting real-time AI functionality. You will manage APIs, scale cloud infrastructure, and ensure system reliability while collaborating with various teams. The ideal candidate should possess strong backend development skills and experience with cloud platforms, shaping the future of AI-powered digital interactions. The salary is competitive, reaching up to $275k depending on qualifications, with a hybrid work model in Los Angeles or San Francisco.

Qualifications

  • Strong experience building and owning backend or distributed systems.
  • Hands-on experience designing APIs, preferably using Python.
  • Experience running ML-backed systems in production environments.

Responsibilities

  • Own backend services and APIs that expose ML-powered features.
  • Design and operate orchestration layers for ML workloads.
  • Scale infrastructure to support high user traffic.

Skills

Backend services development
API design and operation (Python preferred)
Cloud platforms (AWS, GCP)
Performance tuning
Debugging and operational ownership

Tools

FastAPI
Flask
gRPC
Ray Serve
Triton

Job description

Series C Startup | AI-Powered 3D & Avatar Platform | Hybrid (LA or SF)

We’re hiring a Senior ML Infrastructure / Backend Engineer to join a well-funded AI company building the visual and interaction layer for the next generation of AI-powered digital identities.

This team is developing production systems that bring AI characters out of chat boxes and into real-time, interactive 3D experiences.

You’ll own backend and infrastructure systems that serve ML-powered functionality at scale — supporting high-concurrency user traffic, low-latency inference, and rapid iteration as the platform grows.

You’ll work closely with ML researchers, platform engineers, and product teams to take models from experimentation to reliable, scalable production services.

What You’ll Do:

  • Own backend services and APIs that expose ML-powered features to real users
  • Design and operate orchestration layers for ML workloads (routing, batching, retries, concurrency)
  • Deploy and scale ML-backed services in cloud environments
  • Scale infrastructure to support thousands to hundreds of thousands of daily requests
  • Implement observability, monitoring, and alerting to ensure system reliability
  • Partner closely with ML teams to productionize generative and ML models
  • Improve end-to-end efficiency across inference, post-processing, and data pipelines

What We’re Looking For:

  • Strong, production-level experience building and owning backend or distributed systems
  • Hands-on experience designing and operating APIs (Python preferred — FastAPI, Flask, or gRPC)
  • Experience deploying and running ML-backed systems in production environments
  • Proven ability to scale systems under real user traffic with attention to latency and reliability
  • Experience with cloud platforms (AWS, GCP, or similar) and containerized deployments
  • Strong debugging, performance tuning, and operational ownership skills

Nice to have:

  • Experience with ML inference optimization (quantization, mixed precision, ONNX, TensorRT)
  • Familiarity with scalable inference frameworks (Ray Serve, Triton, TorchServe, SageMaker)
  • Exposure to generative models (diffusion or transformer-based systems)
  • Experience running GPU-backed or high-performance workloads in production

This Role Is:

  • Hybrid (Los Angeles or San Francisco office)
  • Salary range: upto $275k base, depending on experience

If you’re excited about building the infrastructure that powers real-time, embodied AI — and want ownership over systems that actually ship — we’d love to talk.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

ML Infrastructure Engineer
ML Infrastructure Engineer

Lattice, Inc. • San Francisco (CA)

On-site
USD 200,000 - 280,000
Competitive salary
Premium health, dental, and vision insurance
Unlimited PTO
+2
ML Infrastructure Engineer
ML Infrastructure Engineer

Objective Partners • San Francisco (CA)

On-site
USD 180,000 - 250,000
Full medical, dental, vision coverage
Flexible PTO
Daily catered lunches
+1
ML / AI Engineer
ML / AI Engineer

Wintermeyer Ventures • San Francisco (CA)

On-site
USD 200,000 - 300,000
Equity
Senior ML Platform Engineer
Senior ML Platform Engineer

techire ai • San Francisco (CA)

On-site
USD 270,000 - 330,000
Stock options
Member of Technical Staff
Member of Technical Staff

kadence • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior Software Engineer - ML Infrastructure
Senior Software Engineer - ML Infrastructure

Claryo • San Francisco (CA), Northern (KY)

On-site
USD 180,000 - 240,000
Medical/Dental/Vision
401k with employer matching
Parental leave
+1
Machine Learning Infrastructure Engineer
Machine Learning Infrastructure Engineer

Alexander Chapman • New York (NY)

On-site
USD 120,000 - 160,000
Equity
Health insurance
Dental & Vision
Lead ML Platform Engineer
Lead ML Platform Engineer

Harnham • New York (NY)

On-site
USD 150,000 - 190,000
Competitive base salary and annual be
Equity participation through RSUs
Opportunity to work on cutting-edge AI
+2
Member of Technical Staff
Member of Technical Staff

Harrison Clarke • San Francisco (CA)

On-site
USD 180,000 - 280,000
Software Engineer (Backend-Focused) $140,000 - $225,000 Posted 12 hours ago
Software Engineer (Backend-Focused) $140,000 - $225,000 Posted 12 hours ago

Fuel Talent LLC • Northern (KY)

On-site
USD 140,000 - 225,000