ML Serving Engineer – Real-Time Voice AI Platform

Unusual Ventures

San Francisco (CA)

On-site

USD 180,000 - 240,000

Full time

6 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Rime is hiring a Software Engineer to own the serving infrastructure that connects inference engines to production. You will work at the intersection of ML systems and cloud infrastructure, building, hardening, and scaling real-time voice model serving.

You'll own the TTS serving stack, optimize multi-node serving, and ensure cross-hardware compatibility from HPC GPUs to cloud deployments. You will shape architecture, CI/CD, and reliability while collaborating with ML teams.

Qualifications

  • Hands-on experience with real-time multinode ML serving infrastructure.
  • Experience with distributed or disaggregated model serving.
  • Strong cloud infrastructure fundamentals and IaC tooling.

Responsibilities

  • Architect and implement the TTS serving infrastructure across GPUs and API surface.
  • Optimize models from single-node to disaggregated fleet serving.
  • Ensure compatibility across NVIDIA hardware for on-prem and cloud deployments.
  • Maintain CI/CD workflows for the model serving pipeline.
  • Manage site reliability: on-call, monitoring, and observability.
  • Plan and manage GPU resource provisioning and costs.

Skills

Real-time ML serving
NVIDIA Dynamo/Triton
vLLM
SGLang
Distributed model serving
Cloud infrastructure
Linux
Docker
Kubernetes
Terraform

Tools

Docker
Kubernetes
Terraform
Packer

Job description

Rime is hiring a Software Engineer to own the serving infrastructure that connects inference engines to production. You will work at the intersection of ML systems and cloud infrastructure, building, hardening, and scaling real-time voice model serving.

You'll own the TTS serving stack, optimize multi-node serving, and ensure cross-hardware compatibility from HPC GPUs to cloud deployments. You will shape architecture, CI/CD, and reliability while collaborating with ML teams.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer, ML Serving
Software Engineer, ML Serving

Unusual Ventures • San Francisco (CA)

On-site
USD 180,000 - 240,000
Speech ML Scientist - TTS & Voice AI (Remote)
Speech ML Scientist - TTS & Voice AI (Remote)

Rime • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Remote-friendly
Visa sponsorship available
Equity upside
+1
Staff ML Engineer — Real-Time Voice Inference Architect
Staff ML Engineer — Real-Time Voice Inference Architect

Together • San Francisco (CA)

On-site
USD 220,000 - 280,000
Staff ML Engineer, Voice AI - Real-Time Inference Lead
Staff ML Engineer, Voice AI - Real-Time Inference Lead

Together AI • San Francisco (CA)

On-site
USD 220,000 - 280,000
Health insurance
Startup equity
Competitive benefits
Principal ML Engineer — Real-Time TTS & High-Performance GPUs
Principal ML Engineer — Real-Time TTS & High-Performance GPUs

Inworld AI • Mountain View (CA), Northern (KY)

Hybrid
USD 270,000 - 500,000
Relocation assistance
Edge ML Engineer: Real-Time Speech & Language for Autonomy
Edge ML Engineer: Real-Time Speech & Language for Autonomy

Instant Teams • New Haven (CT)

On-site
USD 150,000 - 215,000
Equity stake
100% Remote (US)
Comprehensive Health Coverage
+3
Research Engineer: Real-Time Voice ML & Production Systems
Research Engineer: Real-Time Voice ML & Production Systems

Sesame • Bellevue (WA)

On-site
USD 190,000 - 320,000
401(k) max match
Health, vision & dental benefits
Unlimited PTO"
+3
Hands-On Forward-Deployed Voice AI Engineer
Hands-On Forward-Deployed Voice AI Engineer

Rime Labs • San Francisco (CA)

Hybrid
USD 185,000 - 235,000
Competitive compensation
Equity options
Exposure to cutting-edge AI technology
Mountain View, California, USA Staff / Principal Platform Engineer - USA
Mountain View, California, USA Staff / Principal Platform Engineer - USA

Inworld AI • Mountain View (CA)

On-site
USD 150,000 - 210,000
Remote Senior AI Voice Engineer: Real-Time LLM & ML
Remote Senior AI Voice Engineer: Real-Time LLM & ML

amchealth • Annapolis (MD)

On-site
USD 140,000 - 175,000