Senior ML Ops Engineer: Productionize LLMs at Scale

voltai-com

Edison (CA)

On-site

USD 150,000 - 230,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Unlimited PTO
Comprehensive Health Coverage
Free Meals and Snacks
Professional Growth
Visa Sponsorship

Job summary

Voltai is seeking engineers to design, build, and operate scalable ML pipelines for training, evaluation, and deployment of LLMs and retrieval-augmented systems. You will automate the ML lifecycle, monitor model quality, and optimize performance across diverse hardware and cloud environments.

Join a team blending software and hardware expertise to productionize foundation models for enterprise customers, while balancing latency, cost, and security in fast-paced settings.

Qualifications

  • Experience building reliable, scalable ML systems.
  • End-to-end ML lifecycle management experience.
  • Ability to translate research into production-ready systems.

Responsibilities

  • Design, build, and maintain scalable ML pipelines for training, evaluation, and deployment of LLMs.
  • Operationalize evaluation workflows using synthetic and human-labeled data across deployments.
  • Automate the ML Developer lifecycle with data versioning and CI/CD pipelines.
  • Optimize model training and inference for latency, throughput, and cost.
  • Collaborate cross-functionally to productionize foundation models.
  • Deploy and manage both open-source and proprietary models with latency, security, and compliance constraints.
  • Implement real-time monitoring to detect drift and bottlenecks in live systems.
  • Work directly with enterprise customers on deployment strategies and feedback loops.

Skills

Python
Go
Rust
ML Ops
System design
Cross-functional
Performance

Tools

MLflow
Kubeflow
SageMaker
Vertex AI
Apache Airflow
Docker
Kubernetes
Pulumi
Terraform
Weights & Biases
CometML

Job description

Voltai is seeking engineers to design, build, and operate scalable ML pipelines for training, evaluation, and deployment of LLMs and retrieval-augmented systems. You will automate the ML lifecycle, monitor model quality, and optimize performance across diverse hardware and cloud environments.

Join a team blending software and hardware expertise to productionize foundation models for enterprise customers, while balancing latency, cost, and security in fast-paced settings.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

ML Systems Engineer: GPU-Accelerated Training & Inference
ML Systems Engineer: GPU-Accelerated Training & Inference

voltai-com • Edison (CA)

On-site
USD 180,000 - 260,000
Unlimited PTO
Comprehensive Health Coverage
Free Meals and Snacks
+2
Senior ML Engineer - End-to-End LLMs & Production
Senior ML Engineer - End-to-End LLMs & Production

Distil Labs GmbH • United States

Remote
USD 180,000 - 240,000
Remote work within European time zones
Periodic in-person team offsites
Staff ML Ops Engineer: Equity & Platform Lead
Staff ML Ops Engineer: Equity & Platform Lead

LVT (LiveView Technologies) • Seattle (WA)

On-site
USD 213,000 - 272,000
Health, dental, and vision coverage
401k with match
Flexible PTO
Senior ML Research Scientist — Multimodal & LLMs (PTO, Visa)
Senior ML Research Scientist — Multimodal & LLMs (PTO, Visa)

voltai-com • Edison (CA)

On-site
USD 180,000 - 260,000
Unlimited PTO
Comprehensive Health Coverage
Free Meals and Snacks
+2
MLOps Engineer — Production ML Systems & Deployment
MLOps Engineer — Production ML Systems & Deployment

Speria • Dunwoody (GA)

On-site
USD 120,000 - 170,000
MLOps Engineer
MLOps Engineer

InfoVision Inc. • Irving (TX)

On-site
USD 100,000 - 130,000
ML DevOps Architect: Cloud & Large-Scale Compute (Remote)
ML DevOps Architect: Cloud & Large-Scale Compute (Remote)

Ignite Next GmbH • Palo Alto (CA), Northern (KY)

Hybrid
USD 140,000 - 200,000
Remote work
Office visits in Palo Alto, Paris, orW
Flexible relocation options
ML Engineer – Multimodal AI & LLM Production
ML Engineer – Multimodal AI & LLM Production

Voltai • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Production ML Systems Engineer - Scale & Observability
Production ML Systems Engineer - Scale & Observability

Careervitablr • United States

Remote
USD 140,000 - 210,000
Senior MLOps Engineer: End-to-End Production ML
Senior MLOps Engineer: End-to-End Production ML

Insight Global • United States

On-site
USD 194,848,000 - 214,906,000
Medical, dental, and vision insurance
HSA/FSA/DCFSA accounts
401k with employer matching
+1