Senior ML Platform Engineer - Production-Ready AI Systems

Invoca

Austin (TX)

Remote

USD 152,000 - 228,000

Full time

3 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Flexible Time Off
Health benefits
Stock options
401(k) match

Job summary

Invoca is hiring a Senior ML Engineer to own the production ML stack, turning trained models into reliable, high-throughput services used by product teams. You’ll partner with Data Scientists and Applied AI Engineers to ship AI-powered features at scale.

You will own inference infrastructure, optimize latency and throughput, and drive end-to-end ML pipelines, serving models via APIs and monitoring to ensure reliability and speed.

Qualifications

  • 5+ years of ML Engineering with production focus.
  • Advanced Python and DL proficiency (PyTorch, Transformers, spaCy).
  • Experience deploying transformer-based NLP models in production.
  • Fine-tuning SLMs/LLMs (LoRA, QLoRA, PEFT) with optimizations.
  • Proficiency with inference infra: Triton, Baseten, vLLM, TGI, SageMaker or Vertex AI.
  • Building production-grade ML APIs for downstream consumers.
  • Familiarity with MLOps tooling (Braintrust, MLflow).
  • Experience in agentic/multi-step AI workflows is a plus.
  • Bachelor’s in CS/Engineering; advanced degree a plus.
  • Familiarity with RLHF is a bonus.

Responsibilities

  • Architect, implement, and maintain CI/CD pipelines for ML artifacts.
  • Serve as SME for operational excellence across ML stack (uptime, latency, cost).
  • Treat model serving as a product for internal customers and ensure shipping capability.

Skills

ML Engineering
Python & DL
Transformer NLP in production
Fine-tuning SLMs/LLMs
Inference performance
APIs for ML models
MLOps tools
Agentic AI workflows
B.S. in CS/Engineering
RLHF familiarity

Education

B.S. in Computer Science/Engineering

Tools

Triton Inference Server
Baseten
vLLM
TGI
SageMaker
Vertex AI
Braintrust
MLflow

Job description

Invoca is hiring a Senior ML Engineer to own the production ML stack, turning trained models into reliable, high-throughput services used by product teams. You’ll partner with Data Scientists and Applied AI Engineers to ship AI-powered features at scale.

You will own inference infrastructure, optimize latency and throughput, and drive end-to-end ML pipelines, serving models via APIs and monitoring to ensure reliability and speed.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior ML Platform Engineer - Production Systems & LLMs
Senior ML Platform Engineer - Production Systems & LLMs

Invoca • New York (NY)

Remote
USD 152,000 - 228,000
Flexible Time Off
Health Benefits
Remote-First Senior ML Platform Engineer (Production AI)
Remote-First Senior ML Platform Engineer (Production AI)

Invoca • Denver (CO)

Remote
USD 152,000 - 228,000
Stock options
401(k) plan with company match
Health, dental, vision benefits
+2
Senior ML Platform Engineer - Production Systems (Remote)
Senior ML Platform Engineer - Production Systems (Remote)

Invoca • San Francisco (CA)

Remote
USD 152,000 - 228,000
Flexible Time Off
Paid Holidays
Health Benefits
+8
Senior ML Platform Engineer - Production Systems (Remote)
Senior ML Platform Engineer - Production Systems (Remote)

Invoca • San Francisco (CA)

Remote
USD 152,000 - 228,000
Flexible Time Off
Paid Holidays
Health Benefits
+8
Senior ML Engineer — Production ML & Inference (Remote)
Senior ML Engineer — Production ML & Inference (Remote)

Invoca • Chicago (IL)

Remote
USD 152,000 - 228,000
Flexible Time Off
Paid Holidays
Health Benefits
+6
Staff ML Platform Engineer - Production AI Infra (Remote)
Staff ML Platform Engineer - Production AI Infra (Remote)

creditacceptance • United States

Hybrid
USD 154,000 - 226,000
401(K) match
Adoption assistance
Parental leave
+2
Senior ML Inference Platform Engineer
Senior ML Inference Platform Engineer

Atlassian • Austin (TX)

Hybrid
USD 206,000 - 269,000
Health and wellbeing resources
Paid volunteer days
Senior ML Inference Engineer — Scale Production APIs
Senior ML Inference Engineer — Scale Production APIs

AssemblyAI, Inc. • New York (NY)

On-site
USD 190,000 - 225,000
Senior ML Engineer: Production AI & Deployment Lead
Senior ML Engineer: Production AI & Deployment Lead

Candid • United States

On-site
USD 130,000 - 160,000
Health insurance
Retirement plan with match
PSLF eligible employer
+1
Senior ML Platform Engineer - Production & MLOps
Senior ML Platform Engineer - Production & MLOps

Attain • Chicago (IL), Northern (KY)

Hybrid
USD 170,000 - 240,000