Senior ML Platform Engineer - Scale ML Infra & Pipelines

Menlo Ventures

San Francisco (CA)

Hybrid

USD 187,000 - 259,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

4 days in office per week
Fridays from home
Medical, dental, vision benefits
401k match
Generous vacation and paid days off
Parental leave

Job summary

Chime’s Machine Learning Platform (MLP) team builds the infrastructure, tooling, and developer experience that powers machine learning across the company. You will design and build scalable systems that support model training, feature computation, real-time inference, and experimentation, enabling data scientists and ML engineers to move quickly with reliability and cost awareness.

This role focuses on building robust foundations that allow ML teams to move quickly while maintaining governance

Qualifications

  • 5+ years of experience in ML infrastructure, platform engineering, or production ML systems.
  • Knowledge of the ML model development lifecycle including data preprocessing, training, evaluation and deployment.
  • Experience with distributed systems, cloud computing, or large-scale data processing.
  • Hands-on experience with CI/CD pipelines, DevOps practices, and infrastructure as code.

Responsibilities

  • Design, build, and operate scalable ML infrastructure on AWS.
  • Develop distributed training and batch processing systems using Ray.
  • Build and maintain infrastructure-as-code using Terraform.
  • Support and evolve the feature store and feature pipelines.
  • Develop data ingestion and streaming systems (e.g., Kinesis, Kafka, Flink, Spark).
  • Improve CI/CD workflows for ML models and platform components.
  • Enhance observability, reliability, and cost visibility across ML workloads.
  • Partner with Data Science and ML Engineering teams to improve developer experience.
  • Contribute to platform architecture decisions and technical roadmaps.
  • Participate in on-call rotations to support production systems.

Skills

Python
CI/CD
Distributed systems
Software engineering fundamentals
DevOps

Tools

Docker
Kubernetes
Terraform
Spark
Ray
AWS

Job description

Chime’s Machine Learning Platform (MLP) team builds the infrastructure, tooling, and developer experience that powers machine learning across the company. You will design and build scalable systems that support model training, feature computation, real-time inference, and experimentation, enabling data scientists and ML engineers to move quickly with reliability and cost awareness.

This role focuses on building robust foundations that allow ML teams to move quickly while maintaining governance

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior ML Platform Engineer: Scale ML Infra on AWS
Senior ML Platform Engineer: Scale ML Infra on AWS

Chime • San Francisco (CA)

On-site
USD 187,000 - 259,000
Senior ML Platform & Infra Engineer - Scale AI Pipelines
Senior ML Platform & Infra Engineer - Scale AI Pipelines

Monograph • United States

Hybrid
USD 160,000 - 240,000
Competitive base pay
Equity (RSUs)
Benefits
Senior ML Platform Engineer — Scale AI Infrastructure
Senior ML Platform Engineer — Scale AI Infrastructure

Socket.dev • Denver (CO)

Hybrid
USD 160,000 - 240,000
ML Platform Engineer — Scale AI Pipelines
ML Platform Engineer — Scale AI Pipelines

Gusto, Inc. • Denver (CO)

Hybrid
USD 160,000 - 200,000
Equity (RSUs)
Benefits
Staff ML Platform Engineer: Scale ML Infra & MLOps (Hybrid)
Staff ML Platform Engineer: Scale ML Infra & MLOps (Hybrid)

Faire • California (MO)

Hybrid
USD 247,000 - 339,000
Equity
Benefits
Senior ML Infrastructure Engineer - Scale AI Platforms
Senior ML Infrastructure Engineer - Scale AI Platforms

TrulyHired • San Francisco (CA)

On-site
USD 180,000 - 260,000
Equity
401(k) match
Parental leave
+6
Principal ML Platform Engineer - Scalable, Reliable Systems
Principal ML Platform Engineer - Scalable, Reliable Systems

EngineersOfAI • Sunnyvale (CA)

On-site
USD 150,000 - 200,000
Platform Engineer - Scalable ML Inference
Platform Engineer - Scalable ML Inference

Menlo Ventures • San Francisco (CA)

On-site
USD 180,000 - 260,000
Competitive compensation
Ownership culture
World-class team
ML Platform Engineer: Scale AI Infrastructure
ML Platform Engineer: Scale AI Infrastructure

Gusto • San Francisco (CA)

Hybrid
USD 190,000 - 240,000
Lead ML Platform Engineer: Training & Inference at Scale
Lead ML Platform Engineer: Training & Inference at Scale

Paramount • Burbank (CA)

On-site
USD 157,000 - 235,000
Benefits package
On-site & virtual events
Generous PTO