Senior ML Platform Engineer – Distributed GPU & Game AI

WorkGenius Group

Santa Monica (CA)

On-site

USD 124,000 - 207,000

Full time

7 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

WorkGenius Group in Santa Monica, CA seeks a Senior Machine Learning Engineer to build and scale distributed ML training systems across multi-node GPU clusters and to integrate platform components for games.

You will develop automated pipelines, inference serving optimized for latency and cost, and create tools to accelerate the ML lifecycle, while collaborating with game teams and mentoring engineers.

Qualifications

  • Bachelor’s degree in Computer Science or related field or equivalent experience with 3+ years in ML/platform/systems engineering.
  • Experience operating distributed systems and production GPU workloads including distributed ML or agent training with RL/IL.

Responsibilities

  • Build and operate distributed ML training systems across multi-node GPU clusters.
  • Develop automated ML pipelines with data/model versioning, validation, and experiment tracking.
  • Build and optimize inference-serving systems for latency, cost, and reliability.
  • Create developer tools, SDKs, and templates to streamline ML lifecycle.
  • Monitor system health, troubleshoot production issues, and mentor engineers.

Skills

Distributed ML
GPU Clusters
Reinforcement Learning
Imitation Learning
PyTorch
Python
C++
Ray/RLlib
Inference Serving
ML Platform Engineering

Education

Bachelor’s degree in Computer Science or related field

Tools

PyTorch
Ray
RLlib
C++
Python
Multi-node GPU orchestration

Job description

WorkGenius Group in Santa Monica, CA seeks a Senior Machine Learning Engineer to build and scale distributed ML training systems across multi-node GPU clusters and to integrate platform components for games.

You will develop automated pipelines, inference serving optimized for latency and cost, and create tools to accelerate the ML lifecycle, while collaborating with game teams and mentoring engineers.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Machine Learning Engineer - Platform and Integrations
Senior Machine Learning Engineer - Platform and Integrations

WorkGenius Group • Santa Monica (CA)

On-site
USD 124,000 - 207,000
Senior ML Training Systems Engineer - Distributed CUDA
Senior ML Training Systems Engineer - Distributed CUDA

Genesis AI • San Francisco (CA)

On-site
USD 180,000 - 260,000
Senior ML Platform Engineer: Scalable GPU & Production ML
Senior ML Platform Engineer: Scalable GPU & Production ML

adobe • San Jose (CA)

On-site
USD 183,000 - 265,000
Senior ML Infra Engineer for Distributed GPU Training
Senior ML Infra Engineer for Distributed GPU Training

Genesis Molecular AI • City of Utica (NY)

On-site
USD 150,000 - 190,000
Competitive compensation with salary +
Senior ML Platform Engineer: Distributed GPU Training Infra
Senior ML Platform Engineer: Distributed GPU Training Infra

Menlo Ventures • Seattle (WA)

On-site
USD 189,000 - 245,000
Senior ML Infra Engineer - Scale GPU Clusters, Remote
Senior ML Infra Engineer - Scale GPU Clusters, Remote

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 320,000 - 500,000
Equity
Medical/Dental/Vision coverage
Unlimited PTO
+1
Senior ML Platform Engineer - Remote, GPU & Cloud Scale
Senior ML Platform Engineer - Remote, GPU & Cloud Scale

General Motors • Sunnyvale (CA), Northern (KY)

Hybrid
USD 155,000 - 396,000
Medical
Dental
Vision
+9
Senior ML Platform Engineer (Remote)
Senior ML Platform Engineer (Remote)

Sony Interactive Entertainment • San Diego (CA)

Hybrid
USD 187,000 - 266,000
Medical insurance
Dental insurance
Matching 401(k)
+2
Senior ML Infra Engineer: GPU-Optimized Kubernetes Platform
Senior ML Infra Engineer: GPU-Optimized Kubernetes Platform

Hamilton Barnes Associates Limited • San Francisco (CA)

On-site
USD 225,000 - 275,000
Stock options
Lead ML Systems Engineer — Distributed GPU Training & Infra
Lead ML Systems Engineer — Distributed GPU Training & Infra

Nvidia Corporation • Santa Clara (CA)

On-site
USD 224,000 - 431,000
Equity
Benefits