Remote MLOps Engineer for Scalable LLM Systems

Benture

California, Northern (MO, KY)

Hybrid

USD 124,000 - 165,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Benture is seeking an experienced MLOps Engineer, specializing in LLM systems, to join a frontier AI lab's GenAI team on a full-time basis. You will contribute to building foundational large language models and ML infrastructure, with a focus on performance, scalability, and data generation.

The role requires 2+ years in ML infrastructure, hands-on with JAX and PyTorch, CUDA/Triton, and experience in model serving and KPI-driven optimization.

Qualifications

  • 2+ years of hands-on professional experience in ML systems, ML infrastructure, model serving, or GPU/accelerator performance engineering.
  • Production experience with JAX and/or PyTorch; framework-level depth (custom operators, FSDP, DDP, DeepSpeed, Megatron) is a strong plus.
  • Familiarity with modern accelerators (A100, H100, TPU) and the ability to reason about throughput, latency, and memory trade-offs.
  • Practical experience in at least one of: custom GPU kernel development (CUDA, Triton, Pallas); performance profiling and trace analysis; debugging distributed or accelerator-bound workloads; or large-scale LLM serving.

Responsibilities

  • Design and solve challenging MLOps tasks across GPU kernels, performance profiling, debugging, and inference serving for high-quality training data.
  • Guide research and engineering teams to close knowledge gaps and improve AI model performance on ML systems and training infrastructure.
  • Evaluate MLOps tasks and solutions, providing clear, rigorous written technical feedback.
  • Develop rubrics and evaluation frameworks for kernel optimization, profiler output interpretation, and serving throughput/latency trade-offs.
  • Collaborate with experts to ensure consistency and accuracy across training datasets.

Skills

MLOps
GPU perf engineering
JAX
PyTorch
Distributed systems
LLM serving
Written communication

Tools

CUDA
Triton
Pallas
Kineto
torch.profiler
Nsight
XLA/JAX profiler
DeepSpeed
Megatron
Ray Serve
TensorRT-LLM

Job description

Benture is seeking an experienced MLOps Engineer, specializing in LLM systems, to join a frontier AI lab's GenAI team on a full-time basis. You will contribute to building foundational large language models and ML infrastructure, with a focus on performance, scalability, and data generation.

The role requires 2+ years in ML infrastructure, hands-on with JAX and PyTorch, CUDA/Triton, and experience in model serving and KPI-driven optimization.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

MLOps Engineer, LLM Systems · Mercor Mercor · 90-120/hr · remote in US, UK, CA · 2w ago 90-120/hr remote in US, UK, CA 2w ago
MLOps Engineer, LLM Systems · Mercor Mercor · 90-120/hr · remote in US, UK, CA · 2w ago 90-120/hr remote in US, UK, CA 2w ago

Benture • California (MO), Northern (KY)

Hybrid
USD 124,000 - 165,000
Staff ML Systems Engineer - LLM Serving & RL
Staff ML Systems Engineer - LLM Serving & RL

Prime Intellect • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 300,000
Remote option
Visa sponsorship
Relocation support
+2
Remote MLOps Engineer - Build Scalable AI Serving
Remote MLOps Engineer - Build Scalable AI Serving

Bright Vision Technologies • United States

Remote
USD 100,000 - 150,000
MLOps Engineer — Scalable AI Infra & Deployment, Equity
MLOps Engineer — Scalable AI Infra & Deployment, Equity

Fundamental • United States

Remote
USD 180,000 - 260,000
Salary + equity
Health coverage for you and dependents
Parental leave for all
+2
Remote MLOps Engineer: Scalable ML Systems & GPU Kernels
Remote MLOps Engineer: Scalable ML Systems & GPU Kernels

Remote Jobs • United States

Remote
USD 96,000 - 152,000
Remote LLM Engineer: Fine-Tuning & Production ML
Remote LLM Engineer: Fine-Tuning & Production ML

Bright Vision Technologies • United States

Remote
USD 72,000 - 100,000
Senior AI Engineer: Scalable LLMs & MLOps Leader
Senior AI Engineer: Scalable LLMs & MLOps Leader

Compunnel, Inc. • San Francisco (CA)

On-site
USD 120,000 - 160,000
Remote AI Engineer - ML/LLM & Production Systems
Remote AI Engineer - ML/LLM & Production Systems

Zohorecruit • United States

Remote
USD 110,000 - 160,000
On-Site Senior MLOps Engineer: Real-Time LLM Infra
On-Site Senior MLOps Engineer: Real-Time LLM Infra

Nace.AI • Palo Alto (CA)

On-site
USD 210,000 - 280,000
Senior MLOps Engineer - Remote AI/ML & LLM Deployment
Senior MLOps Engineer - Remote AI/ML & LLM Deployment

Fractal • United States

On-site
USD 130,000 - 145,000
Health, dental, vision insurance
Life and disability insurance
401(k) plan