Senior ML Infrastructure Engineer - Large-Scale GPU Serving

Selby Jennings

New York (NY)

On-site

USD 180,000 - 250,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Selby Jennings is building a new machine learning infrastructure initiative and seeks an experienced engineer to own the design and deployment of large-scale model serving platforms. You will influence architecture decisions and work closely with researchers and engineers to push state-of-the-art models into production on GPU farms.

The role emphasizes ownership from day one, with opportunities to shape the platform roadmap and optimize hardware utilization across growing GPU clusters in a

Qualifications

  • 8+ years in software, systems or ML infra.
  • Deep experience with model serving and optimization at scale.
  • Strong knowledge of CUDA, kernel optimization and GPU acceleration.
  • Expert programming in C++ and Python.

Responsibilities

  • Design and build large-scale serving and inference infra for state-of-the-art ML models.
  • Optimize performance, latency and GPU utilization in production.
  • Deploy and scale models across thousands of GPUs; improve efficiency and reliability.
  • Develop compression, quantization, distillation techniques to maximize hardware performance.
  • Move cutting-edge models from research to production with hardware integration.
  • Collaborate with researchers and engineers to speed experimentation and deployment.
  • Shape architecture and roadmap for a rapidly growing ML platform.

Skills

Large-scale infra
CUDA GPU
C++
Python
PyTorch
JAX
TensorFlow
Model serving
Inference optimization
GPU acceleration

Education

BS/MS/PhD in CS/Engineering

Tools

CUDA
Linux

Job description

Selby Jennings is building a new machine learning infrastructure initiative and seeks an experienced engineer to own the design and deployment of large-scale model serving platforms. You will influence architecture decisions and work closely with researchers and engineers to push state-of-the-art models into production on GPU farms.

The role emphasizes ownership from day one, with opportunities to shape the platform roadmap and optimize hardware utilization across growing GPU clusters in a

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior ML Infrastructure Engineer - GPU & Scale
Senior ML Infrastructure Engineer - GPU & Scale

TensorWave • Las Vegas (NV)

On-site
USD 120,000 - 150,000
Competitive Salary
Stock Options
100% paid Medical, Dental, and Vision insurance
+8
Senior ML Infra Engineer - Scale GPU Clusters, Remote
Senior ML Infra Engineer - Scale GPU Clusters, Remote

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 320,000 - 500,000
Equity
Medical/Dental/Vision coverage
Unlimited PTO
+1
Senior ML Engineer, Serving & Optimization
Senior ML Engineer, Serving & Optimization

Selby Jennings • New York (NY)

On-site
USD 180,000 - 250,000
Remote Senior ML Systems Engineer - Model Serving
Remote Senior ML Systems Engineer - Model Serving

Atlassian • Northern (KY)

Hybrid
USD 149,000 - 235,000
Benefits and bonuses
Equity
ML Platform Engineer — Infra for Research on GPU Fleets
ML Platform Engineer — Infra for Research on GPU Fleets

cursor • New York (NY), San Francisco (CA)

On-site
USD 120,000 - 180,000
Senior ML Infra Engineer - Scalable Model Serving (Remote)
Senior ML Infra Engineer - Scalable Model Serving (Remote)

Etsy, Inc. • New York (NY)

Hybrid
USD 182,000 - 246,000
Equity package
Annual bonus
Comprehensive benefits
Senior ML Infra Architect — Large-Scale GPU Training
Senior ML Infra Architect — Large-Scale GPU Training

Hark, Inc. • San Jose (CA)

On-site
USD 180,000 - 450,000
ML Infrastructure Engineer: Build Scalable GPU Clusters
ML Infrastructure Engineer: Build Scalable GPU Clusters

cursor • New York (NY), San Francisco (CA)

On-site
USD 120,000 - 150,000
Senior ML Model Serving Engineer (Remote)
Senior ML Model Serving Engineer (Remote)

NEPSE Trading • Northern (KY)

Hybrid
USD 74,000 - 98,000
Senior AI Infrastructure Architect (GPU & Serving)
Senior AI Infrastructure Architect (GPU & Serving)

Makers Fund • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Equity
Health benefits
Monthly stipends
+1