ML Systems Engineer - Production Training & GPU Tuning

Park Lane Recruitment

New York (NY)

On-site

USD 250,000 - 350,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Relocation support
Sponsorship available
Elite total compensation
401(k) matching
Medical and prescription coverage
Wellness reimbursement
Family building support
Charitable gift matching

Job summary

Park Lane Recruitment is seeking an experienced ML/Infrastructure engineer in New York to build and optimize production ML systems, including training infrastructure and GPU-accelerated workloads.

The role emphasizes collaboration with researchers, rapid prototyping, memory tuning, and durable production deployments. Relocation sponsorship is available, with base salary of $250,000 to $350,000 and a guaranteed first-year bonus plus comprehensive benefits.

Qualifications

  • 3 to 15 years of experience in ML engineering or distributed/infrastructure engineering.
  • Hands-on ability with Python, PyTorch, JAX, CUDA and GPU computing tools.
  • Experience building ML systems that moved to production with training infrastructure and optimization.

Responsibilities

  • We build and optimize machine learning systems that move from research ideas into production use.
  • We write custom GPU code and tune memory usage to reduce training time and remove performance bottlenecks.
  • We design and improve training infrastructure so researchers can run more experiments faster without destabilizing systems or inflating compute costs.
  • We develop first‑pass implementations of promising new techniques and validate them against real data.
  • We help transition prototypes into durable production systems that can run continuously over time.
  • We collaborate closely with researchers to turn conceptual architecture ideas into functioning systems.
  • We influence technical direction by evaluating better approaches and building proof‑of‑concept solutions.
  • We work across a small, senior team where everyone contributes directly and ships meaningful work.

Skills

Research collaboration
Translating ideas to systems

Education

Background in CS/Math/Physics/EE or quantitative discipline

Tools

Python
PyTorch
JAX
CUDA

Job description

Park Lane Recruitment is seeking an experienced ML/Infrastructure engineer in New York to build and optimize production ML systems, including training infrastructure and GPU-accelerated workloads.

The role emphasizes collaboration with researchers, rapid prototyping, memory tuning, and durable production deployments. Relocation sponsorship is available, with base salary of $250,000 to $350,000 and a guaranteed first-year bonus plus comprehensive benefits.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior ML Systems Engineer – Finance (Production, NYC)
Senior ML Systems Engineer – Finance (Production, NYC)

Park Lane Recruitment • New York (NY)

On-site
USD 250,000 - 350,000
401(k) matching
Medical coverage
Wellness reimbursement
+2
Applied ML Systems Engineer - Finance
Applied ML Systems Engineer - Finance

Park Lane Recruitment • New York (NY)

On-site
USD 250,000 - 350,000
Relocation support
Sponsorship available
Elite total compensation
+5
Applied ML Systems Engineer – Finance
Applied ML Systems Engineer – Finance

Park Lane Recruitment • New York (NY)

On-site
USD 250,000 - 350,000
401(k) matching
Medical coverage
Wellness reimbursement
+2
ML Infrastructure Engineer: Build Scalable GPU Clusters
ML Infrastructure Engineer: Build Scalable GPU Clusters

cursor • New York (NY), San Francisco (CA)

On-site
USD 120,000 - 150,000
ML Systems Engineer: Scalable Training & Realtime Inference
ML Systems Engineer: Scalable Training & Realtime Inference

Jobzhr • New York (NY)

On-site
USD 180,000 - 280,000
Production ML Engineer – AI Systems in NYC
Production ML Engineer – AI Systems in NYC

Weekday (YC W21) • New York (NY)

On-site
USD 150,000 - 250,000
Equity/bonus
Health benefits
Paid time off
+2
Staff ML Engineer — Production-Grade AI Systems
Staff ML Engineer — Production-Grade AI Systems

Bjak • New York (NY)

On-site
USD 100,000 - 140,000
Trading ML Systems Engineer - Distributed Training
Trading ML Systems Engineer - Distributed Training

Fintal Partners • New York (NY)

On-site
USD 185,000 - 230,000
Distributed ML Training Engineer - Scale GPUs, Unlimited PTO
Distributed ML Training Engineer - Scale GPUs, Unlimited PTO

Thinking Machines Lab Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 350,000 - 475,000
Health benefits
Unlimited PTO
Parental leave
+1
LLM Systems Engineer & Research
LLM Systems Engineer & Research

Scale AI, Inc. • New York (NY)

On-site
USD 189,000 - 237,000
Comprehensive health coverage
Dental and vision coverage
Retirement benefits
+3