Senior GPU Systems Engineer: Large-Scale Inference & RL

Reflection

New York (NY)

On-site

USD 150,000 - 200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Comprehensive medical, dental, vision, life, and disability insurance
Fully paid parental leave
Daily meals provided
Relocation support
Paid time off

Job summary

Reflection seeks a skilled professional in New York to design and operate large-scale GPU infrastructure for model inference and reinforcement learning. The role demands several years of experience in deploying GPU systems, optimizing model performance, and working with frameworks like SGLang and Megatron. The position offers competitive compensation, comprehensive benefits, and a supportive work environment with daily meals and team activities.

Qualifications

  • Several years of hands-on experience building and running production infrastructure.
  • Strong understanding of GPU performance characteristics and optimization techniques.
  • Experience deploying large-scale GPU systems for inference or model serving.

Responsibilities

  • Design and operate large-scale GPU infrastructure for model inference.
  • Develop systems for synthetic data generation and reinforcement learning pipelines.
  • Optimize GPU utilization for large language model inference.

Skills

GPU system deployment
Model serving
Optimizing throughput
Distributed reinforcement learning
Performance optimization techniques

Tools

SGLang
Megatron

Job description

Reflection seeks a skilled professional in New York to design and operate large-scale GPU infrastructure for model inference and reinforcement learning. The role demands several years of experience in deploying GPU systems, optimizing model performance, and working with frameworks like SGLang and Megatron. The position offers competitive compensation, comprehensive benefits, and a supportive work environment with daily meals and team activities.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Engineer - Large-Scale GPU Inference & RL Infra
Staff Engineer - Large-Scale GPU Inference & RL Infra

Visa Hunt • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Top-tier compensation
Stock options
Health & wellness benefits
+5
Senior GPU ML Infra Engineer — Mid-Training & Inference
Senior GPU ML Infra Engineer — Mid-Training & Inference

Reflection AI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior Inference Systems Engineer — Large-Scale GPUs
Senior Inference Systems Engineer — Large-Scale GPUs

RadixArk • Palo Alto (CA)

On-site
USD 190,000 - 260,000
Competitive compensation
Meaningful equity
Comprehensive benefits
+1
Staff AI Systems Engineer - Pre-Training Infra
Staff AI Systems Engineer - Pre-Training Infra

Reflection • New York (NY)

On-site
USD 130,000 - 180,000
Member of Technical Staff - Mid-Training Infra
Member of Technical Staff - Mid-Training Infra

Reflection • New York (NY)

On-site
USD 150,000 - 200,000
Comprehensive medical, dental, vision, life, and disability insurance
Fully paid parental leave
Daily meals provided
+2
Senior ML Training Systems Engineer - Distributed GPU Infra
Senior ML Training Systems Engineer - Distributed GPU Infra

Baseten • San Francisco (CA)

On-site
USD 150,000 - 200,000
Competitive compensation, including equity
100% coverage of medical, dental, and vision insurance
Generous PTO policy
+2
Staff Engineer, GPU AI Inference & RL Infrastructure
Staff Engineer, GPU AI Inference & RL Infrastructure

B Capital • San Francisco (CA)

On-site
USD 120,000 - 160,000
Top-tier compensation
Comprehensive medical, dental, and vision insurance
Fully paid parental leave
+2
Senior GPU Systems Engineer – Large-Scale AI & HPC
Senior GPU Systems Engineer – Large-Scale AI & HPC

Iceberg • New York (NY)

On-site
USD 200,000 - 300,000
Senior System Software Engineer - GPU AI Inference Equity
Senior System Software Engineer - GPU AI Inference Equity

NVIDIA • United States

On-site
USD 152,000 - 242,000
Distributed RL Systems Engineer — Scale Training & Inference
Distributed RL Systems Engineer — Scale Training & Inference

Luma AI • United States

Remote
USD 180,000 - 240,000