ML Performance Engineer: GPU/CUDA at Scale

Selby Jennings

Chicago (IL)

On-site

USD 140,000 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Selby Jennings in Chicago seeks a Performance Engineer to develop scalable ML infrastructure for training and inference at scale. You will focus on low-level optimization and high-performance data pipelines, enabling rapid experimentation and model deployment.

You should bring strong CUDA/C++ and Python skills, plus hands-on experience with PyTorch, TensorFlow, or JAX. Join a team at the intersection of finance and technology, driving innovation.

Qualifications

  • BS or MS in CS or related field; strong fundamentals in algorithms and systems.
  • Proven experience in GPU-accelerated ML workloads and optimization.
  • Experience with deep learning frameworks and performance profiling.

Responsibilities

  • Build scalable pipelines for deep learning training and inference.
  • Optimize and extend ML frameworks for performance and functionality.
  • Collaborate with researchers to accelerate experimentation and deployment.

Skills

CUDA
C++
Python
PyTorch
TensorFlow
JAX

Education

B.S. or M.S. in Computer Science

Job description

Selby Jennings in Chicago seeks a Performance Engineer to develop scalable ML infrastructure for training and inference at scale. You will focus on low-level optimization and high-performance data pipelines, enabling rapid experimentation and model deployment.

You should bring strong CUDA/C++ and Python skills, plus hands-on experience with PyTorch, TensorFlow, or JAX. Join a team at the intersection of finance and technology, driving innovation.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Performance Engineer: Scale GPU-Driven Training
ML Performance Engineer: Scale GPU-Driven Training

Decisive Point • Sunnyvale (CA)

On-site
USD 180,000 - 240,000
Machine Learning Performance Engineer
Machine Learning Performance Engineer

Selby Jennings • Chicago (IL)

On-site
USD 140,000 - 210,000
ML Performance Engineer: Low-Level Systems & GPUs
ML Performance Engineer: Low-Level Systems & GPUs

Trading Interview • New York (NY)

On-site
USD 170,000 - 210,000
ML Performance Engineer: Scale Training & Throughput
ML Performance Engineer: Scale Training & Throughput

Applied Intuition • Sunnyvale (CA)

On-site
USD 180,000 - 240,000
ML Performance Engineer — Scale Distributed Training & Throughput
ML Performance Engineer — Scale Distributed Training & Throughput

applied • Sunnyvale (CA)

On-site
USD 170,000 - 230,000
GPU Performance Engineer: Scale ML Inference & Systems
GPU Performance Engineer: Scale ML Inference & Systems

Anthropic • New York (NY)

Hybrid
USD 280,000 - 850,000
ML Performance Engineer — Scalable DL Pipelines & Optimization
ML Performance Engineer — Scalable DL Pipelines & Optimization

Optiver US LLC • New York (NY)

On-site
USD 160,000 - 260,000
Competitive compensation package
Global profit-sharing pool
401(k) match up to 50%
+2
Senior ML Performance Engineer
Senior ML Performance Engineer

well-funded deeptech startup • California (MO)

On-site
USD 200,000 - 250,000
Senior ML Performance Engineer: Scale & Throughput
Senior ML Performance Engineer: Scale & Throughput

NLP PEOPLE • Sunnyvale (CA)

On-site
USD 215,000 - 285,000
Machine Learning Performance Engineer
Machine Learning Performance Engineer

Trading Interview • New York (NY)

On-site
USD 170,000 - 210,000