ML Performance Engineer — Scalable DL Pipelines & Optimization

Optiver US LLC

New York (NY)

On-site

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive compensation package
Global profit-sharing pool
401(k) match up to 50%
Comprehensive health, dental, vision coverage
Extensive office perks, including food and social events

Job summary

A global trading firm is seeking a Machine Learning Performance Engineer to focus on AI infrastructure. This role involves building training and inference pipelines, enhancing deep learning frameworks, and tackling performance bottlenecks. Candidates must possess strong GPU programming knowledge with CUDA and expertise in frameworks like PyTorch or TensorFlow. Competitive compensation and extensive benefits, including a profit-sharing pool, are offered in this high-performing environment.

Qualifications

  • Strong knowledge of low-level GPU programming with CUDA, including Tensor Cores.
  • Expertise in internals of deep-learning frameworks like PyTorch, JAX, TensorFlow.
  • Deep understanding of computer architecture.
  • Experience in C++ and Python.

Responsibilities

  • Build scalable and robust training and inference pipelines for deep learning.
  • Enhance functionality of open-source deep learning frameworks.
  • Identify and eliminate performance bottlenecks.
  • Collaborate closely with researchers and other engineers.
  • Develop deep understanding of trading systems.

Skills

Low-level GPU programming with CUDA
Deep learning frameworks expertise (PyTorch, JAX, TensorFlow)
Understanding of computer architecture
Experience in C++
Experience in Python

Tools

CUDA
JAX ecosystem
GPU libraries (Triton, CUB, cuDNN, cuBLAS)

Job description

A global trading firm is seeking a Machine Learning Performance Engineer to focus on AI infrastructure. This role involves building training and inference pipelines, enhancing deep learning frameworks, and tackling performance bottlenecks. Candidates must possess strong GPU programming knowledge with CUDA and expertise in frameworks like PyTorch or TensorFlow. Competitive compensation and extensive benefits, including a profit-sharing pool, are offered in this high-performing environment.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Performance Engineer: Scalable DL Pipelines
ML Performance Engineer: Scalable DL Pipelines

Optiver • New York (NY)

On-site
USD 200,000
Global profit-sharing pool
401(k) match up to 50%
Comprehensive health coverage
+2
Senior ML Performance Engineer: LLM Benchmarking & GPU
Senior ML Performance Engineer: LLM Benchmarking & GPU

Amadeus Search • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Competitive salary
Equity and bonus opportunities
Medical, dental, and vision coverage
+2
ML Performance Engineer – Real-Time Inference
ML Performance Engineer – Real-Time Inference

Odyssey • Palo Alto (CA)

On-site
USD 130,000 - 160,000
Machine Learning Performance Engineer
Machine Learning Performance Engineer

Selby Jennings • Chicago (IL)

On-site
USD 140,000 - 210,000
Staff ML Performance Engineer — Scalable Inference & CUDA
Staff ML Performance Engineer — Scalable Inference & CUDA

Modal • New York (NY)

On-site
USD 120,000 - 160,000
Staff ML Performance Engineer: Scale Training Throughput
Staff ML Performance Engineer: Scale Training Throughput

Wayve • Sunnyvale (CA)

On-site
USD 130,000 - 160,000
Performance Engineer, Large-Scale ML Systems
Performance Engineer, Large-Scale ML Systems

Anthropic • New York (NY)

Hybrid
USD 280,000 - 850,000
Competitive compensation
Optional equity donation matching
Generous vacation and parental leave
+1
ML Performance Engineer: Scale GPU-Driven Training
ML Performance Engineer: Scale GPU-Driven Training

Decisive Point • Sunnyvale (CA)

On-site
USD 180,000 - 240,000
Machine Learning Performance Engineer - Quant Research & Trading
Machine Learning Performance Engineer - Quant Research & Trading

Acquire Me • United States

On-site
USD 200,000 - 350,000
ML Performance Engineer: GPU/CUDA at Scale
ML Performance Engineer: GPU/CUDA at Scale

Selby Jennings • Chicago (IL)

On-site
USD 140,000 - 210,000