ML Performance Engineer – Real-Time Inference

Odyssey

Palo Alto (CA)

On-site

USD 130,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading AI lab in California seeks an experienced software engineer specialized in machine learning performance. Responsibilities include optimizing models used by extensive user bases, developing distributed training strategies, and collaborating closely with researchers to enhance model architectures. Ideal candidates should have over 8 years of software engineering experience, a strong background in ML optimization, and proficiency with tools like PyTorch and NVIDIA ecosystems. Join our innovative team to push the limits of AI technology.

Qualifications

  • 8+ years of software engineering experience with focus on ML performance.
  • Strong understanding of modern machine learning architectures.
  • Ability to solve problems and adapt to new skills as necessary.

Responsibilities

  • Optimize models for real-time usage by hundreds of thousands of users.
  • Design and implement distributed training strategies.
  • Partner with ML researchers to ensure model performance.
  • Develop tools to identify performance bottlenecks.
  • Pioneer innovative approaches to enhance performance metrics.

Skills

Performance optimization
Distributed training
Problem-solving
Deep learning architectures
PyTorch
NVIDIA GPU optimization

Job description

A leading AI lab in California seeks an experienced software engineer specialized in machine learning performance. Responsibilities include optimizing models used by extensive user bases, developing distributed training strategies, and collaborating closely with researchers to enhance model architectures. Ideal candidates should have over 8 years of software engineering experience, a strong background in ML optimization, and proficiency with tools like PyTorch and NVIDIA ecosystems. Join our innovative team to push the limits of AI technology.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Model Performance Engineer - Inference and Acceleration
ML Model Performance Engineer - Inference and Acceleration

Baseten • New York (NY)

On-site
USD 200,000 - 275,000
Senior ML Performance Engineer - Distributed Training
Senior ML Performance Engineer - Distributed Training

Odyssey • Santa Clara (CA)

On-site
USD 120,000 - 160,000
Software Engineer - ML Model Performance
Software Engineer - ML Model Performance

Baseten • San Francisco (CA)

On-site
USD 150,000 - 250,000
ML Performance Engineer — Scalable DL Pipelines & Optimization
ML Performance Engineer — Scalable DL Pipelines & Optimization

Optiver US LLC • New York (NY)

On-site
USD 160,000 - 260,000
Competitive compensation package
Global profit-sharing pool
401(k) match up to 50%
+2
Senior ML Engineer - Real-Time Inference & Systems
Senior ML Engineer - Real-Time Inference & Systems

Inworld AI • Germany (OH)

On-site
USD 120,000 - 180,000
Staff ML Performance Engineer — Scalable Inference & CUDA
Staff ML Performance Engineer — Scalable Inference & CUDA

Modal • New York (NY)

On-site
USD 120,000 - 160,000
Senior ML Performance Engineer: LLM Benchmarking & GPU
Senior ML Performance Engineer: LLM Benchmarking & GPU

Amadeus Search • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Competitive salary
Equity and bonus opportunities
Medical, dental, and vision coverage
+2
Senior ML Engineer - Low-Latency Inference & Systems
Senior ML Engineer - Low-Latency Inference & Systems

Inworld • Germany (OH)

Hybrid
USD 120,000 - 160,000
Staff ML Performance Engineer: Scale Training Throughput
Staff ML Performance Engineer: Scale Training Throughput

Wayve • Sunnyvale (CA)

On-site
USD 130,000 - 160,000
Senior ML Engineer: AI Inference & Performance Optimizer
Senior ML Engineer: AI Inference & Performance Optimizer

Nebius • Palo Alto (CA)

Hybrid
USD 195,000 - 263,000
Health insurance
401(k) plan
Parental leave
+2