Senior Machine Learning Infrastructure Engineer

Morph

California (MO)

On-site

USD 90,000 - 120,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Morph is looking for candidates with exceptional skills in infrastructure to manage high-uptime systems focused on GPU loads. The role requires deep knowledge of inference engines and experience with various infrastructures, including load balancers.

Your contributions to projects like vLLM or SGLang will be a valuable asset. Join us in tackling complex challenges in custom inference stacks.

Qualifications

  • Deep experience with inference engines is essential.
  • Prior contributions to vLLM or SGLang are a plus.
  • Experience with traditional infrastructure like load balancers.

Responsibilities

  • Design and manage systems to achieve high uptime.
  • Serve custom inference stacks with irregular GPU loads.
  • Work with various infrastructure types to meet challenging requirements.

Skills

Infrastructure design
GPU load management
Load balancers
Contribution to inference frameworks

Job description

Goal: 99.99% uptime

We serve custom inference stacks that have irregular GPU load.

We're looking for people that have done genuinely amazing work in infrastructure and are interested in a challenge, working with both traditional infrastructure such as load balancers, NLB, etc., as well as very different infrastructure around inference engines and GPU loads.

This is a role that will inherently require deep experience with inference engines.

Contributions to vLLM, SGLang, trtllm, or inference frameworks a plus.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Inference Infrastructure Architect
ML Inference Infrastructure Architect

Morph • California (MO)

On-site
USD 90,000 - 120,000
Infrastructure Engineer, LLM Inference Optimization
Infrastructure Engineer, LLM Inference Optimization

GMI Cloud • Mountain View (CA)

On-site
USD 170,000 - 230,000
Staff Software Engineer: LLM Inference & GPU Infra (Equity)
Staff Software Engineer: LLM Inference & GPU Infra (Equity)

DevHub • San Francisco (CA)

On-site
USD 180,000 - 240,000
Annual performance bonus
Equity
Software Engineer: ML Infra
Software Engineer: ML Infra

Generalist • Somerville (MA), San Mateo (CA)

On-site
USD 120,000 - 160,000
AI Infrastructure Engineer - Inference Platform
AI Infrastructure Engineer - Inference Platform

Hoonify Technologies Inc. • Albuquerque (NM)

On-site
USD 120,000 - 190,000
Staff ML Infra Engineer: Distributed Training & Inference
Staff ML Infra Engineer: Distributed Training & Inference

Jobtailor • Boston (MA)

On-site
USD 120,000 - 160,000
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 170,000 - 240,000
Member of the Technical Staff- LLMs
Member of the Technical Staff- LLMs

Amadeus Search • San Francisco (CA)

Hybrid
USD 170,000 - 220,000
Staff Software Engineer- Foundation Model Inference
Staff Software Engineer- Foundation Model Inference

DevHub • San Francisco (CA)

On-site
USD 180,000 - 240,000
Annual performance bonus
Equity
Member of Technical Staff (Software Engineer, Inference & Training Platform)
Member of Technical Staff (Software Engineer, Inference & Training Platform)

United States Digital Space LLC • San Francisco (CA)

Hybrid
USD 180,000 - 240,000