ML Systems Engineer — Production-Scale LLM Inference

ChipAgents

San Jose (CA)

On-site

USD 150,000 - 350,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Unlimited PTO
Full benefits (medical, vision, dental, 401k)
Free parking and private gym

Job summary

ChipAgents is looking for a skilled ML Systems Engineer in San Jose, California. In this technical role, you will optimize large language model inference for our agentic AI platform, impacting chip design efficiency.

Your responsibilities include implementing performance optimizations, designing multi-node clusters, and collaborating with top semiconductor firms. With a competitive salary range of $150K–$350K, we also offer substantial benefits and equity options.

Qualifications

  • B.S., M.S., or PhD in Computer Science, Electrical Engineering, or related field.
  • Experience with large‑scale ML systems, GPU computing, or high‑performance inference optimization.
  • Strong proficiency in Python and C++/CUDA.

Responsibilities

  • Design, deploy, and optimize LLM inference systems across multi‑node clusters.
  • Implement and benchmark inference optimizations.
  • Profile and analyze inference bottlenecks at systems level.

Skills

Large-scale ML systems
GPU computing
High-performance inference optimization
Python
C++/CUDA
SGLang
vLLM
PyTorch
Debugging and profiling

Education

B.S., M.S., or PhD in Computer Science, Electrical Engineering, or related field

Tools

Distributed computing frameworks (Ray)

Job description

ChipAgents is looking for a skilled ML Systems Engineer in San Jose, California. In this technical role, you will optimize large language model inference for our agentic AI platform, impacting chip design efficiency.

Your responsibilities include implementing performance optimizations, designing multi-node clusters, and collaborating with top semiconductor firms. With a competitive salary range of $150K–$350K, we also offer substantial benefits and equity options.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Systems Engineer
ML Systems Engineer

ChipAgents • San Jose (CA)

On-site
USD 150,000 - 350,000
Unlimited PTO
Full benefits (medical, vision, dental, 401k)
Free parking and private gym
ML Systems Engineer: Distributed LLM Training & Inference
ML Systems Engineer: Distributed LLM Training & Inference

Scale AI • Seattle (WA), New York (NY), San Francisco (CA)

On-site
USD 200,000 - 251,000
Comprehensive health coverage
Equity-based compensation
Retirement benefits
+3
ML Systems Architect for Production-Grade AI
ML Systems Architect for Production-Grade AI

A1 • Palo Alto (CA)

On-site
USD 347,000 - 490,000
ML Engineer — Production-Grade AI & LLM/VLM Systems
ML Engineer — Production-Grade AI & LLM/VLM Systems

Nace.AI • Palo Alto (CA)

On-site
USD 120,000 - 160,000
ML Inference Systems Engineer
ML Inference Systems Engineer

Gimlet Labs, Inc. • San Francisco (CA)

On-site
USD 120,000 - 160,000
ML Engineer: Production AI & LLM Systems
ML Engineer: Production AI & LLM Systems

CodeGeniusRecruit • United States

On-site
USD 413,000 - 551,000
ML Systems Engineer for RL & Inference Infrastructure
ML Systems Engineer for RL & Inference Infrastructure

Advanced Micro Devices • Santa Clara (CA)

Hybrid
USD 160,000 - 210,000
AMD benefits
Head of ML Systems & Inference
Head of ML Systems & Inference

Doist • San Francisco (CA)

On-site
USD 260,000 - 380,000
Health, dental, vision benefits
401(k) company match
ML Systems Engineer: Scale Training & Inference
ML Systems Engineer: Scale Training & Inference

Doist • San Francisco (CA)

On-site
USD 180,000 - 230,000
Competitive cash compensation
Startup equity
Senior ML Engineer - Low-Latency Inference & Systems
Senior ML Engineer - Low-Latency Inference & Systems

Inworld • Germany (OH)

Hybrid
USD 120,000 - 160,000