Staff Engineer - AI Inference & Benchmarking

Liquid-Ai

Boston (MA)

Hybrid

USD 140,000 - 210,000

Full time

6 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Health insurance
401(k) match
Unlimited PTO

Job summary

Liquid AI in Boston is seeking a Member Of Technical Staff, Infrastructure to own the inference engine layer and benchmarking infrastructure, collaborating with research and product teams and external partners. You will port models across runtimes, extend the llama.cpp/ONNX/MLX stack, design benchmarks, verify outputs, and help drive performance, reliability, and scalability across deployment targets.

This role blends systems work with hands-on verification of numerical correctness and

Qualifications

  • Hands-on experience with inference frameworks going beyond basic usage.
  • Experience designing and building benchmarking pipelines including methodology and reproducibility.
  • Strong C++ and Python in performance-sensitive contexts.
  • Solid understanding of inference fundamentals: quantization, decoding strategies, memory layout.

Responsibilities

  • Design and build benchmark suites that cover inference performance, model quality, and knowledge evaluation across different hardware targets.
  • Run external partner verifications: evaluate their solutions against our benchmarks, identify gaps, and clearly deliver findings.
  • Port models onto different runtimes and frameworks, and verify correctness end-to-end.
  • Maintain and extend the inference engine layer built on llama.cpp, ONNX, and MLX as new model architectures emerge from research.
  • Make benchmark results explainable and verifiable, so internal teams and partners can trust and reproduce them independently.

Skills

C++
Python

Tools

llama.cpp
ONNX Runtime
MLX

Job description

Liquid AI in Boston is seeking a Member Of Technical Staff, Infrastructure to own the inference engine layer and benchmarking infrastructure, collaborating with research and product teams and external partners. You will port models across runtimes, extend the llama.cpp/ONNX/MLX stack, design benchmarks, verify outputs, and help drive performance, reliability, and scalability across deployment targets.

This role blends systems work with hands-on verification of numerical correctness and

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Engineer, AI Inference & Benchmarking
Staff Engineer, AI Inference & Benchmarking

Liquid-Ai • Cambridge (MA)

Hybrid
USD 180,000 - 240,000
Competitive base salary with equity
Health premiums paid (medical, dental,
401(k) matching up to 4%
+1
Member of Technical Staff - Inference Systems
Member of Technical Staff - Inference Systems

Liquid-Ai • Boston (MA)

Hybrid
USD 140,000 - 210,000
Equity
Health insurance
401(k) match
+1
Member of Technical Staff - Inference Systems
Member of Technical Staff - Inference Systems

Liquid-Ai • Cambridge (MA)

Hybrid
USD 180,000 - 240,000
Competitive base salary with equity
Health premiums paid (medical, dental,
401(k) matching up to 4%
+1
Staff Engineer, AI Inference & Benchmarking
Staff Engineer, AI Inference & Benchmarking

S27a • San Francisco (CA), New York (NY)

Hybrid
USD 150,000 - 210,000
Staff, AI Inference Benchmarking (Equity, Onsite SF)
Staff, AI Inference Benchmarking (Equity, Onsite SF)

Artificial Analysis, Inc. • San Francisco (CA)

On-site
USD 120,000 - 180,000
Equity
Staff Engineer, AI Benchmarking & Inference Systems
Staff Engineer, AI Benchmarking & Inference Systems

S27a • San Francisco (CA), New York (NY)

Hybrid
USD 140,000 - 190,000
Generous PTO
Office stipend
Competitive healthcare (medical,Dental
+2
ML Infrastructure Engineer: Training & Inference
ML Infrastructure Engineer: Training & Inference

Physical Superintelligence • Boston (MA)

Hybrid
USD 140,000 - 210,000
Member of Technical Staff - ML Systems & Inference
Member of Technical Staff - ML Systems & Inference

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 280,000
Staff Engineer - Customer-Facing AI Inference Infra
Staff Engineer - Customer-Facing AI Inference Infra

Simplify • San Francisco (CA)

On-site
USD 200,000 - 300,000
Housing stipend
Uber/Waymo rides
Staff Engineer – High-Performance Model Inference
Staff Engineer – High-Performance Model Inference

Pantera Capital • Palo Alto (CA)

On-site
USD 180,000 - 440,000
Equity
Medical coverage
Vision coverage
+5