Founding ML Infra Engineer — Equity & Open-Source AI

Peano AI

Palo Alto (CA)

On-site

USD 180,000 - 260,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Peano AI is building the Agent Cloud infrastructure and seeks a Founding ML Infrastructure Engineer to design and optimize high‑performance model serving for large‑scale open‑source models in Palo Alto.

You will own end‑to‑end ML systems, work with accelerators like TPUs/GPUs, improve latency and throughput, and collaborate with research and product teams to ship production‑grade infrastructure.

Qualifications

  • Strong experience in ML systems, distributed systems, or high-performance computing.
  • Experience optimizing inference or training workloads for large models.
  • Familiarity with TPUs, GPUs, or other accelerators.
  • Experience with CUDA, Triton, NCCL, JAX/XLA, PyTorch internals, vLLM, SGLang, or TensorRT‑LLM.
  • Strong systems debugging skills.
  • Comfort working across model code, runtime, infrastructure, and product requirements.
  • High ownership in an early-stage startup environment.

Responsibilities

  • Optimize large-scale LLM inference and serving systems.
  • Improve tokens per second, latency, throughput, and cost efficiency.
  • Work on serving infrastructure for open-source models across accelerators.
  • Improve batching, scheduling, KV cache management, memory usage, and accelerator utilization.
  • Support long-context inference up to 1M context.
  • Debug performance bottlenecks across model execution, runtime, networking, and infrastructure.
  • Work with frameworks such as JAX/XLA, PyTorch, vLLM, SGLang, TensorRT-LLM, or related systems.
  • Collaborate with the application team to optimize infrastructure for agentic workloads.
  • Help turn research prototypes into reliable production systems.

Skills

ML systems
Distributed systems
High-performance computing
Inference optimization
Training optimization
TPUs
GPUs

Tools

CUDA
Triton
NCCL
JAX/XLA
PyTorch internals
vLLM
SGLang
TensorRT-LLM

Job description

Peano AI is building the Agent Cloud infrastructure and seeks a Founding ML Infrastructure Engineer to design and optimize high‑performance model serving for large‑scale open‑source models in Palo Alto.

You will own end‑to‑end ML systems, work with accelerators like TPUs/GPUs, improve latency and throughput, and collaborate with research and product teams to ship production‑grade infrastructure.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Founding Machine Learning Infrastructure Engineer
Founding Machine Learning Infrastructure Engineer

Peano AI • Palo Alto (CA)

On-site
USD 180,000 - 260,000
Founding ML Systems Engineer – High-Performance AI Infra
Founding ML Systems Engineer – High-Performance AI Infra

Strativ Group • Palo Alto (CA)

On-site
USD 450,000 - 550,000
Founding equity
Direct exposure to founders
Competitive compensation
Machine Learning Systems Engineer
Machine Learning Systems Engineer

Strativ Group • Palo Alto (CA)

On-site
USD 450,000 - 550,000
Founding equity
Direct exposure to founders
Competitive compensation
Founding ML Systems Engineer - End-to-End Infrastructure AI
Founding ML Systems Engineer - End-to-End Infrastructure AI

Meter • San Francisco (CA)

Hybrid
USD 200,000 - 320,000
Founding Staff AI Infra Engineer — ML Platform
Founding Staff AI Infra Engineer — ML Platform

Anduril Industries, Inc. • Costa Mesa (CA), Northern (KY)

Hybrid
USD 220,000 - 292,000
Founding ML Infra Architect
Founding ML Infra Architect

uRun • San Francisco (CA)

On-site
USD 120,000 - 160,000
Health, dental, and vision
401(k)
Paid time off
+2
Systems ML Engineer - High-Scale AI Infra (Equity)
Systems ML Engineer - High-Scale AI Infra (Equity)

Meta • Concord (NH)

On-site
USD 154,000 - 217,000
Senior ML Platform Engineer — Scale Research ML Infra
Senior ML Platform Engineer — Scale Research ML Infra

techire ai • San Francisco (CA)

On-site
USD 270,000 - 330,000
Stock options
Staff ML Infrastructure Architect
Staff ML Infrastructure Architect

Cacheflow • Palo Alto (CA)

On-site
USD 218,000 - 285,000
Principal ML Engineer - AI Security & Scalable Infra
Principal ML Engineer - AI Security & Scalable Infra

Palo Alto Networks • California (MO)

Hybrid
USD 156,000 - 253,000