TPU Systems Engineer — High-Performance ML Inference Equity

RadixArk

Palo Alto (CA)

On-site

USD 180,000 - 250,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Significant founding team equity
Comprehensive health benefits
Flexible work arrangements

Job summary

A leading AI infrastructure firm is seeking a TPU Systems Engineer to develop high-performance systems using JAX, XLA, and Pallas. This role involves pushing large-model workloads on TPU hardware and optimizing performance across the stack. Candidates should have at least 3 years of experience in production ML systems and a strong foundation in Python. Competitive compensation includes equity and comprehensive health benefits, with a US base salary range of $180,000 - $250,000.

Qualifications

  • 3+ years experience building production ML systems with TPU-focused frameworks.
  • Deep understanding of JAX/XLA internals: HLO, fusion, partitioning strategies.
  • Experience with distributed systems like SGLang and training frameworks.

Responsibilities

  • Build high-performance inference and training systems using JAX/XLA/Pallas.
  • Optimize latency and throughput for LLM serving on TPU infrastructure.
  • Collaborate with kernel engineers to achieve performance improvements.

Skills

Building production ML systems with JAX, XLA
Performance tuning across compiler and runtime layers
Proficiency in Python
Distributed inference systems
Kernel development skills

Education

Bachelor's or Master's degree in Computer Science or Electrical Engineering

Job description

A leading AI infrastructure firm is seeking a TPU Systems Engineer to develop high-performance systems using JAX, XLA, and Pallas. This role involves pushing large-model workloads on TPU hardware and optimizing performance across the stack. Candidates should have at least 3 years of experience in production ML systems and a strong foundation in Python. Competitive compensation includes equity and comprehensive health benefits, with a US base salary range of $180,000 - $250,000.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

TPU Kernel Engineer — Lead Low-Latency ML Kernels (Hybrid)
TPU Kernel Engineer — Lead Low-Latency ML Kernels (Hybrid)

Anthropic • New York (NY)

Hybrid
USD 280,000 - 850,000
TPU Kernel Architect
TPU Kernel Architect

Anthropic • Seattle (WA), New York (NY), San Francisco (CA)

On-site
USD 315,000 - 560,000
Competitive compensation
Generous vacation and parental leave
Flexible working hours
TPU Kernel Engineer for High-Performance ML Systems
TPU Kernel Engineer for High-Performance ML Systems

SignalAI • New York (NY)

Hybrid
USD 280,000 - 850,000
Member of Technical Staff — Inference-TPU
Member of Technical Staff — Inference-TPU

RadixArk • Palo Alto (CA)

On-site
USD 180,000 - 250,000
Significant founding team equity
Comprehensive health benefits
Flexible work arrangements
TPU Kernel Engineer
TPU Kernel Engineer

Anthropic • New York (NY)

Hybrid
USD 280,000 - 850,000
TPU Performance Engineer — ML Efficiency & Scale
TPU Performance Engineer — ML Efficiency & Scale

Google • Sunnyvale (CA)

On-site
USD 207,000 - 300,000
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Strativ Group • Palo Alto (CA)

On-site
USD 500,000 - 600,000
Hybrid TPU Kernel Engineer for High-Performance ML
Hybrid TPU Kernel Engineer for High-Performance ML

Neura Market • San Francisco (CA)

Hybrid
USD 280,000 - 850,000
Staff ML Compiler Engineer for TPU Acceleration
Staff ML Compiler Engineer for TPU Acceleration

Google • New York (NY)

On-site
USD 207,000 - 301,000
20% bonus target
Equity options
Comprehensive benefits
Staff TPU Performance Engineer: Optimize Large-Scale ML
Staff TPU Performance Engineer: Optimize Large-Scale ML

Google • Kirkland (WA)

On-site
USD 207,000 - 300,000
Health insurance
Dental, Vision, Life, Disability
401(k) with company match
+5