Member of Technical Staff: AI Inference Compiler Engineer

Gimlet Labs

San Francisco (CA)

On-site

USD 190,000 - 260,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Gimlet Labs is building the first multi-silicon neocloud designed for fast, efficient AI inference. You will shape compiler infrastructure, optimize how workloads are represented and executed across diverse architectures, and influence scheduling, memory movement, and kernel orchestration for production-scale inference.

This ML-systems–hardware hybrid role focuses on delivering low latency and high throughput by partitioning workloads across devices and by enabling new accelerator architectures

Qualifications

  • Experience building compiler, runtime, or execution infrastructure.
  • Experience with IR transformations, compiler passes, lowering, or code generation.
  • Strong systems and performance-engineering fundamentals.
  • Ability to reason about execution behavior, memory systems, scheduling, and hardware efficiency.
  • Strong C++ and/or Python skills.
  • A bachelor's degree in a relevant field or equivalent practical experience.

Responsibilities

  • Improve latency, throughput, and efficiency of production inference workloads.
  • Design execution strategies for partitioning workloads across heterogeneous hardware.
  • Develop compiler optimizations spanning IR transformations, scheduling, memory movement, and kernel orchestration.
  • Enable new models, accelerator architectures, and serving techniques to run efficiently in production.

Skills

C++
Python
Compiler infra

Education

Bachelor's degree in a relevant field

Tools

MLIR
LLVM
XLA
TVM
Triton

Job description

Gimlet Labs is building the first multi-silicon neocloud designed for fast, efficient AI inference. You will shape compiler infrastructure, optimize how workloads are represented and executed across diverse architectures, and influence scheduling, memory movement, and kernel orchestration for production-scale inference.

This ML-systems–hardware hybrid role focuses on delivering low latency and high throughput by partitioning workloads across devices and by enabling new accelerator architectures

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Member of Technical Staff - Compiler Engineer
Member of Technical Staff - Compiler Engineer

Gimlet Labs • San Francisco (CA)

On-site
USD 190,000 - 260,000
Member of Technical Staff - ML Systems & Inference
Member of Technical Staff - ML Systems & Inference

Gimlet Labs • San Francisco (CA)

On-site
USD 120,000 - 160,000
Member of Technical Staff - Applied AI Research
Member of Technical Staff - Applied AI Research

Gimlet Labs, Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Staff Compiler Engineer for AI Systems & Heterogeneous HW
Staff Compiler Engineer for AI Systems & Heterogeneous HW

The Consensus • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 230,000
Member of Technical Staff - ML Systems & Inference
Member of Technical Staff - ML Systems & Inference

Gimlet Labs, Inc. • San Francisco (CA)

On-site
USD 120,000 - 160,000
Member of Technical Staff - Compilers
Member of Technical Staff - Compilers

The Consensus • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 230,000
Senior AI Compiler Engineer - MLIR, GPU Inference, Equity
Senior AI Compiler Engineer - MLIR, GPU Inference, Equity

NVIDIA • Town of Texas (WI)

On-site
USD 152,000 - 288,000
Equity
Benefits package
Lead AI Graph Compiler Engineer
Lead AI Graph Compiler Engineer

EnCharge AI • United States

Remote
USD 190,000 - 255,000
Senior MLIR AI Compiler Engineer for GPU Inference
Senior MLIR AI Compiler Engineer for GPU Inference

NVIDIA • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior AI Compiler Engineer - MLIR & GPU Optimizations
Senior AI Compiler Engineer - MLIR & GPU Optimizations

NVIDIA • Washington

On-site
USD 152,000 - 288,000
Equity
Benefits package