Staff Distributed Systems Engineer - AI Infrastructure

Acceler8 Talent

San Francisco, Northern (CA, KY)

Hybrid

USD 180,000 - 260,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Acceler8 Talent in San Francisco is seeking a Member of Technical Staff to lead distributed systems at the core of our AI infrastructure. You’ll design, build, and operate scheduling, routing, and coordination for AI workloads across CPUs, GPUs, and accelerators.

You’ll work closely with inference, runtimes, compilers, kernels, and hardware teams to push the limits of scale and reliability. This on-site role offers meaningful ownership from day one in a fast-growing startup.

Qualifications

  • Strong distributed systems experience in production.
  • Deep understanding of concurrency, failure modes, scalability, and system trade-offs.
  • Experience with Go, C++, or Python.
  • Exposure to scheduling, queues, RPC, control planes, or Kubernetes internals.
  • Strong evidence of personally building, debugging, or optimising complex systems.

Responsibilities

  • Build distributed scheduling, orchestration, and control-plane systems.
  • Improve reliability, fault tolerance, and workload placement across large-scale infrastructure.
  • Develop RPC, asynchronous messaging, and resource-management systems.
  • Build production APIs that abstract hardware complexity from customers.
  • Work on Kubernetes-adjacent infrastructure beyond simply operating clusters.
  • Collaborate across inference, runtimes, compilers, kernels, and hardware.

Skills

Distributed systems
Go
C++
Python
RPC
Kubernetes

Job description

Acceler8 Talent in San Francisco is seeking a Member of Technical Staff to lead distributed systems at the core of our AI infrastructure. You’ll design, build, and operate scheduling, routing, and coordination for AI workloads across CPUs, GPUs, and accelerators.

You’ll work closely with inference, runtimes, compilers, kernels, and hardware teams to push the limits of scale and reliability. This on-site role offers meaningful ownership from day one in a fast-growing startup.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Engineer, Distributed AI Inference Systems (Equity)
Staff Engineer, Distributed AI Inference Systems (Equity)

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 350,000
Equity
AI Inference Orchestration - Distributed Systems Engineer
AI Inference Orchestration - Distributed Systems Engineer

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 350,000
Member of Technical Staff - Distributed Systems
Member of Technical Staff - Distributed Systems

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 260,000
Member of Technical Staff- Distributed Systems
Member of Technical Staff- Distributed Systems

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 350,000
Equity
Staff Engineer, Distributed AI Systems & Scheduling
Staff Engineer, Distributed AI Systems & Scheduling

Gimlet Labs • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior ML Systems Engineer — Inference & Scale
Senior ML Systems Engineer — Inference & Scale

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 170,000 - 240,000
Software Engineer
Software Engineer

Acceler8 Talent • San Francisco (CA), Northern (KY)

On-site
USD 150,000 - 350,000
Staff Distributed Systems Engineer — AI Infra, Open Source
Staff Distributed Systems Engineer — AI Infra, Open Source

TrustIn • San Francisco (CA)

On-site
USD 400,000 - 500,000
Staff Engineer: AI Infrastructure & Scale Leader
Staff Engineer: AI Infrastructure & Scale Leader

Recruiting From Scratch • San Francisco (CA)

On-site
USD 250,000 - 300,000
Competitive equity
Ownership in infra
Direct collaboration with founders
+1
Member of Technical Staff
Member of Technical Staff

Acceler8 Talent • San Francisco (CA), Northern (KY)

On-site
USD 150,000 - 350,000