AI Systems Architect – Inference & RL at Scale

Togetherai

San Francisco (CA)

On-site

USD 200,000 - 280,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Startup equity
Health insurance
Competitive benefits

Job summary

Togetherai is seeking a qualified engineer to join their Turbo team in San Francisco. The role focuses on advancing inference efficiency and operating Reinforcement Learning pipelines for ML systems. A candidate should have a strong background in systems and algorithms, with experience in Python and distributed systems.

With a commitment to innovation, Togetherai offers competitive compensation, equity, health benefits, and opportunities for growth in a dynamic work environment.

Qualifications

  • 3+ years of experience working on ML systems or equivalent experience.
  • Experience owning complex technical projects end‑to‑end.
  • Demonstrated coding ability in Python.

Responsibilities

  • Design and prototype algorithms for low-latency, high-throughput inference.
  • Profile and optimize performance across multiple layers.
  • Design and operate RL pipelines to enhance inference efficiency.

Skills

Large-scale inference systems
Reinforcement Learning (RL)
Model architecture design
Distributed systems
Python coding

Education

Advanced degree in Computer Science or related field

Tools

GPU performance profiling
TensorRT
SGLang
vLLM

Job description

Togetherai is seeking a qualified engineer to join their Turbo team in San Francisco. The role focuses on advancing inference efficiency and operating Reinforcement Learning pipelines for ML systems. A candidate should have a strong background in systems and algorithms, with experience in Python and distributed systems.

With a commitment to innovation, Togetherai offers competitive compensation, equity, health benefits, and opportunities for growth in a dynamic work environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff AI Systems Engineer — Inference & RL
Staff AI Systems Engineer — Inference & RL

Together • San Francisco (CA)

On-site
USD 200,000 - 280,000
Health insurance
Startup equity
Competitive benefits
RL Systems Engineer: Inference & Training at Scale
RL Systems Engineer: Inference & Training at Scale

xAI • Palo Alto (CA)

On-site
USD 180,000 - 240,000
AI Systems Engineer — RL Environments & Scalable Infra
AI Systems Engineer — RL Environments & Scalable Infra

AI Talent Now • San Francisco (CA)

On-site
USD 120,000 - 150,000
AI Researcher, Core ML (Turbo)
AI Researcher, Core ML (Turbo)

Together • San Francisco (CA)

On-site
USD 200,000 - 280,000
Health insurance
Startup equity
Competitive benefits
ML Systems Engineer – End-to-End Inference & RL
ML Systems Engineer – End-to-End Inference & RL

Togetherai • San Francisco (CA)

On-site
USD 200,000 - 280,000
Startup equity
Health insurance
Competitive benefits
Senior AI Inference Engineer — Scale LLMs & PyTorch
Senior AI Inference Engineer — Scale LLMs & PyTorch

Togetherai • San Francisco (CA)

On-site
USD 160,000 - 230,000
Health insurance
Startup equity
Competitive benefits
Machine Learning Engineer - Inference
Machine Learning Engineer - Inference

Togetherai • San Francisco (CA)

On-site
USD 160,000 - 230,000
Health insurance
Startup equity
Competitive benefits
Senior Backend Engineer, AI Inference Platform
Senior Backend Engineer, AI Inference Platform

Togetherai • San Francisco (CA)

On-site
USD 160,000 - 250,000
Health insurance
Startup equity
Competitive benefits
AI Researcher, Core ML (Turbo)
AI Researcher, Core ML (Turbo)

Togetherai • San Francisco (CA)

On-site
USD 200,000 - 280,000
Startup equity
Health insurance
Competitive benefits
Research Engineer (Core ML)
Research Engineer (Core ML)

Together AI • San Francisco (CA)

On-site
USD 200,000 - 280,000
Startup equity
Health insurance
Competitive benefits