Inference Platform Engineer for Open-Source Models

Together AI

San Francisco (CA)

On-site

USD 200,000 - 290,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Health insurance
Startup equity

Job summary

Together AI in San Francisco is seeking a Research Engineer to help build a platform that lets users customize open-source models, bridging post-training and production inference. You will contribute across Fine-Tuning, RL, and Evaluation services and collaborate with product, research, and engineering teams to keep the API reliable and scalable.

This full-time role offers a competitive US salary range and equity, with focus on building foundational AI infrastructure for developers worldwide.

Qualifications

  • 2+ years of experience building and deploying machine learning-based services in production environment.
  • Hands-on experience with modern inference engines (SGLang, vLLM, TensorRT-LLM).
  • Familiar with fine-tuning LLMs and other AI models.

Responsibilities

  • Design and build Together’s systems for customizing open-source models.
  • Build integrations between the Model Shaping and Inference platforms for seamless post-training to production serving.
  • Add features to inference engines for large-scale post-training experiments and RL workload optimizations.
  • Ensure the service is stable, with on-call rotation and 24/7 platform availability.

Skills

ML services
Inference engines
Python
Go
LLMs fine-tuning

Tools

SGLang
vLLM
TensorRT-LLM

Job description

Together AI in San Francisco is seeking a Research Engineer to help build a platform that lets users customize open-source models, bridging post-training and production inference. You will contribute across Fine-Tuning, RL, and Evaluation services and collaborate with product, research, and engineering teams to keep the API reliable and scalable.

This full-time role offers a competitive US salary range and equity, with focus on building foundational AI infrastructure for developers worldwide.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Systems Engineer, AI Inference Platform
Senior Systems Engineer, AI Inference Platform

Slope • San Francisco (CA)

On-site
USD 180,000 - 260,000
Inference Platform Backend Engineer (Equity & Benefits)
Inference Platform Backend Engineer (Equity & Benefits)

Together • San Francisco (CA)

On-site
USD 160,000 - 250,000
Equity
Health insurance
Competitive compensation
Research Engineer, Post-Training Inference
Research Engineer, Post-Training Inference

Together AI • San Francisco (CA)

On-site
USD 200,000 - 290,000
Health insurance
Startup equity
Software Engineer, Model Inference
Software Engineer, Model Inference

OpenAI • San Francisco (CA)

On-site
USD 325,000 - 490,000
Inference Research Intern: Build Scalable AI Serving
Inference Research Intern: Build Scalable AI Serving

Together AI • San Francisco (CA)

On-site
USD 80,000 - 96,000
Housing stipend
Competitive compensation
Software Engineer, AI Inference Infrastructure Platform
Software Engineer, AI Inference Infrastructure Platform

Slope • San Francisco (CA)

On-site
USD 180,000 - 260,000
Senior Software Engineer - Research Platform, Consumer Devices
Senior Software Engineer - Research Platform, Consumer Devices

OpenAI • San Francisco (CA)

On-site
USD 293,000 - 325,000
Director of Engineering, AI Inference Platform
Director of Engineering, AI Inference Platform

Oho Group • San Francisco (CA)

On-site
USD 260,000 - 420,000
ML Platform Engineer: Inference & Production
ML Platform Engineer: Inference & Production

Xapply • San Francisco (CA)

On-site
USD 190,000 - 230,000
Platform Engineer for AI Inference & Optimization
Platform Engineer for AI Inference & Optimization

OpenAI • Seattle (WA)

On-site
USD 180,000 - 240,000