Inference & Post-Training Engineer (Remote, Equity)

Together

San Francisco (CA)

On-site

USD 270,000 - 300,000

Full time

5 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Startup equity
Health insurance
Remote work flexibility

Job summary

Together AI, the AI Native Cloud, is seeking a Forward Deployed Engineer focused on Inference & Post-Training to partner with our most strategic customers. You will optimize inference engines, tune deployments, and guide LoRA, SFT, DPO, RLHF, and GRPO pipelines from experimentation through production.

You will collaborate with SAs, CX, and Sales to ensure successful platform adoption, meet POC milestones, and influence the product roadmap with field insights.

Qualifications

  • 5+ years in a technical role focused on inference systems.
  • Hands-on experience with open-source LLM deployment.
  • Strong Python skills in production environments.

Responsibilities

  • Inference Engine Optimization: Select, configure, and optimize inference engine based on hardware, model architecture, and workload profile.
  • Configuration & Performance Tuning: Develop configuration updates to win critical POCs, benchmarks, and optimize customer deployments; tune KV cache, apply speculative decoding, determine optimal tensor parallelism, and determine quantization strategy to hit throughput and latency targets.
  • Post-Training & Fine-Tuning: Drive hands-on RL training runs and optimize system design; guide customers through LoRA, SFT, DPO, RLHF, and GRPO pipelines from experimentation through production.
  • Strategic Customer Alignment: Act as the primary technical point of contact for aligned strategic accounts — monitoring and optimizing endpoint configurations, helping customers get the most out of the platform, and collaborating to ensure we hit critical milestones.
  • Opinionated Onboarding: Establish direct alignment with strategic customers at onboarding; ensure the right inference and post-training configurations are in place from day one to improve time-to-value.
  • Product Feedback Loop: Directly influence our software and model roadmap by surfacing insights from the field. Contribute back to the product where needed to support customer requirements or drive a better experience. Drive early feature and research adoption with strategic logos.

Skills

Python
Inference systems
Open-source deployment
Production environments
Model fine-tuning pipelines

Tools

vLLM
TensorRT-LLM
SGLang

Job description

Together AI, the AI Native Cloud, is seeking a Forward Deployed Engineer focused on Inference & Post-Training to partner with our most strategic customers. You will optimize inference engines, tune deployments, and guide LoRA, SFT, DPO, RLHF, and GRPO pipelines from experimentation through production.

You will collaborate with SAs, CX, and Sales to ensure successful platform adoption, meet POC milestones, and influence the product roadmap with field insights.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote AI Inference Support Engineer
Remote AI Inference Support Engineer

Together AI • New York (NY)

On-site
USD 160,000 - 230,000
Health insurance
Startup equity
Remote work
Remote Forward Deployed Engineer for AI Inference
Remote Forward Deployed Engineer for AI Inference

Tenstorrent • Austin (TX)

On-site
USD 100,000 - 500,000
Forward Deployed Engineer – AI Training Data (Post‑Sales)
Forward Deployed Engineer – AI Training Data (Post‑Sales)

Talanto • Redwood City (CA), Northern (KY)

Hybrid
USD 19,000 - 25,000
Health benefits
401(k)
Unlimited PTO
+5
Forward-Deployed AI Engineer: Scale Inference & Ship Solutions
Forward-Deployed AI Engineer: Scale Inference & Ship Solutions

Triwill Group • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Equity
Medical, dental, vision coverage fore
Flexible PTO
+4
Remote Forward-Deployed AI Engineer
Remote Forward-Deployed AI Engineer

Tenstorrent • Santa Clara (CA), Austin (TX)

Hybrid
CAD 139,000 - 694,000
Senior AI Solutions Engineer — Remote & Equity
Senior AI Solutions Engineer — Remote & Equity

Front • San Francisco (CA)

Hybrid
USD 150,000 - 210,000
Competitive salary
Equity
Private health insurance
+4
Remote AI Inference Support Engineer (SRE)
Remote AI Inference Support Engineer (SRE)

Together AI • United States

Remote
USD 160,000 - 230,000
Remote AI Inference Support Engineer (SRE)
Remote AI Inference Support Engineer (SRE)

Together • United States

Remote
USD 160,000 - 230,000
Startup equity
Health insurance
Remote work flexibility
Senior AI Platform Engineer - Remote Cloud-Native ML Infra
Senior AI Platform Engineer - Remote Cloud-Native ML Infra

Bright Vision Technologies • Columbus (OH), Dublin (OH)

On-site
USD 130,000 - 180,000
Remote AI Product Engineer — Forward Deployed (Equity)
Remote AI Product Engineer — Forward Deployed (Equity)

S:23 Recruitment LTD • San Francisco (CA)

On-site
USD 140,000 - 180,000