Enterprise AI Research Engineer LLMs RAG & Agents

Fabrion

San Francisco (CA)

On-site

USD 180,000 - 260,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Equity

Job summary

Fabrion is building the next generation of enterprise AI infrastructure in the San Francisco Bay Area. We seek an ML/AI Research Engineer to lead design, training, evaluation, and optimization of agent-native AI models that sit atop our enterprise data fabric.

You will work at the intersection of LLMs, vector search, graph reasoning, and reinforcement learning, driving end-to-end ML from data curation to deployment with cost-awareness and alignment in scope.

Qualifications

  • Deep experience fine-tuning open-source LLMs using HuggingFace Transformers, DeepSpeed, vLLM, FSDP, LoRA/QLoRA.
  • Worked with base and instruction-tuned models; familiar with SFT, RLHF, DPO pipelines.
  • Experience building enterprise-grade RAG pipelines integrated with real-time or contextual data.

Responsibilities

  • Fine-tune and evaluate open-source LLMs for enterprise use cases with structured and unstructured data.
  • Build and optimize RAG pipelines using LangChain, LangGraph, LlamaIndex or Dust; integrate with vector DBs and internal knowledge graph.
  • Train agent architectures (ReAct, AutoGPT, BabyAGI, OpenAgents) using enterprise task data.
  • Develop embedding-based memory and retrieval chains with token-efficient chunking strategies.
  • Create reinforcement learning pipelines to optimize agent behaviors (RLHF, DPO, PPO).
  • Establish scalable evaluation harnesses for LLM and agent performance, including synthetic evals and explainability tools.
  • Contribute to model observability, drift detection, error classification, and alignment.
  • Optimize inference latency and GPU resource utilization across cloud and on-prem environments.

Skills

LLM training & inference
Agent orchestration
RAG pipelines design
Memory & retrieval
Reinforcement learning
Model evaluation & interpretability
Explainability & safety

Tools

LangChain
LangGraph
LlamaIndex
Weaviate/FAISS
Pinecone
DeepSpeed
vLLM

Job description

Fabrion is building the next generation of enterprise AI infrastructure in the San Francisco Bay Area. We seek an ML/AI Research Engineer to lead design, training, evaluation, and optimization of agent-native AI models that sit atop our enterprise data fabric.

You will work at the intersection of LLMs, vector search, graph reasoning, and reinforcement learning, driving end-to-end ML from data curation to deployment with cost-awareness and alignment in scope.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML/AI Research Engineer — Agentic AI Lab (Founding Team)
ML/AI Research Engineer — Agentic AI Lab (Founding Team)

Fabrion • San Francisco (CA)

On-site
USD 180,000 - 260,000
Equity
Production AI Engineer: LLM, RAG & Multi-Agent Systems
Production AI Engineer: LLM, RAG & Multi-Agent Systems

Zof AI • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Principal AI Researcher: Lead Agentic AI & LLMs
Principal AI Researcher: Lead Agentic AI & LLMs

HR Tech Job • Pleasanton (CA)

Hybrid
USD 228,000 - 342,000
Staff AI Engineer: LLMs, Agents & Enterprise Automation
Staff AI Engineer: LLMs, Agents & Enterprise Automation

Harnham • San Francisco (CA)

Hybrid
USD 250,000 - 350,000
Equity
Full benefits
Remote AI/ML Engineer — LLMs, RAG & Multi-Agent
Remote AI/ML Engineer — LLMs, RAG & Multi-Agent

YO AI Labs • Phoenix (AZ)

Remote
USD 120,000 - 180,000
Enterprise ML Research Scientist - GenAI Data Foundations
Enterprise ML Research Scientist - GenAI Data Foundations

Scale AI • San Francisco (CA)

On-site
USD 265,000 - 331,000
Health, dental & vision coverage
Equity compensation
Generous PTO
+2
Enterprise AI Engineer — LLM Agents for North
Enterprise AI Engineer — LLM Agents for North

Cohere • California (MO)

Hybrid
USD 150,000 - 230,000
Lunch stipend
Health benefits
RRSP matching / 401K
+1
GenAI & LLM Engineer — RAG, Agents, Fine-Tuning (Remote)
GenAI & LLM Engineer — RAG, Agents, Fine-Tuning (Remote)

ACI Infotech • United States

Remote
USD 120,000 - 180,000
Competitive salary and equity
Health, dental, and vision insurance
401(k) with company match
+2
Pioneer AI Research Lead — Enterprise Systems
Pioneer AI Research Lead — Enterprise Systems

Fabrion • San Francisco (CA)

On-site
USD 180,000 - 240,000
Senior AI/ML Engineer: Agentic LLMs, RAG & Production ML
Senior AI/ML Engineer: Agentic LLMs, RAG & Production ML

Orbien LLC • Maryland

On-site
USD 150,000 - 230,000