ML Engineer II: RAG Pipelines & LLM Orchestration

Kensho Technologies

Cambridge (MA)

On-site

USD 140,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical insurance
Dental insurance
Vision insurance
Unlimited PTO
Parental Leave
401(k) with 6% match
Tuition assistance
Conferences/learning opportunities
Dog-friendly office

Job summary

Kensho Technologies in New York, NY, seeks a mid-level Machine Learning Engineer to design and scale RAG pipelines, focusing on retrieval models, LLM orchestration, and system-level thinking. You will work with embedding techniques and proprietary data, contributing to enterprise search, knowledge discovery, and decision-support systems.

The role requires hands-on ML/NLP experience, Python expertise, and production ML tooling.

Qualifications

  • Bachelor's degree in Computer Science, Engineering, or related field.
  • 3+ years hands-on ML/NLP experience including production systems.
  • Strong Python skills and ML frameworks (PyTorch, Transformers, HuggingFace).

Responsibilities

  • Design end-to-end RAG pipelines with embedding models and data retrieval agents.
  • Build and optimize retrieval systems over large proprietary datasets.
  • Develop LLM-based solutions for retrieval, generation, and ranking.
  • Solve challenges in vector search, indexing, and unstructured data retrieval.
  • Collaborate with Product/Design to meet business objectives.
  • Work with ML Ops to manage ML systems lifecycle.

Skills

Python programming
NLP
Information retrieval
LLM orchestration
Vector databases
LangChain/LlamaIndex

Education

Bachelor's degree in CS/Engineering or related field

Tools

LangChain
LLamaIndex
PyTorch
Transformers
HuggingFace
OpenSearch
PostgreSQL/PGVector
LiteLLM

Job description

Kensho Technologies in New York, NY, seeks a mid-level Machine Learning Engineer to design and scale RAG pipelines, focusing on retrieval models, LLM orchestration, and system-level thinking. You will work with embedding techniques and proprietary data, contributing to enterprise search, knowledge discovery, and decision-support systems.

The role requires hands-on ML/NLP experience, Python expertise, and production ML tooling.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

RAG ML Engineer — Retrieval & LLM Orchestration
RAG ML Engineer — Retrieval & LLM Orchestration

S&P Global, Inc. • New York (NY)

On-site
USD 140,000 - 180,000
Medical, Dental, and Vision insurance
Unlimited Paid Time Off
Parental Leave (26 weeks)
RAG ML Engineer II: Retrieval & LLM Orchestration
RAG ML Engineer II: Retrieval & LLM Orchestration

Jobtailor • Massachusetts

On-site
USD 140,000 - 170,000
ML Engineer II: Agentic AI & LLM Orchestration
ML Engineer II: Agentic AI & LLM Orchestration

S&P Global • Cambridge (MA)

On-site
USD 140,000 - 180,000
Medical benefits
Unlimited PTO
Parental Leave
+2
ML Engineer - LLMs, AI Pipelines & Production
ML Engineer - LLMs, AI Pipelines & Production

Abile Group, Inc • Chantilly (VA)

On-site
USD 140,000 - 170,000
AI Engineer: On-Prem LLMs & RAG Pipelines
AI Engineer: On-Prem LLMs & RAG Pipelines

Salvo Software LLC • Northern (KY)

Hybrid
USD 120,000 - 190,000
Machine Learning Engineer II
Machine Learning Engineer II

Kensho Technologies • Cambridge (MA)

On-site
USD 140,000 - 180,000
Medical insurance
Dental insurance
Vision insurance
+6
LLM Applications Engineer - RAG Pipelines & AWS Integration
LLM Applications Engineer - RAG Pipelines & AWS Integration

GAC Solutions • New Jersey

On-site
USD 120,000 - 160,000
LLM-Powered AI Engineer: RAG, Prompts & Production ML
LLM-Powered AI Engineer: RAG, Prompts & Production ML

RiskForce • Northern (KY)

Hybrid
USD 120,000 - 155,000
Machine Learning Engineer II
Machine Learning Engineer II

S&P Global, Inc. • New York (NY)

On-site
USD 140,000 - 180,000
Medical, Dental, and Vision insurance
Unlimited Paid Time Off
Parental Leave (26 weeks)
Junior ML Engineer: Data Pipelines & LLM Deployment
Junior ML Engineer: Data Pipelines & LLM Deployment

Pangram • New York (NY)

On-site
USD 90,000 - 130,000