RAG ML Engineer II: Retrieval & LLM Orchestration

Jobtailor

Massachusetts

On-site

USD 140,000 - 170,000

Full time

12 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Jobtailor is seeking an experienced ML/NLP Engineer to design and implement end-to-end RAG pipelines, combining chunking, embedding models, and vector databases. You will build scalable retrieval systems for large proprietary datasets and orchestrate LLM-based solutions for high-quality, context-aware responses.

You will collaborate with Product, Design, and ML Ops to automate the ML lifecycle from development to deployment in a fast-paced environment.

Qualifications

  • Bachelor's degree or higher in Computer Science, Engineering, or a related field.
  • 3+ years of hands-on industry experience with ML, NLP, information retrieval, and large-scale text processing.
  • Experience designing, shipping, and maintaining production systems.
  • Strong Python programming skills.
  • Experience with PyTorch, Transformers, and HuggingFace.
  • Experience with LangChain and LlamaIndex.
  • Knowledge of vector databases and similarity search techniques.

Responsibilities

  • Design and implement end-to-end RAG pipelines integrating chunking, embedding models, vector databases, and data retrieval agents.
  • Build and optimize retrieval systems over large-scale proprietary datasets using advanced embedding techniques.
  • Develop LLM-based solutions orchestrating retrieval, generation, and ranking for context-aware responses.
  • Investigate challenges in vector search, chunking/indexing strategies, and unstructured data retrieval evaluation.
  • Collaborate with Product and Design teams to build ML-based solutions enhancing user experiences and business objectives.
  • Work with ML Operations to automate management of the full ML systems lifecycle from design to implementation.

Skills

RAG pipelines
End-to-end design
Python programming
LLM orchestration
Information retrieval
NLP
Vector databases
Embedding models
Code optimization

Education

Bachelor's degree or higher in CS/Engineering

Tools

PyTorch
Transformers
HuggingFace
LangChain
LlamaIndex
PostgreSQL/PGVector
OpenSearch
Pinecone

Job description

Jobtailor is seeking an experienced ML/NLP Engineer to design and implement end-to-end RAG pipelines, combining chunking, embedding models, and vector databases. You will build scalable retrieval systems for large proprietary datasets and orchestrate LLM-based solutions for high-quality, context-aware responses.

You will collaborate with Product, Design, and ML Ops to automate the ML lifecycle from development to deployment in a fast-paced environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

RAG ML Engineer — Retrieval & LLM Orchestration
RAG ML Engineer — Retrieval & LLM Orchestration

S&P Global, Inc. • New York (NY)

On-site
USD 140,000 - 180,000
Medical, Dental, and Vision insurance
Unlimited Paid Time Off
Parental Leave (26 weeks)
ML Engineer II: RAG Pipelines & LLM Orchestration
ML Engineer II: RAG Pipelines & LLM Orchestration

Kensho Technologies • Cambridge (MA)

On-site
USD 140,000 - 180,000
Medical insurance
Dental insurance
Vision insurance
+6
End-to-End AI/ML Engineer for LLMs, RAG & Agentic Systems
End-to-End AI/ML Engineer for LLMs, RAG & Agentic Systems

Jobtailor • Colorado

On-site
USD 120,000 - 160,000
Junior AI Engineer: LLM Pipelines & RAG Solutions
Junior AI Engineer: LLM Pipelines & RAG Solutions

Jobtailor • Atlanta (GA)

On-site
USD 150,000 - 190,000
Retrieval-Augmented Generation (RAG) System - Senior Software Developer
Retrieval-Augmented Generation (RAG) System - Senior Software Developer

Elsevier • Philadelphia

On-site
USD 120,000 - 150,000
AI/ML Engineer: LLMs, RAG & Prototyping
AI/ML Engineer: LLMs, RAG & Prototyping

Charter Global • Town of Florida (NY)

On-site
USD 120,000 - 170,000
Senior AI/ML Engineer: LLM, RAG & End-to-End Deployment
Senior AI/ML Engineer: LLM, RAG & End-to-End Deployment

Vizient, Inc. • Centennial (CO)

On-site
USD 102,000 - 179,000
Incentive eligible
Benefits plan
RAG Engineer: AI Knowledge Retrieval Architect
RAG Engineer: AI Knowledge Retrieval Architect

MCI • United States

On-site
USD 120,000 - 180,000
AI Engineer: On-Prem LLMs & RAG Pipelines
AI Engineer: On-Prem LLMs & RAG Pipelines

Salvo Software LLC • Northern (KY)

Hybrid
USD 120,000 - 190,000
Senior ML Engineer: RAG/LLM Systems & Rapid Prototyping
Senior ML Engineer: RAG/LLM Systems & Rapid Prototyping

Grid Dynamics • United States

On-site
USD 140,000 - 190,000
Flexible schedule
Medical insurance
Vision and dental
+2