Data Scientist II: LLM/RAG for Copilot

DataJobs

Washington (District of Columbia)

On-site

USD 102,000 - 219,000

Full time

7 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Microsoft in Washington, DC is seeking a research-focused role on the Turing Team to build model-driven capabilities for Copilot, emphasizing orchestrator reasoning, evaluation, and A/B testing for LLM and agentic product performance. This onsite role focuses on deep learning, NLP, multimodality, and model optimization.

You’ll work on evaluation frameworks, model experiments, diagnosis, prototype AI in containerized environments (AI Foundry, LangChain), and retrieval-augmented generation

Qualifications

  • Bachelor’s degree in relevant field plus 2+ years of data-science experience.
  • Master’s degree plus 1+ year data-science or consulting experience.
  • Doctorate in the same fields.
  • Equivalent experience accepted.

Responsibilities

  • Define evaluation metrics and frameworks for AI/LLM systems and agentic product performance.
  • Experiment with models before selection for tasks within product domains.
  • Diagnose issues in AI/LLM systems and identify opportunities to improve performance.
  • Prototype and develop AI models in containerized environments (AI Foundry, LangChain).
  • Develop and deploy retrieval-augmented generation (RAG) workflows.
  • Collaborate with senior members to refine LLM systems and models.
  • Take model-driven features from concept to production across development, evaluation, metrics, and A/B testing.

Skills

Data Science
NLP
Machine Learning
AB testing
LangChain

Education

Bachelor’s degree in Data Science/Related field
Master’s degree in Data Science/Related field
Doctorate in Data Science/Related field

Tools

LangChain
AI Foundry

Job description

Microsoft in Washington, DC is seeking a research-focused role on the Turing Team to build model-driven capabilities for Copilot, emphasizing orchestrator reasoning, evaluation, and A/B testing for LLM and agentic product performance. This onsite role focuses on deep learning, NLP, multimodality, and model optimization.

You’ll work on evaluation frameworks, model experiments, diagnosis, prototype AI in containerized environments (AI Foundry, LangChain), and retrieval-augmented generation

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data Scientist II, M365 Copilot
Data Scientist II, M365 Copilot

DataJobs • Washington

On-site
USD 102,000 - 219,000
AI Engineer – Microsoft Copilot & Generative AI
AI Engineer – Microsoft Copilot & Generative AI

Matlen Silver • Charlotte (NC)

On-site
USD 140,000 - 190,000
Gen AI & NLP ML Engineer (LLMs/RAG) – Onsite DC
Gen AI & NLP ML Engineer (LLMs/RAG) – Onsite DC

AIToolboard • Washington

On-site
USD 120,000 - 180,000
GenAI & LLM Engineer for Enterprise Copilots
GenAI & LLM Engineer for Enterprise Copilots

Infosys • Irving (TX)

On-site
USD 110,000 - 150,000
GenAI Engineer: Enterprise Copilot & LLM Solutions
GenAI Engineer: Enterprise Copilot & LLM Solutions

Matlen Silver • Charlotte (NC)

On-site
USD 140,000 - 190,000
AI Researcher - Regulatory LLM & RAG Systems
AI Researcher - Regulatory LLM & RAG Systems

Zohorecruit • Silver Spring (MD)

On-site
USD 140,000 - 180,000
PPO/HMO Health Plan
Data Scientist, Consultant
Data Scientist, Consultant

Blue Shield of CA • Springfield Meadows (CA)

Hybrid
USD 140,000 - 230,000
Data Scientist, Consultant
Data Scientist, Consultant

Blue Shield of CA • Huntsville (AL)

Hybrid
USD 120,000 - 170,000
Data Scientist Engineer II — AI/ML & LLM/NLP (TS/SCI)
Data Scientist Engineer II — AI/ML & LLM/NLP (TS/SCI)

Peraton • Maryland

On-site
USD 90,000 - 120,000
Data Scientist, Consultant
Data Scientist, Consultant

Blue Shield of CA • Las Vegas (NV)

Hybrid
USD 120,000 - 190,000