AI Researcher Lead 1

Wipro

San Francisco (CA)

On-site

USD 240,000 - 375,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Medical benefits
Dental benefits
Paid time off
Disability insurance

Job summary

Wipro is seeking a Lead Researcher for its AI Data Foundry in San Francisco to advance measurement and alignment for the next generation of LLMs and SLMs. You will architect preference data pipelines, design proactive benchmarks, and publish novel methodologies on arXiv and at top-tier conferences.

You will collaborate with domain researchers to synthesize ground-truth data into robust evaluation frameworks.

Qualifications

  • Advanced degree in CS/AI/quantitative field.
  • Expertise in RLHF/DPO/RLAIF and evaluation harnesses.
  • Proficient in PyTorch, HuggingFace, and data manipulation.

Responsibilities

  • Design alignment methodologies for human-in-the-loop and synthetic data.
  • Build proactive, domain-specific benchmarks for advanced AI systems.
  • Decontaminate data pipelines to detect overlaps and mitigate reward hacking.
  • Publish research on evaluation frameworks and alignment science.
  • Lead technical collaboration with researchers to produce scoreable evaluations.

Skills

RLHF/DPO/RLAIF experience
Bradley-Terry ranking
Data pipeline design
Benchmark design
Open research publication
Python
PyTorch
HuggingFace

Education

PhD or MSc in Computer Science / AI / Mathematics

Tools

PyTorch
HuggingFace
vLLM
EleutherAI LM Eval
HELM

Job description

Role Overview As an AI Data Foundry, we partner with top-tier labs to build the critical SFT and RLHF datasets that train the next generation of LLMs, SLMs, and Physical AI. We are seeking a Lead Researcher who focuses on the rigorous science of measurement and alignment. In this role, you will architect preference data pipelines, design proactive benchmarks from the ground up, and publish novel methodologies on arXiv and at top-tier conferences (e.g., NeurIPS, ICLR).

Job description

Role Overview As an AI Data Foundry, we partner with top-tier labs to build the critical SFT and RLHF datasets that train the next generation of LLMs, SLMs, and Physical AI. We are seeking a Lead Researcher who focuses on the rigorous science of measurement and alignment. In this role, you will architect preference data pipelines, design proactive benchmarks from the ground up, and publish novel methodologies on arXiv and at top-tier conferences (e.g., NeurIPS, ICLR).

Key Responsibilities
  • Design Alignment Methodologies: Architect frameworks for human-in-the-loop and synthetic data generation. Design pairwise preference rubrics (utilizing Bradley-Terry models) and reward mechanisms for DPO, KTO, and PPO pipelines.
  • Build Proactive Benchmarks: Lead the creation of novel, domain-specific benchmarks, spanning agentic reasoning to physical AI, that evaluate complex capabilities rather than relying on saturated, static datasets.
  • Decontaminate & Validate: Build robust pipelines to detect data contamination (via n-gram overlap, embedding similarity) to ensure deliverables are mathematically sound and mitigate reward hacking.
  • Research Publication: Conduct independent research on evaluation frameworks and alignment science, co-authoring papers to establish our technical authority in the field.
  • Technical Leadership: Work closely with domain-specific researchers to synthesize ground-truth data into cohesive, scoreable, and statistically sound evaluation frameworks.
Candidate Profile
  • Academic Background: Advanced degree (PhD or high-impact MSc) in Computer Science, Artificial Intelligence, Mathematics, or a highly quantitative field.
  • Technical Expertise: Deep familiarity with modern alignment techniques (RLHF, DPO, RLAIF) and standard evaluation harnesses (e.g., EleutherAI LM Eval, HELM, OpenAI Evals).
  • Core Stack: Advanced proficiency in PyTorch, HuggingFace, vLLM, and complex data manipulation.
  • Model Experience: Hands-on experience fine-tuning or evaluating open-weight models (Llama 3, Mistral, Qwen) with a deep understanding of their architectural bottlenecks.

Expected annual pay for this role ranges from$240,000.00 to$375,000.00. Based on the position, the role is also eligible for Wipro’s standard benefits including a full range of medical and dental benefits options, disability insurance, paid time off (inclusive of sick leave), other paid and unpaid leave options

Reinvent your world.We are building a modern Wipro. We are an end-to-end digital transformation partner with the boldest ambitions. To realize them, we need people inspired by reinvention. Of yourself, your career, and your skills. We want to see the constant evolution of our business and our industry. It has always been in our DNA - as the world around us changes, so do we. Join a business powered by purpose and a place that empowers you to design your own reinvention.

Applications from people with disabilities are explicitly welcome.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Researcher Lead 1
AI Researcher Lead 1

Wipro Technologies • San Francisco (CA)

On-site
USD 240,000 - 375,000
Medical and dental benefits
Disability insurance
Paid time off
Domain AI Researcher
Domain AI Researcher

Wipro • San Francisco (CA)

On-site
USD 240,000 - 413,000
Medical and dental benefits
Disability insurance
Paid time off
Domain AI Researcher
Domain AI Researcher

Wipro Technologies • San Francisco (CA)

On-site
USD 240,000 - 413,000
Medical benefits
Dental benefits
Paid time off
Lead AI Alignment & Evaluation Researcher
Lead AI Alignment & Evaluation Researcher

Wipro • San Francisco (CA)

On-site
USD 240,000 - 375,000
Medical benefits
Dental benefits
Paid time off
+1
Senior Data Scientist - AI/ML & Generative AI Lead
Senior Data Scientist - AI/ML & Generative AI Lead

Wipro • Atlanta (GA)

On-site
USD 45,000 - 121,000
Lead AI Alignment & Evaluation Scientist
Lead AI Alignment & Evaluation Scientist

Wipro Technologies • San Francisco (CA)

On-site
USD 240,000 - 375,000
Medical and dental benefits
Disability insurance
Paid time off
Applied Research Engineer
Applied Research Engineer

Sterling Inspired Staffing. • San Francisco (CA)

On-site
USD 120,000 - 160,000
Comprehensive medical plans
Generous parental leave
Unlimited PTO
+3
Principal Applied Scientist- AI
Principal Applied Scientist- AI

UiPath • Bellevue (WA)

Hybrid
USD 218,000 - 265,000
SOLUTION ARCHITECT L2
SOLUTION ARCHITECT L2

Wipro • Orange (CA)

On-site
USD 100,000 - 230,000
Full range of medical and dental benefits
Disability insurance
Paid time off including sick leave
+1
SALES EXECUTIVE - AI Sales
SALES EXECUTIVE - AI Sales

Wipro • Mountain View (CA)

On-site
USD 150,000 - 231,000
Medical & dental benefits
Disability insurance
Paid time off
+1