Staff Data Scientist, AI Evaluation & LLM Quality

Perplexity

United States

Hybrid

USD 100,000 - 150,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Perplexity, based in the United States, is seeking a qualified candidate to enhance our automated evaluation processes. Responsibilities include architecting evaluation pipelines and developing methods to ensure the accuracy and helpfulness of answers across our platforms.

The ideal applicant will possess a PhD or MS in a technical field, along with over 4 years of experience in data science or machine learning, and strong coding skills in Python and SQL. Experience with AWS and Databricks is preferred.

Qualifications

  • PhD or MS in a technical field or equivalent experience.
  • 4+ years of experience in data science or machine learning.
  • Strong proficiency in Python and SQL.

Responsibilities

  • Architect and maintain automated evaluation pipelines for answer quality.
  • Design evaluation sets to measure the impact of tool calls on answer quality.
  • Develop VLM-based solutions for visual answer evaluation.

Skills

Python
SQL
Data science
Machine learning
AWS
Databricks

Education

PhD or MS in a technical field

Job description

Perplexity, based in the United States, is seeking a qualified candidate to enhance our automated evaluation processes. Responsibilities include architecting evaluation pipelines and developing methods to ensure the accuracy and helpfulness of answers across our platforms.

The ideal applicant will possess a PhD or MS in a technical field, along with over 4 years of experience in data science or machine learning, and strong coding skills in Python and SQL. Experience with AWS and Databricks is preferred.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Member of Technical Staff (Data Scientist, Evals)
Member of Technical Staff (Data Scientist, Evals)

Perplexity • United States

Hybrid
USD 100,000 - 150,000
Member of Technical Staff (Answer Quality & Evals)
Member of Technical Staff (Answer Quality & Evals)

Perplexity • San Francisco (CA)

On-site
USD 180,000 - 260,000
Member of Technical Staff (Answer Quality & Evals)
Member of Technical Staff (Answer Quality & Evals)

Perplexity AI • New York (NY)

On-site
USD 120,000 - 190,000
Member of Technical Staff (Answer Quality & Evals)
Member of Technical Staff (Answer Quality & Evals)

Pantera Capital • Palo Alto (CA)

On-site
USD 150,000 - 210,000
Staff Engineer, AI Evaluation & Data Systems
Staff Engineer, AI Evaluation & Data Systems

Perplexity AI • New York (NY)

On-site
USD 120,000 - 190,000
Staff Engineer, AI Evaluation & Observability
Staff Engineer, AI Evaluation & Observability

Pantera Capital • Palo Alto (CA)

On-site
USD 150,000 - 210,000
Staff Data Scientist, LLM Data Quality & Evaluation
Staff Data Scientist, LLM Data Quality & Evaluation

Cohere • San Francisco (CA)

Remote
USD 120,000 - 160,000
Weekly lunch stipend
Full health and dental benefits
401K and pension scheme
+3
Staff Data Scientist — LLM Data Analysis & Evaluation
Staff Data Scientist — LLM Data Analysis & Evaluation

Cohere • New York (NY), San Francisco (CA)

Hybrid
USD 150,000 - 210,000
Lunch stipend
Health benefits
RRSP matching
+6
Member of Technical Staff (Software Engineer, Data Platform)
Member of Technical Staff (Software Engineer, Data Platform)

Perplexity • San Francisco (CA)

On-site
USD 130,000 - 170,000
Applied Scientist, AGI Quality & LLM Evaluation
Applied Scientist, AGI Quality & LLM Evaluation

Amazon • Factoria (WA)

On-site
USD 136,000 - 184,000