Staff Engineer, AI Evaluation & Data Systems

Perplexity AI

New York (NY)

On-site

USD 120,000 - 190,000

Full time

8 days ago
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Perplexity AI invites you to join as an Answer Quality Engineer to help design and implement evaluation systems for high-quality answers. You will work across prompts, tools, search, datasets, and models to improve user experiences by identifying and solving answer-quality issues through data-driven evaluations.

You will collaborate with data scientists, engineers, and product teams to turn evaluation findings into concrete product improvements, building scalable, observable infrastructure for

Qualifications

  • 4+ years of software, data, or ML engineering experience shipping and operating production systems.
  • Strong proficiency in Python and SQL with fundamentals in system design and data modeling.
  • Experience with big-data pipelines and distributed systems.

Responsibilities

  • Build shared evaluation infrastructure that helps teams run reliable evals, analyze results, and make product and model decisions
  • Develop the platform for replaying and analyzing agent traces to reproduce production behavior and diagnose failures
  • Build and operate scalable systems for processing, storing, and monitoring interaction, trace, and evaluation data
  • Partner with data scientists, engineers, and product teams to turn answer-quality problems into evaluations, analyses, and product improvements
  • Operate in a small, high-impact team where your work directly shapes how Perplexity measures and improves Answer Quality

Skills

Python
SQL
System design
Data modeling
Distributed systems

Tools

Databricks
Snowflake
ClickHouse

Job description

Perplexity AI invites you to join as an Answer Quality Engineer to help design and implement evaluation systems for high-quality answers. You will work across prompts, tools, search, datasets, and models to improve user experiences by identifying and solving answer-quality issues through data-driven evaluations.

You will collaborate with data scientists, engineers, and product teams to turn evaluation findings into concrete product improvements, building scalable, observable infrastructure for

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Engineer, AI Evaluation & Observability
Staff Engineer, AI Evaluation & Observability

Pantera Capital • Palo Alto (CA)

On-site
USD 150,000 - 210,000
Staff Engineer - Answer Quality & Evaluation Platforms
Staff Engineer - Answer Quality & Evaluation Platforms

Perplexity • San Francisco (CA)

On-site
USD 180,000 - 260,000
Member of Technical Staff (Answer Quality & Evals)
Member of Technical Staff (Answer Quality & Evals)

Perplexity AI • New York (NY)

On-site
USD 120,000 - 190,000
Member of Technical Staff (Answer Quality & Evals)
Member of Technical Staff (Answer Quality & Evals)

Pantera Capital • Palo Alto (CA)

On-site
USD 150,000 - 210,000
Member of Technical Staff (Answer Quality & Evals)
Member of Technical Staff (Answer Quality & Evals)

Perplexity • San Francisco (CA)

On-site
USD 180,000 - 260,000
Staff Data Scientist, AI Evaluation & LLM Quality
Staff Data Scientist, AI Evaluation & LLM Quality

Perplexity • United States

Hybrid
USD 100,000 - 150,000
AI-Driven Data Systems Engineer
AI-Driven Data Systems Engineer

Perplexity • San Francisco (CA)

On-site
USD 140,000 - 210,000
Member of Technical Staff (Data Scientist, Evals)
Member of Technical Staff (Data Scientist, Evals)

Perplexity • United States

Hybrid
USD 100,000 - 150,000
Member of Data Staff (AI Builder)
Member of Data Staff (AI Builder)

Perplexity • San Francisco (CA)

On-site
USD 140,000 - 210,000
AI-Powered Data Systems Engineer
AI-Powered Data Systems Engineer

Neura Market • San Francisco (CA)

Hybrid
USD 180,000 - 240,000