Staff Software Engineer, Data Eval & Simulation

Perplexity AI

United States

On-site

USD 120,000 - 190,000

Full time

10 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Perplexity AI is seeking an experienced software engineer to build and maintain the data pipelines powering eval verdicts and product decisions. You will own the flow from raw signals to durable datasets that influence decision-making across the company, and you’ll develop a simulator to replay user interactions for LLMs and VLMs.

You will ensure data trust with monitoring and lineage in a fast, high-impact team.

Qualifications

  • 3+ years of software engineering experience shipping production systems.
  • Proficiency in Python and SQL with production-grade code.
  • Experience with AWS and lakehouse ecosystems like Databricks or Spark.
  • Strong fundamentals in data modeling, system design, and debugging distributed systems.
  • Experience with big data systems including distributed compute and large-scale storage.
  • Familiarity with LLM/VLM interfaces, tokenization, and multimodal payloads.
  • Experience with evaluation platforms, experimentation systems, or ML infrastructure.
  • Prior work supporting customer-facing products at scale.

Responsibilities

  • Build the systems and pipelines that enable Search, Product, and other teams to access eval verdicts.
  • Own the evals-to-product loop and turn raw signals into durable datasets.
  • Create a robust simulator pipeline to replay user interactions for LLMs and VLMs.
  • Maintain data trust via monitoring, lineage, and quality checks.
  • Work in a small, high-impact team shaping Perplexity's answer quality.

Skills

Python
SQL
Data modeling
Distributed systems
AWS
Databricks
Spark
Big data
Debugging
ML infra
Data engineering
LLM/VLM interfaces
Evaluation platforms
Customer-facing products

Tools

Databricks
Spark
AWS

Job description

Perplexity AI is seeking an experienced software engineer to build and maintain the data pipelines powering eval verdicts and product decisions. You will own the flow from raw signals to durable datasets that influence decision-making across the company, and you’ll develop a simulator to replay user interactions for LLMs and VLMs.

You will ensure data trust with monitoring and lineage in a fast, high-impact team.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Member of Technical Staff (Software Engineer, Data Flywheel)
Member of Technical Staff (Software Engineer, Data Flywheel)

Perplexity AI • United States

On-site
USD 120,000 - 190,000
Member of Technical Staff (Data Scientist, Evals)
Member of Technical Staff (Data Scientist, Evals)

Perplexity • United States

Hybrid
USD 100,000 - 150,000
Staff Data Scientist, AI Evaluation & LLM Quality
Staff Data Scientist, AI Evaluation & LLM Quality

Perplexity • United States

Hybrid
USD 100,000 - 150,000
Member of Technical Staff (Software Engineer, Data Flywheel)
Member of Technical Staff (Software Engineer, Data Flywheel)

Pantera Capital • San Francisco (CA)

Hybrid
USD 210,000 - 385,000
AI-Driven Data Systems Engineer
AI-Driven Data Systems Engineer

Perplexity • San Francisco (CA)

On-site
USD 140,000 - 210,000
Member of Technical Staff (Software Engineer, Data Platform)
Member of Technical Staff (Software Engineer, Data Platform)

Perplexity • San Francisco (CA)

On-site
USD 130,000 - 170,000
Senior AI Engineer - Production LLM EvalOps
Senior AI Engineer - Production LLM EvalOps

LawPro.ai • Town of Texas (WI)

On-site
USD 140,000 - 210,000
AI-Powered Data Systems Engineer
AI-Powered Data Systems Engineer

Neura Market • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Staff Software Engineer, Safeguards Evaluations & Trust
Staff Software Engineer, Safeguards Evaluations & Trust

Menlo Ventures • San Francisco (CA)

Hybrid
USD 320,000 - 485,000
Flexible working hours
Generous vacation and parental leave
Office space for collaboration
Staff Engineer, Evaluation Infrastructure
Staff Engineer, Evaluation Infrastructure

Simile • San Francisco (CA)

On-site
USD 200,000 - 400,000
Health & Wellness
Equity
Flexible time off