AI Risk & Fraud Evaluation Engineer

Variance

San Francisco (CA)

On-site

USD 170,000 - 230,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive salary
Platinum-level medical, dental, and vision insurance
Unlimited PTO
$100 reimbursement for wellness expenses
401(k) plan

Job summary

A technology firm in San Francisco is seeking a Research Engineer to enhance AI model quality. The ideal candidate will build benchmarks, datasets, and evaluation loops to ensure effective performance on critical tasks. This role requires strong programming skills and a passion for developing rigorous evaluations in high-stakes environments. A commitment to craftsmanship and deep understanding of ML systems are essential. The company offers competitive salary and robust benefits.

Qualifications

  • Experience training, evaluating, or improving modern ML systems.
  • Strong programming skills and comfort working in research-heavy codebases.
  • Experience building benchmarks, datasets, evaluation pipelines, or quality systems.

Responsibilities

  • Build proprietary benchmarks and datasets to evaluate models.
  • Design and run evaluations that measure model performance.
  • Define quality metrics for judgment systems.

Skills

Strong programming skills
Experience with ML systems
Engineering judgment
Interest in fraud and risk

Job description

A technology firm in San Francisco is seeking a Research Engineer to enhance AI model quality. The ideal candidate will build benchmarks, datasets, and evaluation loops to ensure effective performance on critical tasks. This role requires strong programming skills and a passion for developing rigorous evaluations in high-stakes environments. A commitment to craftsmanship and deep understanding of ML systems are essential. The company offers competitive salary and robust benefits.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Benchmarking Engineer — Evaluations & Failure Analysis
AI Benchmarking Engineer — Evaluations & Failure Analysis

Mercor • San Francisco (CA)

On-site
USD 120,000 - 160,000
Generous equity grant vested over 4 years
$10K housing bonus
$1.5K monthly stipend for meals
+2
ML Safety & Benchmarking Research Engineer
ML Safety & Benchmarking Research Engineer

Apple Inc. • San Francisco (CA)

On-site
USD 181,000 - 273,000
Comprehensive medical and dental coverage
Retirement benefits
Employee stock purchase plan
+1
AI Evaluation Lead: Real-World Systems Benchmarking
AI Evaluation Lead: Real-World Systems Benchmarking

SupportFinity™ • San Francisco (CA)

On-site
USD 150,000 - 230,000
Remote AI Safety & Evaluations Engineer
Remote AI Safety & Evaluations Engineer

DeWinter Group • Campbell (CA)

Remote
USD 68,880 - 241,080
ML Evaluation Engineer: Benchmark & Model Quality
ML Evaluation Engineer: Benchmark & Model Quality

Reducto • San Francisco (CA)

On-site
USD 100,000 - 130,000
Unlimited PTO
Daily free lunch
Reimbursed transportation
+3
AI Engineer - Applied Research for Secure AI
AI Engineer - Applied Research for Secure AI

Corridor • San Francisco (CA)

On-site
USD 120,000 - 150,000
Remote AI Engineer, Quality & Evaluation at Enterprise Scale
Remote AI Engineer, Quality & Evaluation at Enterprise Scale

Fieldguide • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Competitive compensation packages
Flexible PTO
401k
+3
AI Model Behavior Engineer—Quality & Evaluation
AI Model Behavior Engineer—Quality & Evaluation

Notion • San Francisco (CA)

On-site
USD 98,000 - 140,000
AI Model Evaluation Engineer — Benchmarking & Validation
AI Model Evaluation Engineer — Benchmarking & Validation

SpreeAI • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Frontier AI Risk Evaluation Scientist
Frontier AI Risk Evaluation Scientist

Scale AI • New York (NY), San Francisco (CA)

On-site
USD 197,000 - 247,000
Comprehensive health insurance
Retirement benefits
Generous PTO
+1