Insurance AI Evaluation Specialist

Obsidian

San Francisco (CA)

Remote

USD 120,000 - 180,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Cincinnatus LLC is seeking an Insurance SME to join a leading AI lab's GenAI team in San Francisco. You will evaluate AI outputs against rubrics and contribute real-world underwriting judgment to training data for foundational AI models.

The role requires 8+ years in insurance, strong communication, and reliable 35-hour weekday availability as part of a W-2 engagement with potential placement at a premier AI lab.

Qualifications

  • 8+ years of dedicated professional experience in insurance (underwriting, claims, actuarial, risk management) at a top-tier organization.
  • Hands-on experience evaluating LLM/AI model outputs against rubrics or structured scoring criteria.
  • Demonstrated career progression (e.g., Underwriter → Senior Underwriter → VP of Underwriting).
  • Ability to engage reliably for at least 35 hours/week during weekdays.

Responsibilities

  • Guide research and engineering teams to close knowledge gaps in underwriting, claims, and risk-assessment reasoning.
  • Design challenging, domain-relevant insurance tasks and write accurate, well-reasoned solutions grounded in real underwriting/claims practice.
  • Evaluate AI model outputs against structured rubrics and provide clear, written feedback on correctness, judgment, and reasoning quality.
  • Develop and refine evaluation guidelines and scoring rubrics specific to insurance tasks.
  • Collaborate with other subject matter experts to ensure consistency and accuracy in training data.

Skills

Insurance expertise
LLM evaluation experience
Communication skills
Availability 35h/wk

Job description

Cincinnatus LLC is seeking an Insurance SME to join a leading AI lab's GenAI team in San Francisco. You will evaluate AI outputs against rubrics and contribute real-world underwriting judgment to training data for foundational AI models.

The role requires 8+ years in insurance, strong communication, and reliable 35-hour weekday availability as part of a W-2 engagement with potential placement at a premier AI lab.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Insurance AI Evaluation Specialist for GenAI Training
Insurance AI Evaluation Specialist for GenAI Training

Mercor • San Francisco (CA)

On-site
USD 150,000 - 190,000
Senior Insurance AI Evaluation & Underwriting Expert
Senior Insurance AI Evaluation & Underwriting Expert

Mercor • United States

On-site
USD 140,000 - 200,000
GenAI Insurance SME — Underwriting & Risk for AI Training
GenAI Insurance SME — Underwriting & Risk for AI Training

Obsidian • San Francisco (CA)

On-site
USD 180,000 - 240,000
AI GenAI Insurance SME - Underwriting & Model Evaluation
AI GenAI Insurance SME - Underwriting & Model Evaluation

Mercor • New York (NY)

On-site
USD 140,000 - 190,000
AI Insurance Underwriter & Model Evaluation Expert
AI Insurance Underwriter & Model Evaluation Expert

Mercor • New York (NY)

Remote
USD 120,000 - 180,000
GenAI Insurance SME — Risk & Underwriting Evaluation
GenAI Insurance SME — Risk & Underwriting Evaluation

Obsidian • New York (NY)

On-site
USD 140,000 - 220,000
AI GenAI Insurance Underwriter & Model Evaluator
AI GenAI Insurance Underwriter & Model Evaluator

Obsidian • New York (NY)

Remote
USD 120,000 - 180,000
Senior Insurance AI Domain Expert
Senior Insurance AI Domain Expert

DigiNo • Northern (KY)

Hybrid
USD 190,000 - 260,000
Insurance SME - Underwriting Expert
Insurance SME - Underwriting Expert

Mercor • San Francisco (CA)

On-site
USD 150,000 - 190,000
Insurance Specialist
Insurance Specialist

Mercor • United States

On-site
USD 140,000 - 200,000