AI Quality Engineer — Model Behavior & Evaluation

Monograph

New York (NY)

Hybrid

USD 98,000 - 140,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading technology company in New York is seeking a candidate to take ownership of quality for AI products. This role blends strategy and operations, focusing on context engineering and enhancing AI product behavior. You'll work closely with data, product, and engineering teams to define quality metrics, evaluate models, and shape strategies for AI implementation. Ideal candidates are proactive problem-solvers with experience in AI or LLMs. The position offers a hybrid work environment and a competitive salary range of $98,000 - $140,000 per year.

Qualifications

  • Drive ownership of AI product quality and initiatives.
  • Explore LLM capabilities and AI's real-world applications.
  • Analyze data to inform AI product improvements.

Responsibilities

  • Design and test criteria for how AI products behave.
  • Engage with production data to identify and debug issues.
  • Develop evaluation strategies for quality measurement.
  • Collaborate with leading AI labs to launch new models.
  • Shape quality narratives with engineering and product teams.
  • Build tools that enhance AI product development.

Skills

Driver mentality
Curiosity
Analytical instinct
Comfortable working with data
Clear communication
Experience with LLMs

Job description

A leading technology company in New York is seeking a candidate to take ownership of quality for AI products. This role blends strategy and operations, focusing on context engineering and enhancing AI product behavior. You'll work closely with data, product, and engineering teams to define quality metrics, evaluate models, and shape strategies for AI implementation. Ideal candidates are proactive problem-solvers with experience in AI or LLMs. The position offers a hybrid work environment and a competitive salary range of $98,000 - $140,000 per year.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Model Behavior Engineer—Quality & Evaluation
AI Model Behavior Engineer—Quality & Evaluation

Notion • San Francisco (CA)

On-site
USD 98,000 - 140,000
AI Quality & Operations Lead
AI Quality & Operations Lead

Welo Data • New York (NY)

On-site
USD <104,000
Free breakfast, lunch, and dinner
Comprehensive Medical, Dental, and Vision coverage
15 days of Paid Sick/Holiday time
+2
AI Quality & Behavior Engineer
AI Quality & Behavior Engineer

Hack Chicago • San Francisco (CA)

On-site
USD 98,000 - 140,000
AI Quality Engineer: End-to-End ML & Data Validation
AI Quality Engineer: End-to-End ML & Data Validation

Cavendish Professionals • Town of Italy (NY)

On-site
USD 95,000 - 120,000
AI Operations Lead, Human-in-the-Loop
AI Operations Lead, Human-in-the-Loop

Welo Data • Washington

On-site
USD <85,000
Free gourmet dining
Comprehensive Medical, Dental, and Vision
15 days of combined Paid Sick/Holiday time
+2
AI Quality & Operations Lead, Human-in-the-Loop
AI Quality & Operations Lead, Human-in-the-Loop

Welo Data • Washington

On-site
USD <86,000
AI Quality Analyst — Hybrid Work & Bias-Resistant AI Testing
AI Quality Analyst — Hybrid Work & Bias-Resistant AI Testing

Zebra Technologies • Lincolnshire (IL)

On-site
USD 122,000 - 185,000
Hybrid work
Flexible hours
Paid time off
+1
AI Quality & Evaluation Scientist for Clinical LLMs
AI Quality & Evaluation Scientist for Clinical LLMs

Bioscope.ai, Inc. • Boston (MA)

On-site
USD 100,000 - 130,000
AI Response Quality Analyst
AI Response Quality Analyst

Crossing Hurdles • United States

On-site
USD 60,000 - 80,000
AI Safety & Data Quality Analyst — Onsite NYC Impact
AI Safety & Data Quality Analyst — Onsite NYC Impact

Welo Data • New York (NY)

On-site
USD <63,000
Gourmet Dining
Endless Snacks
Campus Life
+3