Remote AI Safety & Evaluations Engineer

DeWinter Group

Campbell (CA)

Remote

USD 68,880 - 241,080

Part time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading AI solutions firm is seeking an AI Safety and Evaluations Engineer for a 12-month contract, focusing on designing evaluation frameworks to ensure AI models are bias-free and compliant. The candidate will create automated datasets and develop specific metrics in RAG-based systems. With a requirement of 3+ years in AI Research or Quality Engineering, strong communication skills, and autonomy, this role is integral to maintaining the safety of AI models. This is a remote opportunity with a pay range of $50/hr to $175/hr.

Qualifications

  • 3+ years of experience in AI Research or Quality Engineering.
  • Deep expertise in model evaluation techniques and NLP metrics.
  • Demonstrated ability to work autonomously and manage time effectively.
  • Experience with Python, data analysis tools, and LLM-as-a-Judge frameworks.
  • Strong communication skills for team updates.

Responsibilities

  • Design and build evaluation frameworks for model bias.
  • Create automated datasets to benchmark models before production.
  • Develop metrics for 'Grounding' and 'Faithfulness' in RAG-based systems.
  • Build monitoring tools for harmful AI outputs.
  • Partner with legal and ethics teams for safety constraints.

Skills

AI Research
Quality Engineering
Model evaluation techniques
NLP metrics (ROUGE, BLEU, BERTScore)
Python
Data analysis tools
Self-motivated
Communication skills

Job description

A leading AI solutions firm is seeking an AI Safety and Evaluations Engineer for a 12-month contract, focusing on designing evaluation frameworks to ensure AI models are bias-free and compliant. The candidate will create automated datasets and develop specific metrics in RAG-based systems. With a requirement of 3+ years in AI Research or Quality Engineering, strong communication skills, and autonomy, this role is integral to maintaining the safety of AI models. This is a remote opportunity with a pay range of $50/hr to $175/hr.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Safety & Evaluation Engineer
AI Safety & Evaluation Engineer

DeWinter Group • Campbell (CA)

Remote
USD 68,880 - 241,080
Remote AI Evaluation Engineer: Safeguard & Scale
Remote AI Evaluation Engineer: Safeguard & Scale

DeepRec.ai • Denver (CO)

Remote
USD 180,000
Remote AI Safety Evaluator — Expert Policy & Quality Review
Remote AI Safety Evaluator — Expert Policy & Quality Review

Mercor • Town of Belgium (WI)

On-site
Remote AI Safety Specialist — Expert Evaluator
Remote AI Safety Specialist — Expert Evaluator

Mercor • United States

On-site
AI Risk & Fraud Evaluation Engineer
AI Risk & Fraud Evaluation Engineer

Variance • San Francisco (CA)

On-site
USD 170,000 - 230,000
Competitive salary
Platinum-level medical, dental, and vision insurance
Unlimited PTO
+2
Senior AI Safety Research Engineer
Senior AI Safety Research Engineer

Aisafety • Berkeley (CA)

Hybrid
USD 150,000 - 250,000
Catered lunch and dinner
Visa sponsorship for in-person employees
Work-related travel expenses covered
Remote AI Safety Evaluator (Contract)
Remote AI Safety Evaluator (Contract)

Mercor • United States

On-site
USD 83,000 - 96,000
Remote AI Safety & Biosecurity Software Engineer
Remote AI Safety & Biosecurity Software Engineer

SecureBio, LLC • Boston (MA)

Remote
USD 70,000 - 115,000
Unlimited paid time off
Flexible work hours
Conference sponsorships
+1
AI Safety Specialist - Fully Remote | Upto $70/hr
AI Safety Specialist - Fully Remote | Upto $70/hr

Obsidian • San Francisco (CA)

Remote
USD 120,000 - 180,000
Remote AI Safety Red Team Engineer - Part Time
Remote AI Safety Red Team Engineer - Part Time

Mindrift • San Antonio (TX)

Remote
USD 60,000 - 80,000