AI Model Evaluation Data Scientist

OpenTrain AI, Inc.

United States

Remote

USD 34,000 - 55,000

Part time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

OpenTrain AI, Inc. is seeking an entry-level contractor to join our AI training program. You will design, implement, and evaluate Python-based data analysis workflows, and create datasets and clear explanations of results.

You will collaborate with researchers and annotators on supervised fine-tuning, RLHF, and reward-model refinement. Remote work, global applicants, with 20–40 hours weekly and overlap with Pacific Time.

Qualifications

  • Bachelor's or Master's degree in engineering or computer science (or equivalent).
  • Entry-level experience in Python and data analysis.
  • Strong problem-solving ability and sound business judgment.

Responsibilities

  • Design, develop, and maintain Python code for training and optimizing AI models.
  • Benchmark model performance and analyze evaluation results.
  • Evaluate and rank model responses to user queries using set criteria.
  • Write clear explanations and rationales for evaluation decisions.
  • Create and maintain datasets for supervised fine-tuning.
  • Create and refine responses for clarity, relevance, and technical accuracy.
  • Review code and documentation and provide constructive feedback.
  • Use public datasets from sources such as Kaggle, the United Nations, and the US government.

Skills

Python
Data analysis
Problem-solving
Business judgment
Communication

Education

Bachelor's or Master’s in Engineering/CS

Tools

Jupyter notebooks

Job description

The work

You will use Python and data analysis to evaluate AI models and identify ways to improve their performance. The work combines technical analysis, response review, dataset creation, and clear written reasoning.

You will work with researchers and annotators on supervised fine-tuning, reinforcement learning with human feedback, and reward-model refinement. You will also use public datasets to answer business questions and review code and documentation for issues.

  • Design, develop, and maintain Python code for training and optimizing AI models.
  • Benchmark model performance and analyze evaluation results.
  • Evaluate and rank model responses to user queries using set criteria.
  • Write clear explanations and rationales for evaluation decisions.
  • Create and maintain datasets for supervised fine-tuning.
  • Create and refine responses for clarity, relevance, and technical accuracy.
  • Review code and documentation and provide constructive feedback.
  • Use public datasets from sources such as Kaggle, the United Nations, and the US government.
What it pays and takes

The listing does not specify a pay rate. This is a remote contractor assignment with a one-month contract term.

  • Hours: At least 20 hours per week, with options for 20, 30, or 40 hours per week.
  • Schedule: Four hours of daily overlap with Pacific Time.
  • Location: Fully remote and open worldwide.
  • Language: Fluent conversational and written English.
  • Education: A bachelor's or master's degree in engineering or computer science, or equivalent experience.
  • Experience level: Entry level.
  • Requirements: Proficiency in Python, strong data analysis skills, problem-solving ability, and sound business judgment.
  • Communication: Explain technical reasoning clearly in Jupyter notebooks or comparable formats and communicate effectively with researchers and other stakeholders.
About AI training work

OpenTrain is the hiring and contracting organization for this role. AI training work uses human examples and feedback to improve how artificial intelligence systems respond, and people with data science and technical skills help assess model quality and guide those improvements.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Remote AI Training Lead - Data Quality & Evaluation
Remote AI Training Lead - Data Quality & Evaluation

YO IT Consulting • Atlanta (GA)

Remote
USD 41,328,000 - 82,656,000
AI Trainer - Data Science
AI Trainer - Data Science

Planet Pharma • United States

Remote
GBP 12,398,000 - 33,062,000
Social Scenario AI Evaluation Generalist
Social Scenario AI Evaluation Generalist

OpenTrain AI, Inc. • United States

Remote
USD 34,000 - 48,000
AI Training Experts (Freelance - Cardiff)
AI Training Experts (Freelance - Cardiff)

prolificacademicltd • United States

Remote
USD 28,000 - 34,000
Fully remote
Flexible hours
AI Training Experts (Freelance - Leeds)
AI Training Experts (Freelance - Leeds)

prolificacademicltd • United States

Remote
USD 28,000 - 34,000
Fully remote
Flexible hours
Onboarding after assessment
+1
AI Training Experts (Freelance - Glasgow)
AI Training Experts (Freelance - Glasgow)

prolificacademicltd • United States

Remote
USD 28,000 - 40,000
Fully remote
Flexible hours
Self-employed arrangement
AI Training Experts (Freelance - Edinburgh)
AI Training Experts (Freelance - Edinburgh)

prolificacademicltd • United States

Remote
USD 21,000 - 34,000
Fully remote
Flexible hours
Competitive pay
+1
Fantasy Sports AI Evaluation Expert
Fantasy Sports AI Evaluation Expert

OpenTrain AI, Inc. • United States

Remote
USD 34,000 - 62,000
Code Quality Engineer for LLM Evaluation
Code Quality Engineer for LLM Evaluation

OpenTrain AI, Inc. • United States

Remote
USD 34,000 - 69,000
AI Training Experts (Freelance - London)
AI Training Experts (Freelance - London)

prolificacademicltd • United States

Remote
USD 28,000 - 34,000
Fully remote
Flexible hours
Quick onboarding
+2