Python Data Scientist — AI Model Training & RL Optimization

Hired

Mexico

On-site

PHP 4,389,000 - 7,524,000

Full time

33 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Hired is seeking a Data Scientist/Analyst to join on a full-time basis to develop and maintain Python code for training AI models and to benchmark performance. You will collaborate with researchers and annotators to implement reinforcement learning with human feedback and refine reward models.

The role emphasizes evaluating model outputs, producing clear explanations of evaluations, and leading supervised fine-tuning with high-quality datasets for various tasks.

Qualifications

  • Proficiency in Python for data analysis and model development is required.
  • Experience with data science and machine learning techniques is required.
  • Ability to extract insights from public data using analytical skills is required.
  • Strong bug-fixing skills and ability to create thorough technical documentation are required.
  • Excellent communication skills to articulate reasoning in notebooks or similar formats are required.

Responsibilities

  • Design, develop, and maintain efficient, high-quality code to train and optimize AI models.
  • Conduct evaluations to benchmark model performance and analyze results for continuous improvement.
  • Evaluate and rank AI model responses to user queries across diverse domains, ensuring alignment with predefined criteria.
  • Develop explanations and rationales for evaluations, showcasing reasoning and technical expertise.
  • Lead efforts in supervised fine-tuning, including creating and maintaining high-quality, task-specific datasets.

Skills

Python
Data analysis
Machine learning
Bug fixing
Technical documentation
Communication

Job description

Hired is seeking a Data Scientist/Analyst to join on a full-time basis to develop and maintain Python code for training AI models and to benchmark performance. You will collaborate with researchers and annotators to implement reinforcement learning with human feedback and refine reward models.

The role emphasizes evaluating model outputs, producing clear explanations of evaluations, and leading supervised fine-tuning with high-quality datasets for various tasks.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI/ML Data Scientist (Python) for RL & Model Evaluation
AI/ML Data Scientist (Python) for RL & Model Evaluation

Hire Feed • España

On-site
PHP 5,643,000 - 8,150,000
Data Scientist - Python (Remote)
Data Scientist - Python (Remote)

Hired • Mexico

On-site
PHP 4,389,000 - 7,524,000
Data Scientist - Python (Remote)
Data Scientist - Python (Remote)

Hire Feed • España

On-site
PHP 5,643,000 - 8,150,000
Python Engineer (Remote)
Python Engineer (Remote)

Hired • Mexico

Hybrid
MXN 900,000 - 1,300,000
Senior AI/LLM Engineer: RLHF & Production Lead
Senior AI/LLM Engineer: RLHF & Production Lead

Innodata Inc. • Philippines

On-site
PHP 2,400,000 - 4,800,000
Python (Programming Language)- AI/ML Developer
Python (Programming Language)- AI/ML Developer

Recruitify_HR • Cebu City

On-site
PHP 1,000,000 - 1,500,000
Python (Programming Language)-AI/ML Developer
Python (Programming Language)-AI/ML Developer

Recruitify_HR • Cebu City

On-site
Python Programming Language
Python Programming Language

Accenture in the Philippines • Cebu City

On-site
PHP 900,000 - 1,200,000
Python Developer
Python Developer

Bravissimo Resourcing Inc. • Cebu City

On-site
PHP 600,000 - 1,200,000
Senior Python AI Engineer
Senior Python AI Engineer

Proxify • Mexico

On-site
MXN 1,380,000 - 1,899,000
Guaranteed on‑time monthly payments
Up to 24 flex days off per year
Career-accelerating opportunities
+1