Remote RLHF Specialist: AI Alignment & Evaluation

Odixcity Consulting

South Africa

Remote

ZAR 1,477,000 - 2,626,000

Full time

2 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Odixcity Consulting is seeking an RLHF Specialist to improve AI models through reinforcement learning from human feedback, focusing on data collection, evaluation, and alignment across a global team.

You will design prompts, create annotated datasets, and collaborate with ML engineers to enhance model safety, accuracy, and robust reasoning. Remote worldwide, full-time position with access to modern ML tooling.

Qualifications

  • Minimum of 2 years of experience in Data Annotation, Model Evaluation, Computational Linguistics, or Trust and Safety, specifically working with AI/ML training data.
  • Strong proficiency in Python and deep learning frameworks (PyTorch, JAX, or TensorFlow).
  • Deep understanding of Reinforcement Learning concepts (PPO, Trust Regions, Reward Hacking) and how they apply to language generation.
  • Hands-on experience fine-tuning open-source models (e.g., Llama 2/3, Mistral, gemma) using techniques like LoRA/QLoRA.
  • Experience working with annotation tools (LabelBox, Scale AI, Snorkel) and managing human-in-the-loop workflows.
  • Ability to diagnose why an RL policy collapsed and adjust hyperparameters or reward structure accordingly.
  • Experience with Constitutional AI or Self-Alignment techniques.
  • Contributions to open-source alignment libraries (TRL, Transformer Reinforcement Learning, Axolotl).
  • Experience with cloud Platforms (AWS SageMaker, GCP Vertex AI).

Responsibilities

  • Generate high-quality preference data by comparing multiple model responses and ranking them based on criteria such as helpfulness, honesty, and harmlessness (HHH).
  • Design complex, multi-turn prompts to stress-test model behavior and expose weaknesses in reasoning or safety.
  • Write detailed “chain-of-thought” explanations and rationales to train reward models on why specific responses are superior.
  • Collaborate with Machine Learning Engineers to analyze model failure modes and identify data gaps that, when filled, will improve reinforcement learning outcomes.
  • Develop and iterate on annotation strategies for preference scoring and reinforcement signals, ensuring consistency across a global team.
  • Proactively probe models to identify vulnerabilities, biases, or hallucination patterns, documenting findings for model optimization.
  • Analyze edge cases where the reward model behaves unexpectedly (e.g., over-indexing on verbosity or style over substance). Provide detailed feedback to ML engineers on reward model failure modes and suggest specific data interventions to correct model behavior.
  • Develop and document templated instruction sets for larger annotation teams. Translate complex reinforcement learning concepts into simple, repeatable tasks for junior reviewers, ensuring high-quality data collection at scale.
  • Monitor model performance over time by maintaining a personal test set of prompts. Regularly re-evaluate new model versions against historical benchmarks to track improvements or regressions in reasoning and alignment.

Skills

Python
Deep Learning
PyTorch
JAX
TensorFlow
Reinforcement Learning
PPO
LoRA/QLoRA
Open-source models
Annotation tools

Tools

LabelBox
Scale AI
Snorkel
AWS SageMaker
GCP Vertex AI

Job description

Odixcity Consulting is seeking an RLHF Specialist to improve AI models through reinforcement learning from human feedback, focusing on data collection, evaluation, and alignment across a global team.

You will design prompts, create annotated datasets, and collaborate with ML engineers to enhance model safety, accuracy, and robust reasoning. Remote worldwide, full-time position with access to modern ML tooling.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

RLHF Specialist
RLHF Specialist

Odixcity Consulting • South Africa

Remote
ZAR 1,477,000 - 2,626,000
Remote Computational Linguist & Annotation Specialist
Remote Computational Linguist & Annotation Specialist

Odixcity Consulting • South Africa

Remote
ZAR 350,000 - 550,000
Remote Environmental AI Evaluation Specialist
Remote Environmental AI Evaluation Specialist

Alignerr Corp. • Cape Town

Remote
ZAR 683,000 - 1,366,000
Fully remote
Flexible schedule
Global collaboration
Remote AI Code Annotator & Reviewer
Remote AI Code Annotator & Reviewer

Odixcity Consulting • South Africa

Remote
ZAR 300,000 - 600,000
Remote AI Research Scientist (ML/DL, Equity Eligible)
Remote AI Research Scientist (ML/DL, Equity Eligible)

Placements24 • Centurion

Hybrid
ZAR 1,479,000 - 2,465,000
Equity options
Health insurance
Home office allowance
+2
Remote AI Ethics Specialist
Remote AI Ethics Specialist

Placements24 • Mtubatuba Local Municipality

Hybrid
ZAR 600,000 - 900,000
Remote work stipend
Healthcare coverage
Paid time off
+2
Remote Ruby Engineer for AI Training and Code Evaluation
Remote Ruby Engineer for AI Training and Code Evaluation

Alignerr Corp. • South Africa

On-site
ZAR 904,000 - 1,809,000
Senior AI Engineer
Senior AI Engineer

Placements24 • Rustenburg

Hybrid
ZAR 1,000,000 - 1,600,000
Competitive salary
Remote flexibility
Learning opportunities
+2
Remote AI Ethicist
Remote AI Ethicist

Placements24 • Mtubatuba Local Municipality

Hybrid
ZAR 80,000 - 120,000
Professional development
Conference attendance
Remote-friendly culture
+1
Remote AI Infrastructure Security Analyst
Remote AI Infrastructure Security Analyst

Alignerr Corp. • Johannesburg

Remote
ZAR 917,000 - 1,605,000
Autonomy
Global collaboration
Flexible schedule