Remote Senior AI/ML Engineer — LLM Evaluation & Audits

prolificacademicltd

United States

Remote

USD 82,656,000 - 110,208,000

Part time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Remote, flexible contributor model
Paid hourly tasks up to $80/hr

Job summary

prolificacademic ltd is seeking a Senior AI/ML Engineer to help train and evaluate large language models on a remote, project-based basis. You will provide high-quality technical judgement to research teams building next-generation AI systems, with focus on reliability, efficiency, and safety.

Responsibilities include auditing model code and RLHF workflows, reviewing explanations, and benchmarking outputs against technical taxonomies, while collaborating with researchers across locations.

Qualifications

  • BS, MS, or PhD in Computer Science, Artificial Intelligence, Robotics, or a related quantitative field with a machine learning focus.
  • Hands-on experience building, deploying, or fine-tuning ML models in production environments.
  • Professional-level understanding of neural network architectures (Transformers, CNNs, RNNs) and optimization techniques.
  • Practical experience with prompt engineering, RLHF, or retrieval-augmented generation (RAG) workflows.
  • Ability to audit complex model logic, detect training data contamination, and evaluate mathematical proofs.
  • Sharp attention to detail for spotting hallucinations and biased or flawed AI-generated content.

Responsibilities

  • Review AI-generated explanations of model architectures, loss functions, and backpropagation for technical accuracy.
  • Audit machine learning code, including training loops, data preprocessing scripts, and evaluation notebooks, for correctness and efficiency.
  • Supply structured human feedback used to refine reinforcement learning from human feedback (RLHF) pipelines.
  • Critically assess chain-of-thought reasoning, flagging hallucinations, biases, and logical breakdowns in model outputs.
  • Run comparative benchmarks across model outputs against defined technical taxonomies and performance metrics.

Skills

ML model production
Prompt engineering
RLHF workflows
Code audit
Model evaluation

Education

BS/MS/PhD in CS/AI/Robotics
PhD preferred

Tools

PyTorch
TensorFlow/Keras
Hugging Face Transformers
AWS SageMaker
Vertex AI
Weights & Biases
LangChain
Pinecone
Milvus
Weaviate

Job description

prolificacademic ltd is seeking a Senior AI/ML Engineer to help train and evaluate large language models on a remote, project-based basis. You will provide high-quality technical judgement to research teams building next-generation AI systems, with focus on reliability, efficiency, and safety.

Responsibilities include auditing model code and RLHF workflows, reviewing explanations, and benchmarking outputs against technical taxonomies, while collaborating with researchers across locations.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Remote Senior AI/ML Auditor & Evaluator
Remote Senior AI/ML Auditor & Evaluator

prolificacademicltd • United States

Remote
USD 66,000 - 110,000
Remote work
Flexible scheduling
Senior AI/ML Engineer - LLM Training & Evaluation (Remote)
Senior AI/ML Engineer - LLM Training & Evaluation (Remote)

prolificacademicltd • United States

Remote
USD 69,000 - 110,000
Hourly pay up to $80
Fully remote
Flexible hours
+1
Senior AI Engineer — Remote ML Audits & RLHF Expert
Senior AI Engineer — Remote ML Audits & RLHF Expert

prolificacademicltd • United States

Remote
USD 83,000 - 110,000
Fully remote
Flexible hours
Skills assessment
Senior AI Engineer - Remote Research & RLHF Evaluation
Senior AI Engineer - Remote Research & RLHF Evaluation

prolificacademicltd • United States

Remote
USD 66,000 - 110,000
Remote work
Flexible hours
Short onboarding
Senior AI Engineer - LLM Training & Benchmarking (Remote)
Senior AI Engineer - LLM Training & Benchmarking (Remote)

prolificacademicltd • United States

Remote
USD 66,000 - 110,000
Senior AI/ML Research Engineer (Remote, Flexible Tasks)
Senior AI/ML Research Engineer (Remote, Flexible Tasks)

prolificacademicltd • United States

Remote
USD 66,000 - 110,000
Fully remote engagement
Flexible scheduling
Skills assessment before paid tasks
Remote AI Model Auditor & Evaluator (LLMs, RLHF)
Remote AI Model Auditor & Evaluator (LLMs, RLHF)

prolificacademicltd • United States

Remote
USD 83,000 - 110,000
Fully remote
Flexible scheduling
Short, focused sessions
+1
Senior ML Auditor for RLHF & Model Alignment (Remote)
Senior ML Auditor for RLHF & Model Alignment (Remote)

prolificacademicltd • United States

Remote
USD 66,000 - 110,000
Remote work
Senior AI/ML Auditor for Frontier LLMs (Remote)
Senior AI/ML Auditor for Frontier LLMs (Remote)

prolificacademicltd • United States

Remote
USD 66,000 - 110,000
Remote work
Flexible hours
Technical assessment
Remote Senior AI/ML Engineer – LLM Training & Evaluation
Remote Senior AI/ML Engineer – LLM Training & Evaluation

Prolific • Milwaukee (WI)

On-site
USD 100,000 - 130,000
Competitive pay rates
Flexible hours
Ability to work from home