Machine Learning & NLP Expert

Weekday AI

United States

Remote

USD 110,000 - 152,000

Part time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Fully remote
Weekly payments via Stripe or Wise

Job summary

Weekday AI is seeking experienced Machine Learning & NLP Experts for a part-time, fully remote role. You will design and evaluate ML/NLP tasks, create reference solutions, and assess model performance to improve capabilities.

This independent contractor position requires about 20 hours per week with weekly payments. The role focuses on building rigorous evaluation tasks, leveraging Python and modern ML frameworks, and collaborating with a team of subject matter experts.

Qualifications

  • Deep hands-on experience in ML/NLP through industry, research, or graduate work.
  • Strong Python skills with ML/NLP development experience.
  • Knowledge of model training, evaluation, transformers, and NLP pipelines.
  • Experience with PyTorch, TensorFlow, or HuggingFace Transformers.
  • Ability to commit ~20 hours per week remotely.

Responsibilities

  • Design real-world ML and NLP tasks across ML model development, NLU, NLG, IR, and LLMs.
  • Develop reference solutions and integrate tasks into evaluation environments using Python.
  • Build evaluation frameworks and assess model outputs for correctness.
  • Identify capability gaps and provide detailed analyses.
  • Create evaluation guidelines, rubrics, and quality standards.
  • Collaborate with SMEs to ensure data quality.

Skills

Machine Learning
Natural Language Processing
Python
Model Training & Evaluation
Transformer Architectures
LLMs

Education

PhD or equivalent

Tools

PyTorch
TensorFlow
Hugging Face Transformers

Job description

This role is for one of our clients

Compensation: $80-$110 per hour

Join a cutting-edge AI research initiative and help shape the next generation of frontier AI models. We are seeking experienced Machine Learning & NLP Experts to contribute their technical expertise toward training and evaluating advanced AI systems. In this role, you will design challenging, real-world machine learning and natural language processing tasks, create high-quality reference solutions, and assess AI model performance to identify reasoning gaps and improve overall capabilities.

This is a part-time, fully remote opportunity requiring approximately 20 hours per week.

Key Responsibilities
  • Design challenging, real-world machine learning and natural language processing tasks covering areas such as:
    • Machine Learning Model Development and Evaluation
    • Natural Language Understanding (NLU)
    • Natural Language Generation (NLG)
    • Information Retrieval and Search
    • Applied Machine Learning Pipelines
    • Transformer Models and Large Language Models (LLMs)
  • Develop accurate reference solutions and integrate tasks into agentic development environments using Python.
  • Build executable evaluation frameworks and testing components where appropriate.
  • Evaluate AI model outputs for technical correctness, reasoning quality, and overall performance.
  • Identify capability gaps, classify model failure modes, and provide detailed written analyses.
  • Create and refine evaluation guidelines, scoring rubrics, and quality standards for ML and NLP tasks.
  • Collaborate with fellow subject matter experts to ensure consistency, accuracy, and high-quality training data.
Required Qualifications
  • Deep hands-on experience in Machine Learning and/or Natural Language Processing through industry, research, or graduate/PhD-level work.
  • Strong proficiency in Python with practical experience developing ML or NLP applications.
  • Strong understanding of modern machine learning techniques, including:
    • Model Training and Evaluation
    • Transformer Architectures
    • Large Language Models (LLMs)
    • NLP Pipelines
    • Feature Engineering and Model Optimization
  • Experience with industry-standard frameworks such as PyTorch, TensorFlow, Hugging Face Transformers, or equivalent.
  • Ability to commit approximately 20 hours per week.
  • Excellent written communication skills and the ability to work independently in a remote environment.
Preferred Qualifications
  • Experience in AI model evaluation, AI training data creation, or human-in-the-loop model assessment.
  • Familiarity with Retrieval-Augmented Generation (RAG), vector databases, embedding models, or multimodal AI systems.
  • Experience building benchmarking frameworks, automated evaluation pipelines, or testing infrastructure.
  • Contributions to open-source ML/NLP projects or published research are a plus.
  • Experience working with production-scale machine learning systems.
Role Details
  • Employment Type: Independent Contractor
  • Work Arrangement: Fully Remote
  • Schedule: Approximately 20 hours per week
  • Project Duration: Based on project requirements and performance, with opportunities for extension
Equal Opportunity

We consider all qualified applicants without regard to legally protected characteristics and provide reasonable accommodations upon request.

Contract & Payment Terms
  • You will be engaged as an independent contractor.
  • This is a fully remote role that can be completed on your own schedule.
  • Projects may be extended, shortened, or concluded early depending on business needs and performance.
  • Your work will not involve access to confidential or proprietary information from any employer, client, or institution.
  • Payments are made weekly via Stripe or Wise based on services rendered.
  • Please note: We are unable to support H1-B or STEM OPT candidates at this time.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Machine Learning Engineer - Model Evaluation & Experimentation
Machine Learning Engineer - Model Evaluation & Experimentation

Weekday AI • United States

Remote
USD 83,000 - 124,000
Remote | ML Engineer — $100–$150/hour
Remote | ML Engineer — $100–$150/hour

24-Mag Llc • New York (NY), Northern (KY)

Hybrid
USD 138,000 - 207,000
Remote
Remote

24-MAG • New York (NY)

Remote
USD 138,000 - 207,000
MLOps Engineer (JAX, PyTorch, Pallas/Triton)
MLOps Engineer (JAX, PyTorch, Pallas/Triton)

Weekday AI • United States

Remote
USD 96,000 - 152,000
Remote ML & NLP Expert — Part-Time (20h/w)
Remote ML & NLP Expert — Part-Time (20h/w)

Weekday AI • United States

Remote
USD 110,000 - 152,000
Fully remote
Weekly payments via Stripe or Wise
ML Engineer Specialist - Freelance AI Trainer Project
ML Engineer Specialist - Freelance AI Trainer Project

Meridial • United States

On-site
USD 41,328 - 68,880
Flexibility to set your schedule
Remote work environment
Impact on cutting-edge AI
Remote | Machine Learning Engineer — $80–$140/hour
Remote | Machine Learning Engineer — $80–$140/hour

24-Mag Llc • Northern (KY), New York (NY)

Hybrid
USD 110,000 - 193,000
Remote | Machine Learning Engineer — $80–$140/hour
Remote | Machine Learning Engineer — $80–$140/hour

24-MAG • United States

Remote
USD 110,000 - 193,000
LLM Red Team Specialist - Failure Modes & Edge Cases
LLM Red Team Specialist - Failure Modes & Edge Cases

Weekday AI • United States

Remote
USD 171,924,000 - 257,887,000
Machine Learning Developer (Freelance)
Machine Learning Developer (Freelance)

Mindrift • Missouri

On-site
USD 74,000 - 124,000