Remote | AI Agent Power User — $30–$90/hour

24-Mag Llc

Northern (KY)

Hybrid

USD 41,000 - 124,000

Part time

4 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

24-MAG LLC is offering a specialised part-time consulting opportunity for experienced professionals who use AI assistants like ChatGPT and Claude in real-world business workflows to contribute to an advanced AI training and evaluation project.

Selected candidates will run complex, multi-step business scenarios, evaluate AI outputs, document results, and provide structured feedback to support improvements in AI systems. The role is fully remote and contract-based, with flexible timing.

Qualifications

  • Minimum of 5 years in a technology-enabled business role.
  • Bachelor's degree or higher required.
  • Proficient in ChatGPT/Claude for professional work.
  • Experience with evaluation rubrics or QA scoring.
  • Excellent written English and documentation skills.
  • Ability to work independently in remote, asynchronous settings.

Responsibilities

  • Run multi-step professional scenarios with AI assistants.
  • Recreate realistic business workflows across analysis, communication, documentation, research, and operations.
  • Assess AI performance across workflow stages and provide actionable evaluations.
  • Document quality issues and model-performance patterns.
  • Provide clear, constructive feedback and maintain transparent records.

Job description

We are sharing a specialised part-time consulting opportunity for experienced professionals who use AI assistants extensively in real-world business workflows to contribute to an advanced AI training and evaluation project.

Selected professionals will run realistic, multi-step business scenarios using tools such as ChatGPT and Claude, evaluate the quality and usefulness of AI-generated outputs, document model behaviour, and provide structured feedback that can support improvements in advanced AI systems. The work is designed around authentic professional use cases rather than experimental or casual AI usage.

Key Responsibilities
AI-Assisted Business Workflow Evaluation
  • Run complex, multi-step professional scenarios using AI assistants such as ChatGPT and Claude
  • Recreate realistic business workflows involving analysis, communication, documentation, research, and operational tasks
  • Assess how effectively AI systems perform across each stage of a workflow
  • Identify where AI-generated outputs align with or diverge from genuine professional requirements
  • Evaluate whether AI assistance produces useful, complete, and practically actionable results
Rubric-Based Evaluation & Quality Assurance
  • Evaluate and score AI-generated outputs using defined rubrics, quality criteria, and assessment standards
  • Review responses for completeness, relevance, accuracy, clarity, and professional usefulness
  • Apply evaluation standards consistently across varied business scenarios
  • Identify outputs that appear strong at surface level but fail important workflow or user requirements
  • Document quality issues, inconsistencies, and recurring model-performance patterns
Feedback, Documentation & Model Analysis
  • Write clear, constructive, and actionable feedback based on evaluation findings
  • Document each step of tested workflows to support transparency and reproducibility
  • Explain where and why AI performance succeeds or breaks down
  • Identify recurring strengths, weaknesses, and opportunities for process improvement
  • Maintain detailed written records that can be used by AI research and product teams
AI Tools & Workplace Integrations
  • Connect AI assistants with commonly used workplace platforms where appropriate
  • Test workflows involving tools such as Google Drive, Gmail, Slack, Notion, and related productivity systems
  • Evaluate how AI behaves across connected, multi-tool business processes
  • Apply practical knowledge of AI adoption to assess whether workflows reflect realistic professional usage
  • Help identify opportunities to improve the reliability and usability of AI-assisted work
Ideal Profile
  • Minimum of 5 years of professional experience in a business function involving technology-enabled problem solving
  • Completed Bachelor's degree or higher in any discipline
  • Consistent, hands-on use of ChatGPT and/or Claude for real professional work rather than occasional experimentation
  • Strong familiarity with AI-assisted workflows, prompt-based work, and practical business use cases
  • Experience applying, creating, or reviewing evaluation rubrics, QA scorecards, grading criteria, or content-review guidelines
  • Academic evaluation, quality assurance, hiring assessment, annotation, or similar structured-review experience is advantageous
  • Comfortable integrating AI tools with workplace software and productivity platforms
  • Excellent written English with the ability to produce clear, precise, and actionable documentation
  • Strong attention to detail, analytical judgement, and process-improvement skills
  • Ability to work independently within structured remote and asynchronous project environments
  • Based in the United States, Canada, United Kingdom, Ireland, Australia, or New Zealand, with U.S.-based applicants strongly preferred
  • Authorised to undertake contract work in the relevant country
Engagement Details
  • Part-time independent contractor engagement
  • Fully remote across the United States, Canada, United Kingdom, Ireland, Australia and New Zealand
  • Compensation: $30-$90/hour
  • Work will involve AI-assisted business workflow testing, rubric-based evaluation, quality assurance, and structured written feedback
  • Regular professional use of AI assistants is central to this engagement
  • No prior formal experience in AI research or model training is required
  • Project scope, workload, timing, and duration may evolve depending on project requirements
  • Work must be completed without using confidential or proprietary information belonging to any employer, client, institution, or other third party
About the Platform

This opportunity is available through 24-MAG LLC. We connect experienced professionals with remote consulting opportunities across technical, evaluation, and project-based workstreams.

By submitting this application, you acknowledge that your information may be processed by 24-MAG LLC for recruitment and opportunity matching in accordance with our Privacy Policy: https://www.24-mag.com/privacy-policy

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote | AI Evaluation Specialist — $30–$90/hour
Remote | AI Evaluation Specialist — $30–$90/hour

24-Mag Llc • Northern (KY)

Hybrid
USD 41,000 - 124,000
Remote | Product Manager/Product Owner — $90–$140/hour
Remote | Product Manager/Product Owner — $90–$140/hour

24-Mag Llc • New York (NY), Northern (KY)

Hybrid
USD 110,000 - 179,000
AI Agent Workflow Evaluator
AI Agent Workflow Evaluator

OpenTrain AI • Northern (KY)

Hybrid
USD 41,000 - 124,000
Remote | Software Engineer — $90–$140/hour
Remote | Software Engineer — $90–$140/hour

24-Mag Llc • New York (NY), Northern (KY)

Hybrid
USD 124,000 - 193,000
AI Workflow Evaluator — Part-Time, Remote
AI Workflow Evaluator — Part-Time, Remote

24-Mag Llc • Northern (KY)

Hybrid
USD 41,000 - 124,000
AI Evaluation Specialist
AI Evaluation Specialist

micro1 • United States

Remote
AUD 70,000 - 110,000
AI Agent Evaluation Analyst (Freelance)
AI Agent Evaluation Analyst (Freelance)

Mindrift • Alabama

Remote
Flexible working hours
Competitive pay up to $80/hour
Experience in advanced AI projects
Remote | Technical Writer / Editor — $90–$140/hour
Remote | Technical Writer / Editor — $90–$140/hour

24-Mag Llc • Northern (KY)

Hybrid
USD 124,000 - 193,000
Remote contractor
Open to US, Canada, UK
AI Agent Evaluation Analyst
AI Agent Evaluation Analyst

Mindrift • Dallas (TX)

Remote
Flexible remote work
Competitive pay up to $55/hour
Experience in advanced AI projects
AI Agent Evaluation Analyst (Freelance)
AI Agent Evaluation Analyst (Freelance)

Mindrift • Wisconsin

Remote
Competitive pay up to $60/hour
Flexible remote work
Opportunity to work on advanced AI projects