An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Gramian Consulting Group is seeking an AI Evaluation & Quality Assurance Specialist to review task quality, instructions, and scoring rubrics for AI agents. You will inspect reference solutions, grading logic, execution traces, and generated deliverables to identify defects and unfair penalties.
Comfortable reading Python, SQL, and shell scripts; passable to highly accurate analysis of technical workflows and evaluation logic is required.
Gramian Consultancy is a boutique consultancy specializing in IT professional services and engineering talent solutions. With a strong background in software engineering and leadership, we help companies build high-performing teams by matching them with professionals who truly fit their needs.
We are seeking a detail-oriented AI Evaluation & Quality Assurance Specialist to review the quality, correctness, and fairness of tasks designed to evaluate AI agents. You will inspect reference solutions, grading logic, execution traces, and generated deliverables to identify task defects, evaluation errors, and unjustified model failures. This role requires strong technical fluency, independent analytical judgment, and the ability to produce clear, evidence-based feedback.
LOCATION: Remote – Latin America (LATAM)
CONTRACT: Hourly Contractor
COMMITMENT: 40 hours per week
TIME OVERLAP: 8 hours of mandatory PST overlap
DURATION: 10 weeks
START DATE: Immediately