Remote AI Code Evaluator for Large Language Models

turing

Canada

Remote

CAD 98,000 - 137,000

Full time

2 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Turing in San Francisco seeks a Software Engineering evaluator to create cutting-edge datasets for training, benchmarking, and advancing large language models. You will curate code examples, provide precise solutions, and correct in Python, JavaScript (ReactJS), C/C++, Java, Rust, and Go while collaborating with researchers.

You will evaluate AI-generated code for efficiency and reliability, build verification mechanisms, and work with cross-functional teams to improve enterprise AI-driven

Qualifications

  • 3+ years of software engineering experience.
  • Strong expertise in building full-stack applications and deploying scalable, production-grade software using modern languages and tools.
  • Deep understanding of software architecture, design, development, debugging, and code quality/review assessment.
  • Excellent oral and written communication skills for clear, structured evaluation rationales.

Responsibilities

  • Curate code examples and provide precise solutions in Python, JavaScript (including ReactJS), C/C++, Java, Rust, and Go.
  • Evaluate and refine AI-generated code for efficiency, scalability, and reliability.
  • Collaborate with cross-functional teams to enhance AI-driven coding solutions against industry benchmarks.
  • Build verification mechanisms to automatically verify software engineering tasks.

Skills

3+ years experience
Full-stack development
Software architecture
Code quality/review
Excellent communication

Tools

ReactJS
Python

Job description

Turing in San Francisco seeks a Software Engineering evaluator to create cutting-edge datasets for training, benchmarking, and advancing large language models. You will curate code examples, provide precise solutions, and correct in Python, JavaScript (ReactJS), C/C++, Java, Rust, and Go while collaborating with researchers.

You will evaluate AI-generated code for efficiency and reliability, build verification mechanisms, and work with cross-functional teams to improve enterprise AI-driven

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Remote Software Developer
Remote Software Developer

turing • Canada

Remote
CAD 98,000 - 137,000
Remote Rust Developer
Remote Rust Developer

turing • Canada

Remote
CAD 78,000 - 118,000
Remote Research Engineer: Code Gen & Model Evaluation
Remote Research Engineer: Code Gen & Model Evaluation

YO AI Labs • Ottawa

Remote
CAD 55,000 - 117,000
Remote Senior AI Code Evaluation Engineer
Remote Senior AI Code Evaluation Engineer

YO AI Labs • Calgary

Remote
CAD 90,000 - 131,000
Remote Senior Full-Stack Engineer for AI Code Evaluation
Remote Senior Full-Stack Engineer for AI Code Evaluation

YO AI Labs • Ottawa

Remote
CAD 55,000 - 96,000
Senior Software Engineer - AI/ML for Codebase RL (Remote)
Senior Software Engineer - AI/ML for Codebase RL (Remote)

YO AI Labs • Vancouver

Remote
CAD 38,000 - 76,000
Remote Research Engineer: Code Gen & Model Eval
Remote Research Engineer: Code Gen & Model Eval

YO AI Labs • Montreal (administrative region)

Remote
CAD 83,000 - 131,000
Remote Senior Software Engineer - AI Training Environments
Remote Senior Software Engineer - AI Training Environments

YO AI Labs • Calgary

Remote
CAD 83,000 - 152,000
Data Engineer - AI Model Evaluator
Data Engineer - AI Model Evaluator

Mercor • Toronto

On-site
CAD 193,000 - 248,000
Data Engineer - AI Model Evaluator
Data Engineer - AI Model Evaluator

Obsidian • Toronto

On-site
CAD 413,000 - 689,000