C/C++ AI Code Evaluator

Turing

San Francisco (CA)

Remote

USD 83,000 - 138,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Turing, based in San Francisco, seeks an experienced Software Engineering evaluator to create datasets and refine AI-generated code. The role focuses on C/C++, Python, ReactJS, Java, Rust, and Go, with collaboration across researchers to advance enterprise AI coding solutions.

Engagement is flexible (10–40 hrs/week) as a contractor for 1 month with potential extension based on performance. Strong communication and full-stack expertise are essential.

Qualifications

  • Several years of software engineering experience (3+ years).
  • Experience building full-stack applications and deploying production-grade software.
  • Strong understanding of software architecture, design, development, debugging, and code quality/review assessment.

Responsibilities

  • Curate code examples and provide precise solutions primarily in C/C++, Python, JavaScript (ReactJS), Java, Rust, and Go.
  • Evaluate and refine AI-generated code for efficiency, scalability, and reliability.
  • Collaborate with cross-functional teams to enhance AI-driven coding solutions against industry benchmarks.
  • Build agents that can verify code quality and identify error patterns.
  • Hypothesize steps in the software engineering cycle and evaluate model capabilities on them.
  • Design verification mechanisms that automatically verify solutions to software engineering tasks.

Skills

Full-stack development
Production-grade software
Software architecture
Code review
English communication

Tools

Python
ReactJS
Java
Go
Rust

Job description

About Us:
Based in San Francisco, California, Turing is the world's leading research accelerator for frontier AI labs and a trusted partner for global enterprises deploying advanced AI systems. Turing supports customers in two ways: first, by accelerating frontier research with high-quality data, advanced training pipelines, plus top AI researchers who specialize in software engineering, logical reasoning, STEM, multilinguality, multimodality, and agents; and second, by applying that expertise to help enterprises transform AI from proof of concept into proprietary intelligence with systems that perform reliably, deliver measurable impact, and drive lasting results on the P&L.
Project Overview:
As a Software Engineering evaluator, you will create cutting-edge datasets for training, benchmarking, and advancing large language models, collaborating closely with researchers. This includes curating code examples, providing precise solutions, and making corrections primarily in
C/C++, as well as in Python, JavaScript (including ReactJS), Java, Rust, and Go; evaluating and refining AI-generated code for efficiency, scalability, and reliability; and working with cross-functional teams to enhance enterprise-level AI-driven coding solutions.
What Does a Typical Day Look Like?
  • Working on AI model training initiatives by curating code examples, building solutions, and correcting code — primarily in C/C++, along with Python, JavaScript (including ReactJS), Java, Rust, and Go.
  • Evaluate and refine AI-generated code to ensure that it is efficient, scalable, and reliable.
  • Collaborate with cross-functional teams to enhance AI-driven coding solutions against industry performance benchmarks.
  • Build agents that can verify the quality of the code and identify error patterns.
  • Hypothesize on steps in the software engineering cycle (prototyping, architecture design, API design, production implementation, launch, experiments, monitoring, operational maintenance) and evaluate model capabilities on them.
  • Design verification mechanisms that can automatically verify a solution to a software engineering task.
Required Skills:
  • Several years of software engineering experience (3 years or more)
  • Strong expertise in building full-stack applications and deploying scalable, production-grade software using modern languages and tools.
  • Deep understanding of software architecture, design, development, debugging, and code quality/review assessment.
  • Excellent oral and written communication skills for clear, structured evaluation rationales.
Engagement Details:
  • Commitment: flexible engagement, minimum 10 hrs/week, up to 40 hrs/week
  • Type: Contractor (no medical/paid leave)
  • Duration: 1 month (potential extensions based on performance and fit)
  • Location: Candidates must be based out of the US, Canada, or WEU countries (Austria, Belgium, France, Germany, ...)
Evaluation Process:
  • The application process takes 15-30 minutes.
  • Completion of an AI video interview is required.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

C/C++ AI Code Evaluator (Contract)
C/C++ AI Code Evaluator (Contract)

Turing • New York (NY)

On-site
USD 83,000 - 138,000
C/C++ Developer
C/C++ Developer

Turing • New York (NY)

On-site
USD 83,000 - 138,000
Remote C/C++ Developer
Remote C/C++ Developer

Turing • Seattle (WA)

Remote
USD 14,000 - 55,000
Contractor (no medical/paid leave)
Remote C/C++ Developer
Remote C/C++ Developer

Turing • San Francisco (CA)

Remote
USD 83,000 - 138,000
C/C++ Developer
C/C++ Developer

Turing • San Francisco (CA)

On-site
USD 83,000 - 124,000
AI-Focused Java Engineer: Code Curation & Evaluation
AI-Focused Java Engineer: Code Curation & Evaluation

Turing • San Francisco (CA)

On-site
USD 69,000 - 96,000
Flexible hours
Contractor engagement
Remote Software Developer
Remote Software Developer

Turing • San Francisco (CA)

Remote
USD 55,000 - 96,000
Java Developer
Java Developer

Turing • San Francisco (CA)

On-site
USD 69,000 - 96,000
Flexible hours
Contractor engagement
Java Developer
Java Developer

Turing • New York (NY)

On-site
USD 83,000 - 165,000
Remote Java Developer
Remote Java Developer

turing • United States

Remote
USD 83,000 - 124,000