Remote AI Software Engineer - Code & Model Evaluator

Turing

New York (NY)

Remote

USD 55,104 - 110,208

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Turing is seeking a Software Engineering Evaluator to create cutting‑edge datasets for training, benchmarking, and advancing large language models, collaborating closely with researchers.

This includes curating code examples, providing precise solutions, and making corrections in Python, JavaScript (including ReactJS), C/C++, Java, Rust, and Go; evaluating and refining AI-generated code for efficiency, scalability, and reliability; and working with cross‑functional teams to enhance

Qualifications

  • 3+ years of software engineering experience.
  • Strong expertise in building full‑stack applications and deploying scalable, production‑grade software using modern languages and tools.
  • Deep understanding of software architecture, design, development, debugging, and code quality/review assessment.
  • Excellent oral and written communication skills for clear, structured evaluation rationales.

Responsibilities

  • Curate code examples across Python, JavaScript (including ReactJS), C/C++, Java, Rust, and Go for AI model training.
  • Evaluate and refine AI-generated code to ensure it is efficient, scalable, and reliable.
  • Collaborate with cross‑functional teams to enhance AI‑driven coding solutions against industry performance benchmarks.
  • Build agents that verify code quality and identify error patterns.
  • Hypothesize on steps in the software engineering cycle and evaluate model capabilities on them.
  • Design verification mechanisms that automatically verify solutions to software tasks.

Skills

Software engineering
Full-stack
Architecture
Communication

Job description

Turing is seeking a Software Engineering Evaluator to create cutting‑edge datasets for training, benchmarking, and advancing large language models, collaborating closely with researchers.

This includes curating code examples, providing precise solutions, and making corrections in Python, JavaScript (including ReactJS), C/C++, Java, Rust, and Go; evaluating and refining AI-generated code for efficiency, scalability, and reliability; and working with cross‑functional teams to enhance

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

C++ Software Engineer - AI Model Evaluation & Datasets
C++ Software Engineer - AI Model Evaluation & Datasets

Turing • Austin (TX)

On-site
USD 55,000 - 110,000
Python AI Software Engineer - Data & Code Evaluation
Python AI Software Engineer - Data & Code Evaluation

Turing • Austin (TX)

On-site
USD 83,000 - 124,000
AI Dataset & Code Evaluation Engineer
AI Dataset & Code Evaluation Engineer

Turing • New York (NY)

On-site
USD 69,000 - 124,000
AI Code Evaluator & Data Curation Engineer
AI Code Evaluator & Data Curation Engineer

Turing • Austin (TX)

On-site
USD 69,000 - 117,000
Senior Software Engineer - AI Coding Evaluator
Senior Software Engineer - AI Coding Evaluator

Turing • Austin (TX)

On-site
USD 14,000 - 55,000
AI Systems Engineer: Code, Evaluation & Research
AI Systems Engineer: Code, Evaluation & Research

Turing • Chicago (IL)

On-site
USD 14,000 - 55,000
Java Software Engineer for AI Data & Evaluation
Java Software Engineer for AI Data & Evaluation

Turing • San Francisco (CA)

On-site
USD 83,000 - 165,000
Java Software Engineer for AI Code Evaluation (Contract)
Java Software Engineer for AI Code Evaluation (Contract)

Turing • Chicago (IL)

On-site
USD 28,000 - 55,000
AI-Focused Software Engineer & Code Evaluator
AI-Focused Software Engineer & Code Evaluator

Turing • San Francisco (CA)

On-site
USD 83,000 - 165,000
Go Software Engineer for AI Code Evaluation
Go Software Engineer for AI Code Evaluation

Turing • New York (NY)

On-site
USD 55,000 - 96,000