Remote Software Engineer - Python/Typescript

Turing

United States

Remote

USD 83,000 - 152,000

Part time

6 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Turing is seeking experienced software engineers to evaluate and improve AI coding models. You’ll review AI-generated code and agent behavior across real-world repositories, judge correctness, and identify failure modes to help refine model performance.

In this contractor role, you’ll collaborate with researchers and engineers, design rubrics, and produce clear evaluation data and actionable recommendations to advance coding models.

Qualifications

  • 5+ years hands-on software engineering experience
  • Proficiency in Python, TypeScript/JavaScript, Go or another major production language
  • Experience reviewing AI-generated code and coding agents

Responsibilities

  • Evaluate AI-generated code across real-world repositories
  • Review agent behavior, tool usage, and code changes for correctness
  • Identify technical errors, weak approaches, and recurring model failure modes
  • Compare model outputs and explain why one solution is better than another
  • Create and refine rubrics and evaluation criteria for coding tasks
  • Produce high-quality evaluation data and recommendations for the team
  • Build and maintain pipelines and infrastructure for data generation, collection, and evaluation
  • Synthesize findings into clear write-ups and updates for the team
  • Collaborate with researchers and engineers to translate qualitative judgment into scalable processes
  • Share clear, actionable findings with AI researchers and engineers

Skills

5+ years hands-on software engineering
Python
TypeScript/JavaScript
Go
Code review
Written communication
AI coding tools
LLM evaluation

Tools

Git
CI/CD
LLM tooling

Job description

About Turing

Turing is one of the world’s leading AGI infrastructure companies, working with frontier AI labs to accelerate model development through high-quality training data, evaluations, and engineering talent.

About the Role

We’re looking for experienced, hands-on software engineers to help evaluate and improve AI coding models.

Rather than primarily building production applications, you’ll work with coding agents across real-world repositories and assess the quality of their work. You’ll review generated code and agent behavior, determine whether solutions are technically correct, identify failure modes, and create the evaluation signals and feedback used to improve model performance.

Think of the coding agent as another engineer whose work you’re reviewing: Can it understand the task? Did it choose the right approach? Is the resulting code correct, robust, and maintainable? Can you explain precisely where it succeeded or failed?

What You’ll Do
  • Evaluate AI-generated code and solutions across real-world software repositories
  • Review agent behavior, tool usage, and code changes for correctness and quality
  • Identify technical errors, weak approaches, and recurring model failure modes
  • Compare model outputs and explain why one solution is better than another
  • Create and refine rubrics and evaluation criteria for coding tasks
  • Produce high-quality evaluation and preference data used to improve coding models
  • Build and maintain pipelines and infrastructure supporting data generation, collection, and evaluation workflows
  • Synthesize findings from data work into clear write-ups, updates, and recommendations for the team
  • Collaborate closely with researchers and engineers to translate qualitative judgment into scalable processes
  • Share clear, actionable findings with AI researchers and engineers
What We’re Looking For
  • 5+ years of hands-on software engineering experience
  • Strong proficiency in Python, TypeScript/JavaScript, Go, or another major production language
  • Experience working in substantial real-world codebases
  • Strong code-review skills and technical judgment
  • Ability to clearly explain why an implementation is correct, incorrect, or could be improved
  • Strong written communication
  • Experience using modern LLMs or AI coding tools

Experience with LLM evaluation, coding agents, RLHF, preference data, rubric design, or post-training is a plus, but not required.

Engagement Details
  • Compensation: Market rate; please provide a specific hourly rate expectation
  • Availability: 40 hours/week preferred, with at least 6 hours of Pacific Time overlap
  • Type: Independent contractor
  • Duration: Approximately 3 months
  • Start: As soon as possible
Evaluation Process
  • AI interview (~25 minutes)
  • Practical code/AI evaluation exercise (~30 minutes)
  • Hiring manager interview (~20 minutes)

The practical exercise focuses on your ability to review and evaluate AI-generated code, not competitive programming or algorithm puzzles.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Software Engineer - Python/Typescript
Software Engineer - Python/Typescript

Turing • Seattle (WA)

On-site
USD 110,000 - 193,000
Software Engineer
Software Engineer

turing • California (MO)

Hybrid
USD 55,000 - 96,000
Remote Senior Python Engineer – LLM Evaluation (US-based)
Remote Senior Python Engineer – LLM Evaluation (US-based)

Turing • Chicago (IL)

On-site
Software Developer – AI Research & Evaluation
Software Developer – AI Research & Evaluation

turing • California (MO)

Hybrid
USD 14,000 - 55,000
Java Developer
Java Developer

Turing • New York (NY)

On-site
USD 83,000 - 165,000
Remote Software Developer
Remote Software Developer

Turing • Minneapolis (MN)

Remote
Python Developer
Python Developer

turing • California (MO)

Hybrid
USD 83,000 - 124,000
Java Developer
Java Developer

Turing • San Francisco (CA)

On-site
USD 69,000 - 96,000
Flexible hours
Contractor engagement
Senior Software Developer
Senior Software Developer

turing • California (MO)

Hybrid
USD 83,000 - 165,000
Remote Software Developer (US)
Remote Software Developer (US)

Turing • Austin (TX)

Remote