AI TRAINER – CODING

Planet Pharma

San Francisco (CA)

Sur place

USD 90 000 - 130 000

Plein temps

Il y a 16 heures
Soyez parmi les premiers à postuler
Générateur de candidature

Une candidature complète en une minute — un CV et une lettre de motivation personnalisés, prêts à envoyer.

Passez les filtres ATS

Résumé du poste

Planet Pharma is seeking an AI Trainer – Coding in San Francisco, CA, to design and evaluate realistic software engineering tasks for AI agents. You will craft scenarios from your practice, including bug fixes, specs, and PR reviews, then run them through frontier AI agents and judge outputs against a professional standard.

The role is on-site in San Francisco and requires strong engineering judgment. The position emphasizes explicit written reasoning and working with real production-like files,

Qualifications

  • 2+ years of professional software engineering experience preferred.
  • In progress Bachelor’s degree or higher.
  • Expertise in at least one major stack (JS/TS, Python, Java, Go, Rust, C++, C#, Ruby, PHP, Swift, Kotlin).
  • Proficient in terminal usage, Git, tests, and a debugger; Docker and CI experience is a plus.
  • Strong ability to read unfamiliar codebases and articulate reasoning in writing.
  • English: full professional proficiency in speaking and writing.

Responsabilités

  • Design realistic software engineering tasks drawn from day-to-day work.
  • Run tasks through AI coding agents and evaluate accuracy against professional standards.
  • Compare agent outputs on identical specs and document gaps.
  • Create tests, rubrics, and success criteria for correct deliverables.
  • Flag failures with concrete evidence and align on fix strategies.
  • Contribute across the stack and review tasks created by others.

Connaissances

JavaScript/TypeScript
Python
Java
Go
Rust
C++
C#
Ruby
PHP
Swift
Kotlin

Formation

Bachelor’s degree or higher

Outils

Docker
CI/CD
Git

Description du poste

Job Description

AI Trainer – Coding

Location: San Francisco, California

Category: Technology

Salary: Apply for details

Country: United States

Employment: Direct Hire/Perm

Worksite: On-Site

About the role

is looking for experienced software engineers to evaluate how AI coding agents handle real engineering work inside realistic codebases: implementing a change to spec, fixing a bug with a regression test, refactoring without breaking behavior, reviewing a pull request, debugging from a trace. Coding agents are strong in many ways and frustratingly bad in others: they miss edge cases, make brittle changes, and need close supervision from an experienced engineer. You bring the judgment you have built over a career shipping production systems. We bring the agent output that judgment is needed to grade.

In this role, you will design challenging, realistic tasks drawn from your own practice, such as a bug fix with a regression test, a feature implemented to spec with tests, a refactor with behavior preserved, a pull request review, a debugging write-up from logs and traces, or an API contract with examples, run them through frontier AI agents, and evaluate what comes back against a professional standard.

You will work with realistic professional files, the kind a practitioner in your field actually handles, which you assemble yourself. Some tasks are compact, built around a handful of files; others are larger scenarios that take several days to build. In every case the goal is the same: a task a competent professional in your field would complete correctly and a frontier model currently gets wrong.

This is not a traditional software engineering role. You will be helping build better AI by putting your knowledge to work in a structured, flexible, fully remote environment. The work is long-form and self-directed, and clear written reasoning matters as much as technical depth.

Responsibilities

Design challenging, realistic software engineering tasks drawn from your own day-to-day work: the scenario, a spec phrased the way you would brief a trusted teammate, and the codebase or supporting files an engineer would need (repositories you construct or adapt, tests, logs, API contracts, design notes), which you author yourself.

Run those tasks through AI coding agents and evaluate the deliverable they produce (the diff, the tests, the design, the review) against the standard you would hold a teammate to.

Compare two agent outputs on identical specs and codebases, decide which performed better, and document where each fell short.

Write tests, rubrics, and success criteria that specify what a correct deliverable must contain (the right behavior preserved, the right edge cases handled, the right tests added, the right conventions followed), and explain in writing why a submission passes or fails each one.

Flag concrete failures with evidence: brittle changes, missed edge cases, tests that pass for the wrong reason, thrashing in the trace, fabricated or ignored files, and off-spec interpretation of the ask.

Contribute across your stack and adjacent ones, and review and refine tasks built by other engineers.

Qualifications

2+ years of professional software engineering experience preferred, shipping and maintaining production code as part of a team.

In progress Bachelor’s degree or higher.

Expertise in at least one common stack: TypeScript or JavaScript, Python, Java, Go, Rust, C++, C#, Ruby, PHP, Swift, or Kotlin. Specialists are welcome: a strong frontend-only, backend-only, or mobile engineer is a good referral.

Comfortable in a terminal with git, tests, and a debugger; able to read an unfamiliar codebase and orient quickly. Docker and CI experience is a plus.

Strong judgment about what good code and a real model failure look like, and the ability to explain both in writing.

No prior AI or machine learning experience is required. Engineering judgment and attention to detail matter most.

Hands-on practitioner: you currently do (or recently did) the work yourself at an individual-contributor level, not solely in a managerial capacity.

Full professional or native-level written and spoken English; you can articulate why a result is wrong, not only that it is.

General familiarity with AI and LLM tools: you have used models like Claude or ChatGPT in professional work and can tell a well-reasoned answer from a plausible-sounding but incorrect one.

Baseline tech literacy: comfortable with cloud file tools (e.g., Google Workspace), managing browser profiles, downloading and installing desktop apps (e.g., Claude), and everyday file handling (e.g., converting between Excel and Google Sheets, zipping files for sharing).

A computer science degree is a plus but not required; shipped professional work outweighs credentials.

Equal Opportunity Employer

We are proud to be an equal opportunity employer. We welcome and encourage applications from all qualified candidates regardless of race, sex, gender identity or expression, disability, age, religion or belief, sexual orientation, or any other characteristic protected by applicable laws and regulations. It is our policy not to discriminate against any applicant or employee, and we are committed to fostering a diverse, inclusive, and respectful work environment across all locations in which we operate. We believe that diversity, equity, and inclusion are fundamental to our mission and enhance our ability to serve clients globally. If you have a disability or require any reasonable accommodations during the application or interview process, please inform your recruiter or contact us(opens in new tab) directly so that we can explore the appropriate arrangements.

Obtenez votre examen gratuit et confidentiel de votre CV.

ou faites glisser et déposez votre fichier ici.

Similar jobs

Postes similaires à comparer

AI TRAINER – HEALTHCARE OPERATIONS
AI TRAINER – HEALTHCARE OPERATIONS

Planet Pharma • San Francisco (CA)

Sur place
USD 120 000 - 170 000
AI Engineer
AI Engineer

Willis Towers Watson • Nashville (TN)

Sur place
USD 210 000 - 250 000
Health benefits
401(k) plan
Paid time off
AI TRAINER – FINANCE
AI TRAINER – FINANCE

Planet Pharma • San Francisco (CA)

Sur place
USD 140 000 - 190 000
Software Engineer, Teacher Experience
Software Engineer, Teacher Experience

CodeAI • Washington

Sur place
USD 127 000 - 162 000
Technology subsidy
Remote work
Paid time off 5 weeks
+3
AI Software Engineer
AI Software Engineer

Ledgent Technology • Livermore (CA)

Sur place
USD 103 320 - 144 648
Software Engineer, Teacher Experience
Software Engineer, Teacher Experience

Code.org • Seattle (WA)

Sur place
USD 127 000 - 162 000
Technology subsidy for BYOD
Remote-friendly work environment
Paid time off: 5 weeks
+3
AI Trainer – BIOLOGY
AI Trainer – BIOLOGY

Planet Pharma • San Francisco (CA)

Sur place
USD 120 000 - 180 000
AI-Native Software Engineer (New Grad)
AI-Native Software Engineer (New Grad)

TrulyHired • Mountain View (CA)

Hybride
USD 59 000 - 81 000
AI Trainer - PHYSICS
AI Trainer - PHYSICS

Planet Pharma • San Francisco (CA)

À distance
USD 120 000 - 180 000
AI Trainer - CHEMISTRY
AI Trainer - CHEMISTRY

Planet Pharma • San Francisco (CA)

À distance
USD 120 000 - 180 000
Fully remote
Flexible schedule