AI Trainer – BIOLOGY

Planet Pharma

San Francisco (CA)

Remote

USD 120,000 - 180,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Planet Pharma (remote) seeks experienced scientists and engineers to design and evaluate frontier AI models using real test, measurement, and process data. You will craft realistic tasks from your practice, run them through AI agents, and assess outputs against professional standards.

Strong coding and scientific computing are required for some projects, with long-form, self-directed work in a fully remote setting.

Qualifications

  • 2+ years of applied experience in core sciences or engineering or data science.
  • In progress Bachelor’s degree or higher; 2+ years outside undergrad preferred.
  • Comfort writing and verifying work in code (Python or similar) and using command line.
  • Hands-on with real data and analysis tools (MATLAB, Python, R; SPICE, PCB tools; FEA/CFD; GIS).
  • Strong understanding of statistical experimentation and inference.

Responsibilities

  • Design challenging, realistic technical tasks drawn from your day-to-day work.
  • Run tasks through frontier AI models and evaluate deliverables against professional standards.
  • Compare two model outputs and document which performed better and why.
  • Write detailed grading rubrics and explain pass/fail criteria.
  • Flag concrete failures with evidence: unit errors, misread data, or boundary issues.
  • Review and refine tasks built by other experts across disciplines.

Skills

Applied experience
Python
MATLAB
R
Data science
Experiment design
Statistics
English writing

Education

Bachelor’s degree or higher

Tools

MATLAB
Python
SPICE
PCB tools
FEA
CFD
GIS
Chromatography software

Job description

About the role

is looking for experienced scientists and engineers to evaluate how frontier AI models handle real technical work: analyzing test, measurement or process data, sizing and verifying a design, designing an experiment and reading it out, interpreting a simulation, writing the technical report. You bring the judgment you have built catching the unit error in a test file, the multiple comparisons trap in a process improvement study, and the assumption that does not hold at the boundary. We bring the model output that judgment is needed to grade.

In this role, you will design challenging, realistic tasks drawn from your own practice, such as a calculation package with stated assumptions and checks, a test plan and data analysis, a failure or deviation investigation, a design trade study, an experimental protocol with acceptance criteria, a circuit or system design review, or a technical report, run them through frontier AI agents, and evaluate what comes back against a professional standard.

Across our STEM and Data Science programs, tasks are grounded in real day-to-day workflows and checked against frontier models so only genuinely hard tasks make it through. Some projects are authored and verified in code, so coding / scientific computing is a strong plus and is required on those projects.

You will work with realistic professional files, the kind a practitioner in your field actually handles, which you assemble yourself. Some tasks are compact, built around a handful of files; others are larger scenarios that take several days to build. In every case the goal is the same: a task a competent professional in your field would complete correctly and a frontier model currently gets wrong.

This is not a traditional science or engineering role. You will be helping build better AI by putting your knowledge to work in a structured, flexible, fully remote environment. The work is long-form and self-directed, and clear written reasoning matters as much as technical depth.

Responsibilities
  • ? Design challenging, realistic technical tasks drawn from your own day-to-day work: the scenario, a prompt phrased the way you would brief a trusted colleague, and the supporting files an engineer or scientist would need (test data, drawings or schematics, specifications, simulation outputs, lab records, reports), which you author yourself.
  • ? Run those tasks through frontier AI models and evaluate the deliverable they produce (the calculation, analysis, design review or report) against the standard you would hold a colleague to.
  • ? Compare two model outputs on identical prompts and files, decide which performed better, and document where each fell short.
  • ? Write detailed grading rubrics that specify what a correct deliverable must contain (the right assumptions stated, the right method, the right magnitudes and units, the right failure modes considered), and explain in writing why a response passes or fails each one.
  • ? Flag concrete failures with evidence: unit and scaling errors, misread data, unsupported conclusions, fabricated or ignored source files, missed safety or boundary conditions, and off-brief interpretation of the ask.
  • ? Contribute across your discipline and adjacent ones, and review and refine tasks built by other experts.
Qualifications
  • ? 2+ years of applied experience preferred in one of the core sciences (mathematics, physics, chemistry, or biology) or in an engineering discipline (electrical and electronics, civil and structural, materials, chemical and process, environmental and earth sciences, or similar) or in data science.
  • ? In progress Bachelor’s degree or higher. We prefer 2+ years of experience outside undergraduate study.
  • ? Coding / scientific computing is a strong plus and is required on some projects: comfortable writing and verifying work in code (Python or similar) and working at the command line.
  • ? Hands-on with real data and the tools of your field: instrument and test data, measurement files, simulation outputs, schematics or drawings, and the analysis tools that go with them (MATLAB, Python or R; SPICE and PCB tools; FEA or CFD; GIS; chromatography or spectroscopy software).
  • ? Comfort with statistical and experimental reasoning: experiment design, measurement error and uncertainty, and the common statistical traps. The assessment leans on this.
  • ? Working understanding of adjacent sub-disciplines, enough to assess work outside your own specialty and point out what was done correctly or incorrectly.
  • ? Hands-on practitioner: you currently do (or recently did) the work yourself at an individual-contributor level, not solely in a managerial capacity.
  • ? Full professional or native-level written and spoken English; you can articulate why a result is wrong, not only that it is.
  • ? General familiarity with AI and LLM tools: you can tell a well-reasoned answer from a plausible-sounding but incorrect one.
  • ? Baseline tech literacy: comfortable with cloud file tools (e.g., Google Workspace), managing browser profiles, downloading and installing desktop apps (e.g., Claude), and everyday file handling (e.g., converting between Excel and Google Sheets, zipping files for sharing).
  • ? Advanced degree or professional licensure (PE, chartered status) is a plus but not required. Practical applied work outweighs credentials.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Trainer – MATH
AI Trainer – MATH

Planet Pharma • San Francisco (CA)

Remote
USD 120,000 - 180,000
AI Trainer – PHYSICS
AI Trainer – PHYSICS

Planet Pharma • San Francisco (CA)

Remote
USD 60,000 - 110,000
AI Trainer – CHEMISTRY
AI Trainer – CHEMISTRY

Planet Pharma • San Francisco (CA)

Remote
USD 120,000 - 180,000
AI Trainer – Mechanical Engineering
AI Trainer – Mechanical Engineering

Planet Pharma • San Francisco (CA)

Remote
USD 90,000 - 150,000
AI Trainer – Electrical Engineering
AI Trainer – Electrical Engineering

Planet Pharma • San Francisco (CA)

Remote
USD 120,000 - 180,000
AI Trainer – Medical (Nurses, Advanced Practice, Physicians, Pharmacists)
AI Trainer – Medical (Nurses, Advanced Practice, Physicians, Pharmacists)

Planet Pharma • San Francisco (CA)

Remote
USD 120,000 - 180,000
AI Trainer – Accounting
AI Trainer – Accounting

Planet Pharma • San Francisco (CA)

Remote
USD 120,000 - 180,000
AI Trainer – Legal
AI Trainer – Legal

Planet Pharma • San Francisco (CA)

Remote
USD 140,000 - 190,000
Member of Technical Staff (Applied AI Research)
Member of Technical Staff (Applied AI Research)

Artificial Analysis, Inc. • San Francisco (CA)

On-site
USD 170,000 - 210,000
Member of Technical Staff (Language Model Evaluations)
Member of Technical Staff (Language Model Evaluations)

Artificial Analysis • San Francisco (CA)

On-site
USD 180,000 - 260,000
Equity