AI Trainer - PHYSICS

Planet Pharma

San Francisco (CA)

A distancia

USD 120.000 - 180.000

Jornada completa

hace 18 horas
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Convierte este puesto en una entrevista — un currículum y una carta de presentación creados pensando en lo que quiere el empleador.

Supera los filtros ATS

Descripción de la vacante

Planet Pharma seeks experienced scientists and engineers to evaluate frontier AI models on real technical work, including data analysis, design verification, experiments, and simulations. You will design challenging tasks drawn from your practice, with assumptions and checks, then run them through frontier AI agents and assess results to professional standards.

The role emphasizes self-directed, long-form work in a fully remote setting, with explicit focus on writing clear reasoning and robust

Formación

  • Bachelor’s degree or higher; ideally 2+ years outside undergrad.
  • 2+ years of applied experience in core sciences or engineering or data science.
  • Comfortable writing and verifying work in code (Python or similar) and using the command line.
  • Hands-on with real data and your field’s tools (MATLAB, Python, R; SPICE; FEA/CFD; GIS; spectroscopy software).
  • Strong statistical and experimental reasoning; understanding of experimental design and uncertainty.

Responsabilidades

  • Design realistic technical tasks drawn from day-to-day work, with clear brief and supporting files.
  • Run tasks through frontier AI models and evaluate outputs against professional standards.
  • Compare model outputs on identical prompts and document performance gaps.
  • Write detailed grading rubrics and explain pass/fail criteria in writing.
  • Flag concrete failures with evidence and ensure proper interpretation of the ask.
  • Collaborate across disciplines to refine tasks and reviewer processes.

Conocimientos

Python
Data analysis
Experiment design
Statistical reasoning
Cloud tools

Educación

Bachelor's degree or higher

Herramientas

Python
MATLAB
R
SPICE
FEA/CFD
GIS
Chromatography software
Spectroscopy software
PCB tools

Descripción del empleo

About The Role

is looking for experienced scientists and engineers to evaluate how frontier AI models handle real technical work: analyzing test, measurement or process data, sizing and verifying a design, designing an experiment and reading it out, interpreting a simulation, writing the technical report. You bring the judgment you have built catching the unit error in a test file, the multiple comparisons trap in a process improvement study, and the assumption that does not hold at the boundary. We bring the model output that judgment is needed to grade.

About The Role

is looking for experienced scientists and engineers to evaluate how frontier AI models handle real technical work: analyzing test, measurement or process data, sizing and verifying a design, designing an experiment and reading it out, interpreting a simulation, writing the technical report. You bring the judgment you have built catching the unit error in a test file, the multiple comparisons trap in a process improvement study, and the assumption that does not hold at the boundary. We bring the model output that judgment is needed to grade.

In this role, you will design challenging, realistic tasks drawn from your own practice, such as a calculation package with stated assumptions and checks, a test plan and data analysis, a failure or deviation investigation, a design trade study, an experimental protocol with acceptance criteria, a circuit or system design review, or a technical report, run them through frontier AI agents, and evaluate what comes back against a professional standard.

Across our STEM and Data Science programs, tasks are grounded in real day-to-day workflows and checked against frontier models so only genuinely hard tasks make it through. Some projects are authored and verified in code, so coding / scientific computing is a strong plus and is required on those projects.

You will work with realistic professional files, the kind a practitioner in your field actually handles, which you assemble yourself. Some tasks are compact, built around a handful of files; others are larger scenarios that take several days to build. In every case the goal is the same: a task a competent professional in your field would complete correctly and a frontier model currently gets wrong.

This is not a traditional science or engineering role. You will be helping build better AI by putting your knowledge to work in a structured, flexible, fully remote environment. The work is long-form and self-directed, and clear written reasoning matters as much as technical depth.

Responsibilities
  • Design challenging, realistic technical tasks drawn from your own day-to-day work: the scenario, a prompt phrased the way you would brief a trusted colleague, and the supporting files an engineer or scientist would need (test data, drawings or schematics, specifications, simulation outputs, lab records, reports), which you author yourself.
  • Run those tasks through frontier AI models and evaluate the deliverable they produce (the calculation, analysis, design review or report) against the standard you would hold a colleague to.
  • Compare two model outputs on identical prompts and files, decide which performed better, and document where each fell short.
  • Write detailed grading rubrics that specify what a correct deliverable must contain (the right assumptions stated, the right method, the right magnitudes and units, the right failure modes considered), and explain in writing why a response passes or fails each one.
  • Flag concrete failures with evidence: unit and scaling errors, misread data, unsupported conclusions, fabricated or ignored source files, missed safety or boundary conditions, and off-brief interpretation of the ask.
  • Contribute across your discipline and adjacent ones, and review and refine tasks built by other experts.
Qualifications
  • 2+ years of applied experience preferred in one of the core sciences (mathematics, physics, chemistry, or biology) or in an engineering discipline (electrical and electronics, civil and structural, materials, chemical and process, environmental and earth sciences, or similar) or in data science.
  • In progress Bachelor’s degree or higher. We prefer 2+ years of experience outside undergraduate study.
  • Coding / scientific computing is a strong plus and is required on some projects: comfortable writing and verifying work in code (Python or similar) and working at the command line.
  • Hands-on with real data and the tools of your field: instrument and test data, measurement files, simulation outputs, schematics or drawings, and the analysis tools that go with them (MATLAB, Python or R; SPICE and PCB tools; FEA or CFD; GIS; chromatography or spectroscopy software).
  • Comfort with statistical and experimental reasoning: experiment design, measurement error and uncertainty, and the common statistical traps. The assessment leans on this.
  • Working understanding of adjacent sub-disciplines, enough to assess work outside your own specialty and point out what was done correctly or incorrectly.
  • Hands‑on practitioner: you currently do (or recently did) the work yourself at an individual-contributor level, not solely in a managerial capacity.
  • Full professional or native‑level written and spoken English; you can articulate why a result is wrong, not only that it is.
  • General familiarity with AI and LLM tools: you can tell a well‑reasoned answer from a plausible‑sounding but incorrect one.
  • Baseline tech literacy: comfortable with cloud file tools (e.g., Google Workspace), managing browser profiles, downloading and installing desktop apps (e.g., Claude), and everyday file handling (e.g., converting between Excel and Google Sheets, zipping files for sharing).
  • Advanced degree or professional licensure (PE, chartered status) is a plus but not required. Practical applied work outweighs credentials.
Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

AI Trainer - CHEMISTRY
AI Trainer - CHEMISTRY

Planet Pharma • San Francisco (CA)

A distancia
USD 120.000 - 180.000
Fully remote
Flexible schedule
AI Trainer – PHYSICS
AI Trainer – PHYSICS

Planet Pharma • San Francisco (CA)

Presencial
USD 60.000 - 110.000
AI Trainer – BIOLOGY
AI Trainer – BIOLOGY

Planet Pharma • San Francisco (CA)

Presencial
USD 120.000 - 180.000
AI Trainer – MATH
AI Trainer – MATH

Planet Pharma • San Francisco (CA)

Presencial
USD 120.000 - 180.000
AI Trainer – CHEMISTRY
AI Trainer – CHEMISTRY

Planet Pharma • San Francisco (CA)

Presencial
USD 120.000 - 180.000
AI Trainer – Mechanical Engineering
AI Trainer – Mechanical Engineering

Planet Pharma • San Francisco (CA)

A distancia
USD 90.000 - 150.000
AI Trainer – Accounting
AI Trainer – Accounting

Planet Pharma • San Francisco (CA)

Presencial
USD 120.000 - 180.000
AI Trainer – Electrical Engineering
AI Trainer – Electrical Engineering

Planet Pharma • San Francisco (CA)

A distancia
USD 120.000 - 180.000
AI Trainer - Healthcare Operations
AI Trainer - Healthcare Operations

Planet Pharma • San Francisco (CA)

A distancia
USD 120.000 - 180.000
AI Trainer – Medical (Nurses, Advanced Practice, Physicians, Pharmacists)
AI Trainer – Medical (Nurses, Advanced Practice, Physicians, Pharmacists)

Planet Pharma • San Francisco (CA)

Presencial
USD 120.000 - 180.000