Part-Time AI Benchmark Tasks Architect for Terminal Science

OpenTrain AI, Inc.

United States

Remote

USD 55,000 - 110,000

Part time

6 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

OpenTrain AI, Inc. seeks a part-time contractor to create realistic terminal-based scientific tasks for AI training and evaluation. You will build reproducible benchmark environments and assess agents’ ability to reason, use tools, and debug calculations.

The role requires strong scientific programming expertise and independent validation of workflows, running in Linux/terminal environments with fixed dependencies.

Qualifications

  • PhD or equivalent advanced technical experience in physical sciences.
  • Ability to independently create and validate computational scientific workflows.
  • Strong programming skills in Python, C/C++, Julia, Bash, or similar.

Responsibilities

  • Design multi-step terminal tasks based on real physical-science workflows.
  • Build self-contained computational environments with fixed dependencies and scientific software.
  • Create datasets, molecular structures, simulation settings, and model configurations.
  • Write expert solutions using Python, Bash, C/C++, Julia, or domain-specific tools.
  • Develop automated tests for numerical accuracy, physical consistency, convergence, and output structure.
  • Create tasks involving simulations, numerical modeling, data fitting, optimization, spectroscopy, molecular analysis, and visualization.
  • Define numerical tolerances, units, boundary conditions, and expected scientific behavior.
  • Validate reproducibility and debug dependency, precision, solver stability, performance, and file-format problems.

Skills

Python
C/C++
Julia
Bash
Command-line

Education

PhD or equivalent advanced technical experience
Postdoctoral experience

Tools

Docker
Conda
Git
CI
HPC environments

Job description

OpenTrain AI, Inc. seeks a part-time contractor to create realistic terminal-based scientific tasks for AI training and evaluation. You will build reproducible benchmark environments and assess agents’ ability to reason, use tools, and debug calculations.

The role requires strong scientific programming expertise and independent validation of workflows, running in Linux/terminal environments with fixed dependencies.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Physical Sciences AI Benchmark Task Expert
Physical Sciences AI Benchmark Task Expert

OpenTrain AI, Inc. • United States

Remote
USD 55,000 - 110,000
Earth Sciences AI Benchmarking Engineer
Earth Sciences AI Benchmarking Engineer

Gramian Consulting Group • United States

Remote
USD 90,000 - 150,000
Remote Technical Writer for AI Benchmark Tasks
Remote Technical Writer for AI Benchmark Tasks

YO AI Labs • Town of Texas (WI)

Remote
USD 60,000 - 90,000
Remote Technical Writer - AI Benchmark Projects
Remote Technical Writer - AI Benchmark Projects

YO AI Labs • Town of Schroeppel (NY)

Remote
USD 55,000 - 117,000
AI Benchmarking: Physical Sciences Task Designer
AI Benchmarking: Physical Sciences Task Designer

Gramian Consulting Group • United States

Remote
USD 120,000 - 190,000
Remote AI Benchmark Technical Writer (Contractor)
Remote AI Benchmark Technical Writer (Contractor)

YO AI Labs • California (MO)

Remote
USD 55,000 - 110,000
Remote Technical Writer for AI Benchmark & Documentation
Remote Technical Writer for AI Benchmark & Documentation

YO AI Labs • San Francisco (CA)

Remote
USD 55,000 - 117,000
Remote AI Benchmark Technical Writer
Remote AI Benchmark Technical Writer

YO AI Labs • Maryland

Remote
USD 55,000 - 96,000
Remote AI Benchmark Technical Writer
Remote AI Benchmark Technical Writer

YO AI Labs • Seattle (WA)

Remote
USD 83,000 - 131,000
AI Benchmark Technical Writer - Remote Contract
AI Benchmark Technical Writer - Remote Contract

YO AI Labs • Washington

Remote
USD 55,000 - 83,000