A complete application in a minute — tailored resume and cover letter, ready to send.
AI Ethics Network is seeking evaluators to perform high-stakes testing of robotics, teen-facing products, and post-incident scenarios. You will red-team models and agents, run prompt-injection tests, assess sensor and navigation safety, and document findings for executives.
The role requires a Master’s degree (PhD preferred in your track), clear written English, and the ability to defend methods in front of counsel. U.S.
$100–$160 per hour on contract. Rate follows domain depth. Robotics, teen-facing, or post-incident work sits at the top of the band.
Typical assessment: 20–60 hours. If conversion is offered after paid work: $145,000–$185,000 base equivalent plus project bonuses. No equity at the consulting stage.
Eval Labs does not let the training team grade itself. When a client is weeks from launch — or managing an active incident — the evaluator must hold an advanced degree in a field related to the system under test and be able to defend methods under questioning from counsel, leadership, or insurers. We hire for that standard.
Skilled in at least one track. Two preferred. Three is priced accordingly.
Degree: MS or PhD in CS, machine learning, cybersecurity, or equivalent.
LLM/agent evaluation, log analysis, eval harnesses, and failure modes explained without theater.
Degree: MS or PhD in robotics, mechanical or electrical engineering, controls, or computer vision.
Sensors, actuators, stop conditions, and real-world safety-interlock failure — not only simulators.
Degree: MS or PhD in HCI, cognitive science, special education, clinical/counseling with HCI crossover, or documented ND accessibility practice.
Paid reviewer work, instruction-load evaluation, and a method you can defend.
Required. Master’s required; PhD preferred in a field related to your primary track. Evidence you personally ran the work — not summarized papers. Clear written English. Authorized to work as a U.S. contractor (no sponsorship). Able to refuse scope outside your qualification. No active conflict with an Eval Labs prospect or a competing audit of the same client system.
Strongly preferred. Expert-witness writing, incident reports, or board/insurer packets. Public red-team write-ups, eval suites, or robotics safety publications. Companion, teen-facing, or voice-agent experience. Pacific Time hours for client calls.
We do not hire on enthusiasm for “AI ethics” alone. We hire people who can defend a method when a lawyer is in the room.