Stand out for this role — generate a tailored resume and cover letter in about a minute.
OpenTeams seeks a Senior AI/ML Test and Evaluation Engineer to build and operate benchmarking and evaluation capabilities at the core of our AI platform. You will design automated metrics and human judgments to compare models and agentic workflows, surface failure modes, and ensure evaluations stay valid as models evolve.
This hands-on role uses an open‑source toolchain, reports to senior stakeholders, and helps decide which capabilities are ready to field, emphasizing robust, transparently
OpenTeams seeks a Senior AI/ML Test and Evaluation Engineer to build and operate benchmarking and evaluation capabilities at the core of our AI platform. You will design automated metrics and human judgments to compare models and agentic workflows, surface failure modes, and ensure evaluations stay valid as models evolve.
This hands-on role uses an open‑source toolchain, reports to senior stakeholders, and helps decide which capabilities are ready to field, emphasizing robust, transparently