Stand out for this role — generate a tailored resume and cover letter in about a minute.
prolificacademic ltd seeks a senior AI/ML engineer to evaluate frontier LLMs, audit training code, and critique model outputs. You will focus on deep technical judgment and error tracing rather than building models, working remotely with flexible hours in the US.
You will analyze prompts, assist in RLHF workflows, and benchmark reasoning quality using defined metrics and taxonomies. A strong ML background is required.
This is a paid participant opportunity for senior AI and machine learning engineers based in Memphis who want to channel their production ML experience into evaluating and refining frontier large language models. The role emphasizes deep technical judgment, code auditing, and analytical critique rather than model development. It suits engineers seeking flexible, remote-friendly work that fits alongside other commitments.