Transformez ce poste en entretien — un CV et une lettre de motivation conçus selon ce que cet employeur recherche.
Odixcity Consulting seeks an experienced LLM Evaluator to assess, analyze, and improve large language model performance. You will evaluate AI-generated content for factuality, coherence, safety, and alignment with guidelines.
The role involves ranking outputs, providing justified choices, and reporting recurring failures to help patch vulnerabilities. You will collaborate with QA to refine evaluation guidelines, participate in cross-checking sessions, and explore deeper causes behind errors to
Job Title: LLM Evaluator (Model Response Analyst)
Location: Remote (Worldwide)
Job Summary: We are seeking a detail-oriented and analytical LLM Evaluator to assess, analyze, and improve the performance of large language models (LLMs). In this role, you will evaluate AI-generated content for accuracy, coherence, factual reliability, bias, safety, and alignment with defined guidelines.