Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
Get past ATS filters
Job summary
A leading evaluation firm in Canada is seeking an experienced evaluator to assess the accuracy and relevance of LLM-generated responses. The ideal candidate will have native-level fluency in French, strong English proficiency, and proven experience with large language models. Responsibilities include evaluating response quality, fact-checking, and creating structured feedback. Candidates should possess excellent writing skills and analytical thinking, and be able to work across various domains. This role offers competitive compensation and the opportunity to contribute to AI evaluations.
Qualifications
Native-level or near-native fluency in French with strong English proficiency.
Proven experience using large language models.
Excellent writing skills and ability to provide structured feedback.
Strong attention to detail and analytical thinking.
Ability to work across multiple topics and domains.
Responsibilities
Evaluate LLM-generated responses for accuracy, relevance, and effectiveness.
Perform fact-checking using reliable sources and tools.
Create high-quality evaluation data by annotating responses.
Assess reasoning quality, tone, clarity, and completeness of outputs.
Ensure responses follow expected conversational behavior and guidelines.
Skills
Fluency in French
Writing skills
Attention to detail
Analytical thinking
Experience with large language models
Job description
A leading evaluation firm in Canada is seeking an experienced evaluator to assess the accuracy and relevance of LLM-generated responses. The ideal candidate will have native-level fluency in French, strong English proficiency, and proven experience with large language models. Responsibilities include evaluating response quality, fact-checking, and creating structured feedback. Candidates should possess excellent writing skills and analytical thinking, and be able to work across various domains. This role offers competitive compensation and the opportunity to contribute to AI evaluations.