Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
Leading AI Lab is seeking experienced software engineers to join an evaluation and annotation team focused on real-world software engineering, model evaluation, and AI-driven improvements.
Ideal candidates have deep hands-on coding experience, strong Python skills, and a track record in structured evaluation and high-quality feedback. This is a senior, contract-friendly role with collaboration across private LLMs and external models.
This is a contracting engagement - initially 6 months - with potential for long term engagement.
Location: Paris-based preferred; alternatively Europe remote for strong candidates
We are building and evaluating state-of-the-art large language models (LLMs) and are looking for experienced software engineers to join our evaluation and annotation team. This role sits at the intersection of real-world software engineering, model evaluation, and applied AI, and is critical to improving model reliability, reasoning, and code quality.
You will design challenging coding tasks, evaluate model outputs against rigorous benchmarks, identify failure modes, and contribute to reinforcement learning and model improvement workflows.
This is not a junior annotation role. We are looking for practitioners with deep hands-on coding experience who can think like both an engineer and an evaluator.
Company: Leading AI Lab