Lead Engineer, AI Platform

Lever, Inc.

France

À distance

EUR 128 000 - 174 000

Plein temps

Il y a 44 heures
Soyez parmi les premiers à postuler
Générateur de candidature

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Passez les filtres ATS

Avantages offerts par ce poste

Annual equity refresh
Fully remote work
35 days PTO
Competitive benefits
Global team retreats

Résumé du poste

Lever, Inc. is seeking a Lead Engineer, AI Platform to lead the engineering foundation behind reliable AI-powered features in a fully remote role based in France. You will build evaluation frameworks, observability tooling, and diagnostic infrastructure for AI agents operating in production environments.

You’ll combine hands-on software development with technical leadership and people management, shaping the AI Quality function while collaborating with AI Core teams across time zones.

Qualifications

  • 7+ years of experience building and shipping production software, ideally with LLM-powered agents.

Responsabilités

  • Design, build, and own evaluation infrastructure for AI agents across tool calls and multi-turn conversations.
  • Develop observability and diagnostic capabilities to identify quality issues across planning, execution, and tool use.
  • Investigate failures in AI workflows and translate findings into prototypes or priorities for AI teams.
  • Build datasets via human annotation, AI-generated examples, and simulated conversations to expand coverage.
  • Develop experimentation frameworks for prompts, models, and agent harnesses with production baselines.
  • Identify opportunities to reduce AI system cost and latency through optimization strategies.
  • Set technical direction and manage day-to-day priorities for the AI Quality team while staying hands-on.
  • Collaborate with AI Core teams to measure changes and improve quality.
  • Establish engineering practices and evaluation approaches for reliable production AI.

Connaissances

LLM-powered agents
Technical leadership
People management
English proficiency (C2)

Outils

Ruby on Rails
Python
Braintrust
LangSmith

Description du poste

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Lead Engineer, AI Platform based in France.

This role offers the opportunity to lead the engineering foundation behind reliable, measurable, and scalable AI-powered features. You’ll build evaluation frameworks, observability tooling, and diagnostic infrastructure that reveal how AI agents perform in real production environments. The position combines hands‑on software engineering with technical leadership and people management. You’ll investigate quality issues across complex agent workflows, develop datasets and evaluation systems, and run experiments across models, prompts, and agent architectures. You’ll also help optimize AI systems for cost, latency, reliability, and overall user experience. Working in a highly remote and asynchronous environment, you’ll collaborate closely with AI engineering teams while shaping the technical direction of a growing AI Quality function.

Accountabilities
  • Design, build, and own evaluation infrastructure, including CI/CD pipelines, scorers, datasets, and systems for assessing AI agents from individual tool calls through complete multi-turn conversations.
  • Develop observability and diagnostic capabilities to identify exactly where quality issues occur across planning, execution, tool selection, and complex agent trajectories.
  • Investigate failures across sophisticated AI workflows and turn findings into technical prototypes, improvements, or clearly defined priorities for AI engineering teams.
  • Build and expand datasets through human annotation, AI-generated examples, and simulated conversations to increase evaluation coverage efficiently.
  • Develop structured experimentation frameworks for prompts, models, and agent harnesses, including evaluation of new and open-source models against production baselines.
  • Identify opportunities to improve AI system cost and latency through model selection, caching, routing, and other optimization strategies.
  • Set the technical direction and manage day‑to‑day priorities for the AI Quality engineering team while remaining actively involved in hands‑on development.
  • Partner closely with AI Core engineering teams to ensure changes to AI products can be measured effectively and demonstrably improve quality.
  • Establish engineering practices and evaluation approaches that support reliable, efficient, and scalable production AI systems.
Requirements
  • 7+ years of experience building and shipping production software, ideally including LLM-powered agents capable of taking real actions within products.
  • Experience working with complex, tool‑using AI systems involving multiple tools, planning, orchestration, or sub-agents rather than only simple, single‑turn assistants.
  • Strong ability to demonstrate shipped software and explain how its effectiveness and reliability were measured.
  • Experience with Ruby on Rails and/or Python, with the ability to become productive quickly in technologies that may be new to you.
  • Experience building evaluation or observability infrastructure for ML/AI systems, including evaluation pipelines, scorers, dashboards, or CI/CD systems for evaluations.
  • Familiarity with evaluation frameworks such as Braintrust, LangSmith, or similar tools.
  • Experience designing datasets, annotation workflows, or labeling pipelines for machine learning or AI evaluation.
  • Ability to learn quickly, experiment extensively, and use empirical results to guide technical decisions.
  • Comfortable operating in a fast‑paced environment with ambiguity and changing technical requirements.
  • Strong technical leadership and people‑management capabilities, with the ability to balance team leadership and hands‑on engineering.
  • Excellent English proficiency in spoken, written, and reading communication, equivalent to CEFR C2 / ILR 5.
  • Strong alignment with a collaborative, ownership‑oriented engineering culture.
Benefits
  • Annual cash compensation of $170,000 USD, benchmarked to U.S. compensation levels regardless of location.
  • Equity in the company, including ongoing refresh grants.
  • 35 days of paid time off per year.
  • Fully remote work environment.
  • Significant flexibility and autonomy in how you organize your work.
  • Twice‑yearly company retreats in international destinations.
  • Benefits supporting health, wellbeing, and professional development.
  • Opportunity to lead and grow an AI Quality engineering team while remaining hands‑on technically.
  • Exposure to advanced AI agents, evaluation infrastructure, observability, experimentation, and production AI optimization.
  • Opportunity to work with a globally distributed team across multiple countries and time zones.
Obtenez votre examen gratuit et confidentiel de votre CV.

ou faites glisser et déposez votre fichier ici.

Similar jobs

Postes similaires à comparer

Lead QA Engineer (AI native)
Lead QA Engineer (AI native)

Lever, Inc. • France

À distance
EUR 90 000 - 140 000
Fully remote
Vacation: 28 days per year
Wellness days: 7 per year
+7
Senior AI Engineer (f/m/x)
Senior AI Engineer (f/m/x)

Lever, Inc. • France

À distance
EUR 90 000 - 130 000
Remote-first
Office Vienna (dog-friendly)
Learning budget
+4
Senior AI Software Engineer
Senior AI Software Engineer

GetVocal AI Ltd. • Paris

Sur place
EUR 90 000 - 130 000
25 days holiday
Private healthcare
Diversified, international team
+1
Staff Platform Engineer
Staff Platform Engineer

Lever, Inc. • France

À distance
EUR 110 000 - 150 000
Equity participation
Remote work option
Collaboration with AI professionals
+3
(Senior or Staff) Backend Engineer, AI tooling
(Senior or Staff) Backend Engineer, AI tooling

Lever, Inc. • France

Sur place
EUR 120 000 - 150 000
RSUs
Remote Romania
25 days leave
+2
Lead AI engineer
Lead AI engineer

Mirakl • Paris

Sur place
EUR 90 000 - 100 000
Permanent contract (CDI)
Hybrid work schedule (4 days on-site)
Remote AI Platform Engineer — Quality & Observability Lead
Remote AI Platform Engineer — Quality & Observability Lead

Lever, Inc. • France

À distance
EUR 128 000 - 174 000
Annual equity refresh
Fully remote work
35 days PTO
+2
AI Architect / AI engineering lead
AI Architect / AI engineering lead

Lever, Inc. • France

À distance
EUR 110 000 - 170 000
Fully remote
Lead AI architecture
LLM and RAG work
+4
Expert Team Lead, SWE
Expert Team Lead, SWE

Lever, Inc. • France

Hybride
EUR 90 000 - 140 000
Competitive compensation
Equity participation
Comprehensive health coverage
+1
Fullstack Software Engineer (x/f/m) - Ops AI Platform
Fullstack Software Engineer (x/f/m) - Ops AI Platform

Alan • Limoges

Sur place
EUR 55 000 - 75 000
Generous equity packages
Flexible office options
Comprehensive health insurance
+2