Remote AI Benchmark Engineer - French Language Specialist

Lilt

France

Hybride

EUR 60 000 - 90 000

Temps partiel

Il y a 9 jours
Générateur de candidature

Une candidature complète en une minute — un CV et une lettre de motivation personnalisés, prêts à être envoyés.

Passez les filtres ATS

Avantages offerts par ce poste

Remote freelance
Flexible schedule
Competitive rates
Fast payments
Global project exposure

Résumé du poste

LILT is seeking experienced native-speaking software engineers to design, build, and validate multilingual benchmarks for Terminal-Bench tasks. You will create high-signal tasks and assets in your native language to measure robustness without English translation crutches.

This is a remote, freelance opportunity. You’ll work with a global team, set your own schedule, and receive competitive rates with prompt payments. No fixed hours; project-based collaboration across languages and locales.

Qualifications

  • 5+ years of industry experience in software engineering.
  • Proven track record at leading technology companies and/or graduation from top-tier engineering universities.
  • Native or near-native fluency, deep understanding of grammar and phrasing rules. High English proficiency.
  • Strong proficiency in Python, shell scripting, and data processing.
  • Extensive experience with Terminal/CLI-based development workflows and familiarity with coding agents.
  • Deep technical understanding of multilingual text processing pitfalls (Unicode, locale conventions, bidi handling).

Responsabilités

  • Task Engineering: Evaluating Coding Agents.
  • Asset Creation: Build environments with datasets and files in your native language, assets must stay in the target language.
  • Prompting & Translation: identify failure points in your native language.
  • Implementation & Verification: support reference implementations and deterministic verifier scripts.
  • Calibration & Execution: analyze logs and calibrate task difficulty across model tiers.
  • Quality Assurance: participate in a 4-layer QA process and automated checks.

Connaissances

Python
Shell scripting
Data processing
Terminal/CLI workflows
Coding agents
Multilingual text processing

Formation

Top-tier engineering universities

Description du poste

LILT is seeking experienced native-speaking software engineers to design, build, and validate multilingual benchmarks for Terminal-Bench tasks. You will create high-signal tasks and assets in your native language to measure robustness without English translation crutches.

This is a remote, freelance opportunity. You’ll work with a global team, set your own schedule, and receive competitive rates with prompt payments. No fixed hours; project-based collaboration across languages and locales.

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

Bilingual SME for AI Benchmarking in Professional Services
Bilingual SME for AI Benchmarking in Professional Services

LILT (Production) • France

À distance
EUR 50 000 - 70 000
Flexible work schedule
Prompt payments
Access to diverse projects
+2
Senior AI SME & Localization Engineer
Senior AI SME & Localization Engineer

LILT AI • Paris

Sur place
EUR 50 000 - 70 000
AI Benchmark Engineer | Native Language Specialist - French (France) - Remote
AI Benchmark Engineer | Native Language Specialist - French (France) - Remote

Lilt • France

Hybride
EUR 60 000 - 90 000
Remote freelance
Flexible schedule
Competitive rates
+2
Robotics Subject Matter Expert (French) - AI Benchmarking
Robotics Subject Matter Expert (French) - AI Benchmarking

LILT (Production) • France

À distance
EUR 40 000 - 60 000
Subject Matter Expert – Quantative/Scientific/Corporate (English/French) – Remote
Subject Matter Expert – Quantative/Scientific/Corporate (English/French) – Remote

LILT AI • Paris

Sur place
EUR 50 000 - 70 000
Remote Natural Sciences SME for AI Benchmarking
Remote Natural Sciences SME for AI Benchmarking

LILT (Production) • France

À distance
EUR 30 000 - 50 000
Flexible schedule
Prompt payments
Opportunity to work on impactful projects
+2
Robotics SME (French) - AI Benchmarking Expert
Robotics SME (French) - AI Benchmarking Expert

LILT AI • Paris

Sur place
EUR 60 000 - 80 000
Remote AI Trainer French - STEM Specialist
Remote AI Trainer French - STEM Specialist

Meridial • France

Sur place
Remote French Language Specialist – Freelance AI Trainer
Remote French Language Specialist – Freelance AI Trainer

Meridial • France

Sur place
Language Specialist (French) | $50/hr Remote
Language Specialist (French) | $50/hr Remote

Crossing Hurdles • France

À distance