Remote AI Benchmark Engineer — French Language Specialist

LILT (Production)

France

Hybride

EUR 74 000 - 129 000

Temps partiel

Il y a 6 jours
Soyez parmi les premiers à postuler
Générateur de candidature

Obtenez une réponse de cet employeur — un CV et une lettre de motivation adaptés exactement à ce qu’il recherche.

Passez les filtres ATS

Résumé du poste

Lilt is seeking a native-speaking software engineer for a remote, freelance engagement to design, build, and validate multilingual benchmarks. You will craft high-signal tasks, datasets, and deterministic verifier scripts, using rubric-based judging only where necessary.

Responsibilities cover task engineering, asset creation in your native language, prompting and translation, implementation and verification, calibration of task difficulty, and rigorous quality assurance across a 4-layer review

Qualifications

  • 5+ years of software engineering experience.
  • Native or near-native fluency with deep understanding of grammar, register and phrasing rules.
  • Strong proficiency in Python, shell scripting, and data processing.

Responsabilités

  • Task Engineering: evaluate coding agents and build realistic task environments.
  • Asset Creation: develop language-specific assets that stay in the target language.
  • Prompting & Translation: identify failure points in native language.
  • Implementation & Verification: create robust reference implementations and deterministic verifier scripts.
  • Calibration & Execution: adjust task difficulty across model tiers with standard configurations.
  • Quality Assurance: participate in a 4-layer human and automated QC process.

Connaissances

Python
Shell scripting
Data processing
Terminal/CLI workflows
Multilingual NLP
Unicode normalization
Locale handling
RTL handling

Outils

Verifier scripts

Description du poste

Lilt is seeking a native-speaking software engineer for a remote, freelance engagement to design, build, and validate multilingual benchmarks. You will craft high-signal tasks, datasets, and deterministic verifier scripts, using rubric-based judging only where necessary.

Responsibilities cover task engineering, asset creation in your native language, prompting and translation, implementation and verification, calibration of task difficulty, and rigorous quality assurance across a 4-layer review

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

Remote AI Benchmark Engineer - French Language Specialist
Remote AI Benchmark Engineer - French Language Specialist

Lilt • France

Hybride
EUR 60 000 - 90 000
Remote freelance
Flexible schedule
Competitive rates
+2
Remote AI Benchmark Engineer — French Language Specialist
Remote AI Benchmark Engineer — French Language Specialist

LILT (Production) • France

Hybride
EUR 69 000 - 138 000
Bilingual SME for AI Benchmarking in Professional Services
Bilingual SME for AI Benchmarking in Professional Services

LILT (Production) • France

À distance
EUR 50 000 - 70 000
Flexible work schedule
Prompt payments
Access to diverse projects
+2
AI Benchmark Engineer - Native Language Specialist - French (France) - Remote
AI Benchmark Engineer - Native Language Specialist - French (France) - Remote

LILT (Production) • France

Hybride
EUR 74 000 - 129 000
Senior AI SME & Localization Engineer
Senior AI SME & Localization Engineer

LILT AI • Paris

Sur place
EUR 50 000 - 70 000
Robotics Subject Matter Expert (French) - AI Benchmarking
Robotics Subject Matter Expert (French) - AI Benchmarking

LILT (Production) • France

À distance
EUR 40 000 - 60 000
Robotics SME (French) - AI Benchmarking Expert
Robotics SME (French) - AI Benchmarking Expert

LILT AI • Paris

Sur place
EUR 60 000 - 80 000
AI Benchmark Engineer | Native Language Specialist - French (France) - Remote
AI Benchmark Engineer | Native Language Specialist - French (France) - Remote

Lilt • France

Hybride
EUR 60 000 - 90 000
Remote freelance
Flexible schedule
Competitive rates
+2
Remote Natural Sciences SME for AI Benchmarking
Remote Natural Sciences SME for AI Benchmarking

LILT (Production) • France

À distance
EUR 30 000 - 50 000
Flexible schedule
Prompt payments
Opportunity to work on impactful projects
+2
Subject Matter Expert – Quantative/Scientific/Corporate (English/French) – Remote
Subject Matter Expert – Quantative/Scientific/Corporate (English/French) – Remote

LILT AI • Paris

Sur place
EUR 50 000 - 70 000