Lead Engineer, AI Platform

Lever, Inc.

Italia

Remoto

EUR 136.000 - 167.000

Tempo pieno

46 ore fa
Candidati tra i primi
Generatore di candidature

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Supera i filtri ATS

Vantaggi offerti da questo lavoro

Annual $170,000 USD cash compensation
Equity with refresh grants
Fully remote work environment
35 days PTO per year
Company retreats

Descrizione del lavoro

Lever, Inc. is seeking a Lead Engineer, AI Platform in Italy. The role blends hands-on software engineering with technical leadership to build evaluation frameworks, observability tooling, and diagnostic infrastructure for AI agents in production across models and prompts.

You will guide the AI Quality team, drive experiments, and optimize cost, latency, and user experience while collaborating with globally distributed AI teams in a highly remote and asynchronous environment.

Competenze

  • 7+ years of experience building and shipping production software, ideally with LLM-powered agents.
  • Experience with complex tool-using AI systems involving planning, orchestration, or sub-agents.
  • Strong ability to demonstrate shipped software and how its reliability was measured.
  • Excellent English communication (CEFR C2).

Mansioni

  • Design, build, and own evaluation infrastructure and scoring systems for AI agents.
  • Develop observability and diagnostic capabilities across planning, tool selection, and agent trajectories.
  • Investigate failures and drive prototypes or prioritised improvements for AI products.
  • Expand datasets via annotation, AI-generated samples, and simulated conversations to improve evaluation coverage.
  • Lead and mentor the AI Quality engineering team while remaining hands-on.

Conoscenze

Leadership
LLM-powered agents
Python
Ruby on Rails
English proficiency

Strumenti

Braintrust
LangSmith
CI/CD pipelines

Descrizione del lavoro

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Lead Engineer, AI Platform based in Italy.

This role offers the opportunity to lead the engineering foundation behind reliable, measurable, and scalable AI-powered features. You’ll build evaluation frameworks, observability tooling, and diagnostic infrastructure that reveal how AI agents perform in real production environments. The position combines hands‑on software engineering with technical leadership and people management. You’ll investigate quality issues across complex agent workflows, develop datasets and evaluation systems, and run experiments across models, prompts, and agent architectures. You’ll also help optimize AI systems for cost, latency, reliability, and overall user experience. Working in a highly remote and asynchronous environment, you’ll collaborate closely with AI engineering teams while shaping the technical direction of a growing AI Quality function.

Accountabilities:
  • Design, build, and own evaluation infrastructure, including CI/CD pipelines, scorers, datasets, and systems for assessing AI agents from individual tool calls through complete multi-turn conversations.
  • Develop observability and diagnostic capabilities to identify exactly where quality issues occur across planning, execution, tool selection, and complex agent trajectories.
  • Investigate failures across sophisticated AI workflows and turn findings into technical prototypes, improvements, or clearly defined priorities for AI engineering teams.
  • Build and expand datasets through human annotation, AI-generated examples, and simulated conversations to increase evaluation coverage efficiently.
  • Develop structured experimentation frameworks for prompts, models, and agent harnesses, including evaluation of new and open-source models against production baselines.
  • Identify opportunities to improve AI system cost and latency through model selection, caching, routing, and other optimization strategies.
  • Set the technical direction and manage day-to-day priorities for the AI Quality engineering team while remaining actively involved in hands‑on development.
  • Partner closely with AI Core engineering teams to ensure changes to AI products can be measured effectively and demonstrably improve quality.
  • Establish engineering practices and evaluation approaches that support reliable, efficient, and scalable production AI systems.
Requirements:
  • 7+ years of experience building and shipping production software, ideally including LLM-powered agents capable of taking real actions within products.
  • Experience working with complex, tool-using AI systems involving multiple tools, planning, orchestration, or sub-agents rather than only simple, single-turn assistants.
  • Strong ability to demonstrate shipped software and explain how its effectiveness and reliability were measured.
  • Experience with Ruby on Rails and/or Python, with the ability to become productive quickly in technologies that may be new to you.
  • Experience building evaluation or observability infrastructure for ML/AI systems, including evaluation pipelines, scorers, dashboards, or CI/CD systems for evaluations.
  • Familiarity with evaluation frameworks such as Braintrust, LangSmith, or similar tools.
  • Experience designing datasets, annotation workflows, or labeling pipelines for machine learning or AI evaluation.
  • Ability to learn quickly, experiment extensively, and use empirical results to guide technical decisions.
  • Comfortable operating in a fast-paced environment with ambiguity and changing technical requirements.
  • Strong technical leadership and people-management capabilities, with the ability to balance team leadership and hands‑on engineering.
  • Excellent English proficiency in spoken, written, and reading communication, equivalent to CEFR C2 / ILR 5.
  • Strong alignment with a collaborative, ownership-oriented engineering culture.
Benefits:
  • Annual cash compensation of $170,000 USD, benchmarked to U.S. compensation levels regardless of location.
  • Equity in the company, including ongoing refresh grants.
  • 35 days of paid time off per year.
  • Fully remote work environment.
  • Significant flexibility and autonomy in how you organize your work.
  • Twice-yearly company retreats in international destinations.
  • Benefits supporting health, wellbeing, and professional development.
  • Opportunity to lead and grow an AI Quality engineering team while remaining hands‑on technically.
  • Exposure to advanced AI agents, evaluation infrastructure, observability, experimentation, and production AI optimization.
  • Opportunity to work with a globally distributed team across multiple countries and time zones.

We appreciate your interest and wish you the best!

Ottieni la revisione del curriculum gratis e riservata.

o trascina qui il file.

Similar jobs

Offerte di lavoro simili che vale la pena confrontare

Expert Team Lead, Engineering
Expert Team Lead, Engineering

Lever, Inc. • Italia

In loco
EUR 70.000 - 120.000
Remote working options where permitted
Equity opportunities
Medical/dental/vision coverage
Lead QA Engineer (AI native)
Lead QA Engineer (AI native)

Lever, Inc. • Italia

Remoto
EUR 90.000 - 130.000
Remote work
Vacation days 28
Wellness days 7
+4
Staff Platform Engineer
Staff Platform Engineer

Lever, Inc. • Italia

Remoto
EUR 90.000 - 140.000
Equity participation
Global remote work
Ownership culture
AI Engineer Senior - Fully Remote
AI Engineer Senior - Fully Remote

Aiclo • Lazio

Ibrido
EUR 70.000 - 110.000
Fully remote work
Autonomy in technical decisions
Direct collaboration with CTO
+3
Applied AI Engineer
Applied AI Engineer

Jobgether • Lombardia

In loco
EUR 55.000 - 75.000
Unlimited paid time off
Comprehensive health insurance
Professional development stipend
AI Platform Lead Engineer - Remote & Impactful
AI Platform Lead Engineer - Remote & Impactful

Lever, Inc. • Italia

Remoto
EUR 136.000 - 167.000
Annual $170,000 USD cash compensation
Equity with refresh grants
Fully remote work environment
+2
Senior AI Engineer (Agentic Systems)
Senior AI Engineer (Agentic Systems)

Frontiere s.r.l. • Milano, Roma, Napoli

In loco
EUR 90.000 - 130.000
Health insurance
Meal vouchers for office workdays
Dynamic work environment
+2
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Skillvue • Milano

In loco
EUR 70.000 - 90.000
Competitive compensation
Flexible work
Budget for conferences and training
+1
AI Research Engineer (Agentic Post-training)
AI Research Engineer (Agentic Post-training)

Jobgether • Milano

In loco
EUR 70.000 - 110.000
Remote-first environment
Global team
Competitive compensation
Engineering Manager AI
Engineering Manager AI

TEAMSYSTEM SPA • Milano

Ibrido
EUR 57.000 - 78.000
Wellbeing wallet
Personal wellbeing budget
Hybrid work
+2