Lead Engineer, AI Platform

Lever, Inc.

Deutschland

Remote

EUR 129.000 - 174.000

Vollzeit

vor 41 Stunden
Sei unter den ersten Bewerbenden
Bewerbungsgenerator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Schaffe es an den ATS-Filtern vorbei

Benefits dieser Stelle

Fully remote work
35 days paid time off
Equity / stock options
Global team collaboration
Flexible schedule
Professional development support

Zusammenfassung

Lever, Inc. in Germany seeks a Lead Engineer, AI Platform to drive evaluation infrastructure, observability tooling, and diagnostic systems for AI agents in production.

You will balance hands-on software development with technical leadership and people management, shaping the AI Quality function within a global, async team. You will design evaluation pipelines, develop datasets, run experiments, and optimize AI systems for cost, latency, and user experience while collaborating with AI Core teams

Qualifikationen

  • 7+ years building and shipping production software, incl. LLM-powered agents.
  • Experience with tool-using AI systems involving orchestration or sub-agents.
  • Ability to quantify effectiveness and reliability of software.
  • Experience with Rails or Python, quickly productive in new tech.
  • Experience designing evaluation/observability infrastructure for ML/AI systems.
  • Familiarity with Braintrust, LangSmith or similar tools.
  • Strong leadership and people management in fast-paced, ambiguous env.
  • Excellent English proficiency (CEFR C2).

Aufgaben

  • Design, build, and own evaluation infrastructure and datasets for AI agents.
  • Develop observability and diagnostic capabilities across planning, execution, and tool selection.
  • Investigate failures and translate findings into prototypes or priorities.
  • Expand datasets via annotation and AI-generated simulations.
  • Develop experimentation frameworks for prompts, models, and agent harnesses.
  • Identify cost and latency optimization opportunities for AI systems.
  • Set technical direction and manage day-to-day AI Quality engineering priorities.
  • Collaborate with AI Core teams to measure impact of changes.
  • Establish engineering practices for reliable, scalable production AI.

Kenntnisse

LLM-powered agents
Observability engineering
Leadership
English proficiency (CEFR C2)

Tools

Ruby on Rails
Python
Braintrust
LangSmith

Jobbeschreibung

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Lead Engineer, AI Platform based in Germany.

This role offers the opportunity to lead the engineering foundation behind reliable, measurable, and scalable AI-powered features. You’ll build evaluation frameworks, observability tooling, and diagnostic infrastructure that reveal how AI agents perform in real production environments. The position combines hands-on software engineering with technical leadership and people management. You’ll investigate quality issues across complex agent workflows, develop datasets and evaluation systems, and run experiments across models, prompts, and agent architectures. You’ll also help optimize AI systems for cost, latency, reliability, and overall user experience. Working in a highly remote and asynchronous environment, you’ll collaborate closely with AI engineering teams while shaping the technical direction of a growing AI Quality function.

Accountabilities:
  • Design, build, and own evaluation infrastructure, including CI/CD pipelines, scorers, datasets, and systems for assessing AI agents from individual tool calls through complete multi-turn conversations.
  • Develop observability and diagnostic capabilities to identify exactly where quality issues occur across planning, execution, tool selection, and complex agent trajectories.
  • Investigate failures across sophisticated AI workflows and turn findings into technical prototypes, improvements, or clearly defined priorities for AI engineering teams.
  • Build and expand datasets through human annotation, AI-generated examples, and simulated conversations to increase evaluation coverage efficiently.
  • Develop structured experimentation frameworks for prompts, models, and agent harnesses, including evaluation of new and open-source models against production baselines.
  • Identify opportunities to improve AI system cost and latency through model selection, caching, routing, and other optimization strategies.
  • Set the technical direction and manage day-to-day priorities for the AI Quality engineering team while remaining actively involved in hands-on development.
  • Partner closely with AI Core engineering teams to ensure changes to AI products can be measured effectively and demonstrably improve quality.
  • Establish engineering practices and evaluation approaches that support reliable, efficient, and scalable production AI systems.
Requirements:
  • 7+ years of experience building and shipping production software, ideally including LLM-powered agents capable of taking real actions within products.
  • Experience working with complex, tool-using AI systems involving multiple tools, planning, orchestration, or sub-agents rather than only simple, single-turn assistants.
  • Strong ability to demonstrate shipped software and explain how its effectiveness and reliability were measured.
  • Experience with Ruby on Rails and/or Python, with the ability to become productive quickly in technologies that may be new to you.
  • Experience building evaluation or observability infrastructure for ML/AI systems, including evaluation pipelines, scorers, dashboards, or CI/CD systems for evaluations.
  • Familiarity with evaluation frameworks such as Braintrust, LangSmith, or similar tools.
  • Experience designing datasets, annotation workflows, or labeling pipelines for machine learning or AI evaluation.
  • Ability to learn quickly, experiment extensively, and use empirical results to guide technical decisions.
  • Comfortable operating in a fast-paced environment with ambiguity and changing technical requirements.
  • Strong technical leadership and people-management capabilities, with the ability to balance team leadership and hands-on engineering.
  • Excellent English proficiency in spoken, written, and reading communication, equivalent to CEFR C2 / ILR 5.
  • Strong alignment with a collaborative, ownership-oriented engineering culture.
Benefits:
  • Annual cash compensation of $170,000 USD, benchmarked to U.S. compensation levels regardless of location.
  • Equity in the company, including ongoing refresh grants.
  • 35 days of paid time off per year.
  • Fully remote work environment.
  • Significant flexibility and autonomy in how you organize your work.
  • Twice-yearly company retreats in international destinations.
  • Benefits supporting health, wellbeing, and professional development.
  • Opportunity to lead and grow an AI Quality engineering team while remaining hands‑on technically.
  • Exposure to advanced AI agents, evaluation infrastructure, observability, experimentation, and production AI optimization.
  • Opportunity to work with a globally distributed team across multiple countries and time zones.
How Jobgether works:

We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-filling candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.

We appreciate your interest and wish you the best!

Data Privacy Notice:

By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

AI Augmented Software Engineer [gn] Data Intelligence Platform
AI Augmented Software Engineer [gn] Data Intelligence Platform

Jobgether • Deutschland

Vor Ort
EUR 90.000 - 140.000
Competitive salary
Remote or hybrid options
AI-assisted engineering exposure
+2
AI Trainers Network - English
AI Trainers Network - English

Lever, Inc. • Deutschland

Remote
EUR 30.000 - 54.000
Flexible remote projects
Global contributor network
Varied project types
Junior Legal Counsel, AI & GDPR
Junior Legal Counsel, AI & GDPR

Lever, Inc. • Deutschland

Remote
EUR 35.000 - 45.000
Flexible WorkFlex
Well-Being Day
Gym subsidies
+3
AI Agent Evaluation Analyst (Freelance)
AI Agent Evaluation Analyst (Freelance)

Mindrift • Deutschland

Vor Ort
EUR 29.551 - 61.468
Competitive pay up to $52/hour
Flexible schedule
Experience with advanced AI projects
AI Process Forward Deployed Engineer
AI Process Forward Deployed Engineer

United States Digital Space LLC • Deutschland

Remote
EUR 90.000 - 130.000
Fully remote environment
Home office budget
Flexible time off
+4
Customer Success - Strategic - EMEA
Customer Success - Strategic - EMEA

Lever, Inc. • Deutschland

Vor Ort
EUR 90.000 - 140.000
AI-first environment
Global team
Professional development stipend
+2
Lead AI Engineer / Senior AI Engineer
Lead AI Engineer / Senior AI Engineer

AI.IMPACT • Hamburg

Vor Ort
EUR 120.000 - 180.000
Collaborative team environment
Competitive compensation with equity
Backend Engineer (AI Agents/Workflows)
Backend Engineer (AI Agents/Workflows)

Andercore • Berlin

Vor Ort
EUR 55.000 - 75.000
Senior AI Engineer (German Speaking)
Senior AI Engineer (German Speaking)

Lufinity AI • München

Hybrid
EUR 90.000 - 130.000
Backend Engineer (AI Agents/Workflows)
Backend Engineer (AI Agents/Workflows)

Andercore GmbH • Berlin

Vor Ort
EUR 60.000 - 80.000