Engineering Manager, AI Evaluation Systems

Cervin

Northern (KY)

Hybrid

USD 163,000 - 224,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

RSUs
Health insurance
Vision insurance
Dental insurance
Mental health benefits

Job summary

LaunchDarkly is seeking an Engineering Manager for the AgentControl Evaluations team. You will lead a group building offline evaluations, playgrounds, LLM-as-judges, and annotated datasets, shaping the core product's evaluation stack.

You'll partner with product and design to scope work, own delivery, and drive a culture of high performance. Expect collaboration across AI, observability, and feature management teams, with Go, Python, and Typescript at the core.

Qualifications

  • 3+ years of software engineering experience
  • 2+ years managing a backend or infrastructure team
  • Ability to translate business goals into engineering plans with reliable estimations
  • Strong communication skills in distributed, cross-timezone teams
  • Experience coaching and developing engineers, including performance management
  • Familiarity with observability practices (metrics, tracing, alerting, logging)
  • Familiarity with AI providers (Anthropic, OpenAI, Bedrock, Gemini) and AI tooling (Langchain, Arize)
  • Experience with Go and Python or similar backend languages (Java, Rust, Ruby)

Responsibilities

  • Lead and develop a team of engineers with coaching and career growth.
  • Partner with PM and Design to scope and sequence roadmap items with reliable estimates.
  • Own delivery of features within timelines and quality requirements.
  • Collaborate with adjacent teams on infrastructure decisions and cross-team dependencies.
  • Communicate roadmaps, risks, and progress to engineering leadership and stakeholders.
  • Establish team norms to promote collaboration and reduce silos.
  • Improve reliability, observability, and incident response for production systems.
  • Participate in hiring to grow the team and raise engineering bar.

Skills

Team leadership
Roadmap collaboration
Observability
AI tooling familiarity
Go & Python

Tools

Go
Python
Typescript

Job description

LaunchDarkly is seeking an Engineering Manager for the AgentControl Evaluations team. You will lead a group building offline evaluations, playgrounds, LLM-as-judges, and annotated datasets, shaping the core product's evaluation stack.

You'll partner with product and design to scope work, own delivery, and drive a culture of high performance. Expect collaboration across AI, observability, and feature management teams, with Go, Python, and Typescript at the core.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Engineering Manager, AgentControl Evaluations
Engineering Manager, AgentControl Evaluations

LaunchDarkly Group • Northern (KY)

Hybrid
USD 163,000 - 224,000
RSUs
Health insurance
Dental insurance
+2
Engineering Manager, AgentControl
Engineering Manager, AgentControl

Cervin • Northern (KY)

Hybrid
USD 163,000 - 224,000
RSUs
Health insurance
Vision insurance
+2
Engineering Manager, AgentControl New Remote - US
Engineering Manager, AgentControl New Remote - US

LaunchDarkly Group • Northern (KY)

Hybrid
USD 163,000 - 224,000
RSUs
Health insurance
Dental insurance
+2
Lead AI Solutions Engineer – AgentControl
Lead AI Solutions Engineer – AgentControl

LaunchDarkly • United States

On-site
USD 182,000 - 252,000
GenAI Full-Stack Engineer — Scale AI Configs & APIs
GenAI Full-Stack Engineer — Scale AI Configs & APIs

LaunchDarkly • United States

Hybrid
USD 145,000 - 236,000
Engineering Manager, AI Evaluation Systems
Engineering Manager, AI Evaluation Systems

Cursor • San Francisco (CA)

On-site
USD 130,000 - 160,000
Senior AI Engineer
Senior AI Engineer

arosplatforms | AI Consulting & Services • North Township (IN)

On-site
USD 120,000 - 170,000
Senior ML / Evaluation Engineer
Senior ML / Evaluation Engineer

Intellias • Town of Poland (NY)

On-site
USD 140,000 - 200,000
LLM Evaluation Engineering Lead
LLM Evaluation Engineering Lead

DeepRec.ai • Redwood City (CA)

On-site
USD 180,000 - 240,000
High autonomy
Strong technical peers
Meaningful equity
Staff Engineer - Next-Gen Experience Platform & UX
Staff Engineer - Next-Gen Experience Platform & UX

LaunchDarkly • United States

Remote
USD 183,000 - 251,000
RSUs
Health insurance
Dental insurance
+1