Lead Engineer, AI Platform

Lever, Inc.

India

Remote

INR 14,451,000 - 18,304,000

Full time

38 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Annual equity refresh
Fully remote work environment
Health & wellbeing benefits
35 days PTO
Global company retreats

Job summary

Lever, Inc. is seeking a Lead Engineer, AI Platform based in India.

You will lead the engineering foundation behind reliable, scalable AI-powered features, building evaluation frameworks, observability tooling, and diagnostic infrastructure for production AI agents. You’ll combine hands-on software engineering with technical leadership and people management, guiding experiments across models, prompts, and agent architectures while optimizing cost and latency.

Qualifications

  • 7+ years of production software experience, including LLM-powered agents.
  • Experience with multi-tool AI systems involving planning and orchestration.
  • Proven track record shipping reliable software with measurable impact.
  • Experience with Ruby on Rails and/or Python.
  • Built evaluation or observability infrastructure for ML/AI systems.
  • Familiarity with Braintrust, LangSmith, or similar tools.

Responsibilities

  • Design, build, and own evaluation infrastructure, including CI/CD pipelines, scorers, datasets, and systems for evaluating AI agents across multi-turn conversations.
  • Develop observability and diagnostic capabilities to locate quality issues in AI workflows.
  • Investigate failures and turn findings into prototypes or prioritized improvements for AI teams.
  • Expand datasets via human annotation, AI-generated examples, and simulated conversations to improve coverage.
  • Develop experimentation frameworks for prompts, models, and agent harnesses, and evaluate new/open-source models against production baselines.
  • Identify cost and latency improvements through model selection, caching, and routing strategies.
  • Set technical direction and manage day-to-day priorities for the AI Quality team while staying hands-on.
  • Collaborate with AI Core teams to measure product changes and improve quality.

Skills

Leadership
Experimentation
Empirical decisions
CEFR C2 English

Tools

Ruby on Rails
Python
Braintrust
LangSmith
ML evaluation pipelines

Job description

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Lead Engineer, AI Platform based in India.

This role offers the opportunity to lead the engineering foundation behind reliable, measurable, and scalable AI-powered features. You’ll build evaluation frameworks, observability tooling, and diagnostic infrastructure that reveal how AI agents perform in real production environments. The position combines hands‑on software engineering with technical leadership and people management. You’ll investigate quality issues across complex agent workflows, develop datasets and evaluation systems, and run experiments across models, prompts, and agent architectures. You’ll also help optimize AI systems for cost, latency, reliability, and overall user experience. Working in a highly remote and asynchronous environment, you’ll collaborate closely with AI engineering teams while shaping the technical direction of a growing AI Quality function.


Accountabilities:
  • Design, build, and own evaluation infrastructure, including CI/CD pipelines, scorers, datasets, and systems for assessing AI agents from individual tool calls through complete multi-turn conversations.
  • Develop observability and diagnostic capabilities to identify exactly where quality issues occur across planning, execution, tool selection, and complex agent trajectories.
  • Investigate failures across sophisticated AI workflows and turn findings into technical prototypes, improvements, or clearly defined priorities for AI engineering teams.
  • Build and expand datasets through human annotation, AI-generated examples, and simulated conversations to increase evaluation coverage efficiently.
  • Develop structured experimentation frameworks for prompts, models, and agent harnesses, including evaluation of new and open-source models against production baselines.
  • Identify opportunities to improve AI system cost and latency through model selection, caching, routing, and other optimization strategies.
  • Set the technical direction and manage day-to-day priorities for the AI Quality engineering team while remaining actively involved in hands‑on development.
  • Partner closely with AI Core engineering teams to ensure changes to AI products can be measured effectively and demonstrably improve quality.
  • Establish engineering practices and evaluation approaches that support reliable, efficient, and scalable production AI systems.
Requirements:
  • 7+ years of experience building and shipping production software, ideally including LLM-powered agents capable of taking real actions within products.
  • Experience working with complex, tool-using AI systems involving multiple tools, planning, orchestration, or sub-agents rather than only simple, single-turn assistants.
  • Strong ability to demonstrate shipped software and explain how its effectiveness and reliability were measured.
  • Experience with Ruby on Rails and/or Python, with the ability to become productive quickly in technologies that may be new to you.
  • Experience building evaluation or observability infrastructure for ML/AI systems, including evaluation pipelines, scorers, dashboards, or CI/CD systems for evaluations.
  • Familiarity with evaluation frameworks such as Braintrust, LangSmith, or similar tools.
  • Experience designing datasets, annotation workflows, or labeling pipelines for machine learning or AI evaluation.
  • Ability to learn quickly, experiment extensively, and use empirical results to guide technical decisions.
  • Comfortable operating in a fast-paced environment with ambiguity and changing technical requirements.
  • Strong technical leadership and people‑management capabilities, with the ability to balance team leadership and hands‑on engineering.
  • Excellent English proficiency in spoken, written, and reading communication, equivalent to CEFR C2 / ILR 5.
  • Strong alignment with a collaborative, ownership-oriented engineering culture.
Benefits:
  • Annual cash compensation of $170,000 USD, benchmarked to U.S. compensation levels regardless of location.
  • Equity in the company, including ongoing refresh grants.
  • 35 days of paid time off per year.
  • Fully remote work environment.
  • Significant flexibility and autonomy in how you organize your work.
  • Twice‑yearly company retreats in international destinations.
  • Benefits supporting health, wellbeing, and professional development.
  • Opportunity to lead and grow an AI Quality engineering team while remaining hands‑on technically.
  • Exposure to advanced AI agents, evaluation infrastructure, observability, experimentation, and production AI optimization.
  • Opportunity to work with a globally distributed team across multiple countries and time zones.

How Jobgether works:

We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top‑fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.

We appreciate your interest and wish you the best!

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Architect / AI engineering lead
AI Architect / AI engineering lead

Lever, Inc. • India

Remote
INR 2,500,000 - 5,000,000
Fully remote contract position
International remote working
Sr. Gen AI Engineer
Sr. Gen AI Engineer

Lever, Inc. • India

Remote
INR 1,000,000 - 6,500,000
Freelance / contract engagement
Remote work within India
Exposure to LLMs and AI technologies
+2
Senior AI Engineer – Agentic Systems & LLM Applications
Senior AI Engineer – Agentic Systems & LLM Applications

Jobgether • India

On-site
INR 1,800,000 - 2,400,000
Fully remote
Global-first team
Autonomous agents projects
+1
Lead QA Engineer (AI native)
Lead QA Engineer (AI native)

Lever, Inc. • India

Remote
INR 4,500,000 - 6,000,000
28 days vacation per year
7 wellness days per year
Referral bonuses up to $5,000
+2
Senior QA AI Engineer
Senior QA AI Engineer

Lever, Inc. • India

Remote
INR 2,200,000 - 4,200,000
Remote work in India
AI testing focus
Mentorship program
+1
Senior AI Engineer
Senior AI Engineer

Lever, Inc. • India

On-site
INR 1,800,000 - 3,000,000
Full-time position in India
Remote work eligible
Healthcare benefits
+3
Expert Team Lead, SWE
Expert Team Lead, SWE

Lever, Inc. • India

Remote
INR 2,600,000 - 6,000,000
Competitive compensation
Equity participation
Medical, dental, and vision coverage
+3
AI Engineer (Remote, Full-Time) [HR163]
AI Engineer (Remote, Full-Time) [HR163]

Smart Working • Delhi

On-site
INR 1,200,000 - 1,800,000
Fixed Shifts
No Weekend Work
Day 1 Benefits: Laptop and full medical insurance
+2
AI Engineer (Remote, Full-Time) [HR163]
AI Engineer (Remote, Full-Time) [HR163]

Smart Working • Chennai District

On-site
INR 1,200,000 - 1,500,000
Fixed Shifts: 12:00 PM - 9:30 PM IST
No Weekend Work
Day 1 Benefits: Laptop and full medical insurance
+2
AI Engineer (Remote, Full-Time) [HR163]
AI Engineer (Remote, Full-Time) [HR163]

Smart Working • Ernakulam

On-site
INR 1,000,000 - 1,500,000
Fixed shifts
No weekend work
Laptop and full medical insurance
+2