Quality Manager, Applied AI

LILT

Boston (MA)

Hybrid

USD 140,000 - 170,000

Full time

8 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

LILT is a Boston-based AI translation company seeking a senior(quality) leader to oversee quality programs in ML/LLM contexts. The role focuses on building quality workflows, defining tooling and metrics, and scaling quality across multiple concurrent programs.

The candidate will ensure go/no-go sign-offs before shipping and drive remediation of customer disputes with data-driven analysis. Boston hybrid work arrangement; U.S. work authorization required.

Qualifications

  • Education: B.S./M.S. in a quantitative field (CS, Stats, Data Science, Eng, Linguistics).
  • Applied AI / evaluation experience: 6+ years building or operating quality programs for ML/LLM systems.
  • Quality systems leadership: design and run end-to-end quality frameworks across programs.
  • Measurement & statistics: reliability and agreement methods; reporting for execs.
  • Rater/annotator operations: calibration, adjudication, retraining, drift monitoring, integrity.
  • Tooling & automation: hands-on labeling/eval tooling and deterministic validations.
  • Data governance: dataset/version management, traceability, privacy/compliance.
  • Stakeholder management: translate goals into measurable targets and negotiate tradeoffs.

Responsibilities

  • Ensure Workflows Are Built for Quality, with measurable outcomes.
  • Own Technical Tooling & Quality Infrastructure, collecting metrics.
  • Own Operational Quality at Scale across concurrent programs.
  • Delivery sign-off: documented go/no-go decisions and post-sign-off actions.
  • Own Customer Quality Remediation with quantitative analysis and closed disputes.

Skills

Applied AI
Quality leadership
Measurement statistics
Rater/annotator ops
Tooling automation
Data governance
Stakeholder management
Multilingual
Programming language

Education

B.S./M.S. in quantitative field

Tools

Label Studio

Job description

About LILT

AI is changing how the world communicates - and LILT is leading that transformation.

We're on a mission to make the world's information accessible to everyone, regardless of the language they speak. We use cutting-edge AI, machine translation, and human-in-the-loop expertise to translate content faster, more accurately, and more cost-effectively without compromising on brand, voice, or quality.

At LILT, we empower our teammates with leading tools, global collaboration, and growth opportunities to do their best work. Our company virtues-Work together, win together; Find a way or make one; Dance in the customer's shoes; Quicker than they expect; Quality is Job 1-guide everything we do. We are trusted by Intel Corporation, Canva, the United States Department of Defense, the United States Air Force, ASICS, and hundreds of global Enterprises. Backed by Sequoia, Intel Capital, and Redpoint, we're building a category-defining company in a $50B+ global translation market being redefined by AI.

Key Responsibilities
  • Ensure Workflows Are Built for Quality

    • Outcome: Every Applied AI deliverable ships with measured quality: rater agreement meets agreed thresholds per dimension, rubric scores are reproducible across raters, and reported quality figures hold up under customer audit.

  • Own Technical Tooling & Quality Infrastructure

    • Outcome: Quality metrics for every active program (defect rate, acceptance rate, rework, reviewer reliability, throughput, SLA adherence) are available without manual compilation, and systemic quality issues are detected from workflow data before a customer reports them.

  • Own Operational Quality at Scale

    • Outcome: Concurrent programs meet their quality targets within agreed cost and throughput. Acceptance rates hold at or above target, rework declines over time, and standard QC on standardized workflows is executed and interpreted by Production staff without Quality Manager involvement by the end of the second quarter in role.

  • Delivery sign-off: Every deliverable has a documented go/no-go decision before shipping, based on programmatic QA and a readiness memo with flags. Post-sign-off quality escapes stay below an agreed threshold, and each escape results in a documented process change.

  • Own Customer Quality Remediation

    • Outcome: Customer quality disputes are closed with quantitative analysis, normally at first response. Re-adjudication completes within SLA, recurring defect classes decline release over release, and annotator integrity issues are detected internally before they reach a deliverable.

Qualifications
  • Education: B.S./M.S. in a quantitative field (CS, Statistics, Data Science, Engineering, Linguistics, HCI) or equivalent practical experience.

  • Applied AI / evaluation experience: 6+ years building or operating quality programs for ML/LLM systems (data labeling, benchmark creation, model evaluation, or human-in-the-loop pipelines).

  • Quality systems leadership: Proven ability to design and run end-to-end quality frameworks (rubrics, sampling plans, QC gates, acceptance criteria, escalation paths) across multiple concurrent programs.

  • Measurement & statistics: Strong grasp of reliability and agreement methods (e.g., Krippendorff's α, Cohen's κ), power/sample-size intuition, error analysis, and reporting for executive stakeholders.

  • Rater/annotator operations: Experience with calibration, adjudication, targeted retraining, drift monitoring, and integrity/fraud detection in high-throughput human evaluation programs.

  • Tooling & automation: Hands-on with labeling/eval tooling (e.g., Label Studio or similar) and designing deterministic validations; comfort partnering with engineering/research to implement quality instrumentation.

  • Data governance: Strong understanding of dataset/version management, traceability, privacy/compliance constraints, and audit-ready documentation for deliverables.

  • Stakeholder management: Excellent written and verbal communication; able to translate ambiguous research goals into measurable quality targets and negotiate tradeoffs with Applied AI, product, and external partners.

Preferred Skills
  • Fluency in multiple human languages.
  • Proficiency in a programming language and experience working with modern AI frameworks.
Where You’ll Work

This position is based in our Boston office and will be expected to work in the office in a hybrid capacity. LILT is hybrid with hubs in SF, NY, Indianapolis, Boston, London, and Berlin.

Authorization to work in the U.S. is a precondition of employment.

  • Boston (highly preferred)
  • Others: SF Bay Area, New York, Indianapolis, London
Our Story

Our founders, Spence and John met at Google working on Google Translate. As researchers at Stanford and Berkeley, they both worked on language technology to make information accessible to everyone. While together at Google, they were amazed to learn that Google Translate wasn’t used for enterprise products and services inside the company.The quality just wasn’t there. So they set out to build something better. LILT was born.

LILT has been a machine learning company since its founding in 2015. At the time, machine translation didn’t meet the quality standard for enterprise translations, so LILT assembled a cutting-edge research team tasked with closing that gap. While meeting customer demand for translation services, LILT has prioritized investments in Large Language Models, human-in-the-loop systems, and now agentic AI.

With AI innovation accelerating and enterprise demand growing, the next phase of LILT’s journey is just beginning.

Our Tech

What sets our platform apart:

  • Brand-aware AI that learns your voice, tone, and terminology to ensure every translation is accurate and consistent
  • Agentic AI workflows that automate the entire translation process from content ingestion to quality review to publishing
  • 100+ native integrations with systems like Adobe Experience Manager, Webflow, Salesforce, GitHub, and Google Drive to simplify content translation
  • Human-in-the-loop reviews via our global network of professional linguists, for high-impact content that requires expert review
LILT in the News
  • Featured in The Software Report's Top 100 Software Companies!
  • LILT makes it onto the Inc. 5000 List.
  • LILT's continues to be an intellectual powerhouse, holding numerous patents that help power the most efficient and sophisticated AI and language models in the industry.
  • Check out all our news on our website.

Information collected and processed as part of your application process, including any job applications you choose to submit, is subject to LILT's Privacy Policy at https://lilt.com/legal/privacy.

At LILT, we are committed to a fair, inclusive, and transparent hiring process. As part of our recruitment efforts, we may use artificial intelligence (AI) and automated tools to assist in the evaluation of applications, including résumé screening, assessment scoring, and interview analysis. These tools are designed to support human decision-making and help us identify qualified candidates efficiently and objectively. All final hiring decisions are made by people. If you have any concerns, require accommodations, or would like to opt-out of the use of AI in our hiring process, please let us know at recruiting@lilt.com.

LILT is an equal opportunity employer. We extend equal opportunity to all individuals without regard to an individual's race, religion, color, national origin, ancestry, sex, sexual orientation, gender identity, age, physical or mental disability, medical condition, genetic characteristics, veteran or marital status, pregnancy, or any other classification protected by applicable local, state or federal laws. We are committed to the principles of fair employment and the elimination of all discriminatory practices.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Production Manager, Applied AI
Production Manager, Applied AI

LILT • Boston (MA)

Hybrid
USD 120,000 - 180,000
Quality Manager, Applied AI
Quality Manager, Applied AI

lilt-corporate • United States

On-site
USD 150,000 - 210,000
Senior Forward Deployed Engineer
Senior Forward Deployed Engineer

LILT • Washington

Hybrid
USD 120,000 - 180,000
Competitive compensation and equity
Medical, dental, vision benefits
Paid parental leave
Head of Talent Acquisition
Head of Talent Acquisition

Lilt---Lega-Italiana-Per-La-Lotta-Contro-I-Tumori-1 • San Francisco (CA)

On-site
USD 180,000 - 240,000
Senior Technical Program Manager, Applied AI
Senior Technical Program Manager, Applied AI

LILT • Washington

Hybrid
USD 180,000 - 240,000
Hybrid work model
Competitive salary
Growth opportunities
Head of Talent Acquisition
Head of Talent Acquisition

LILT AI • San Francisco (CA)

On-site
USD 180,000 - 260,000
Senior Forward Deployed Engineer
Senior Forward Deployed Engineer

Lilt • Washington, Northern (KY)

Hybrid
USD 140,000 - 190,000
Competitive compensation + equity
401(k) matching
Flexible time off
+3
Production Manager, Applied AI
Production Manager, Applied AI

lilt-corporate • United States

On-site
USD 90,000 - 130,000
Account Development Representative (ADR)
Account Development Representative (ADR)

LILT AI • New York (NY)

On-site
USD 60,000 - 95,000
AI Training Contributor - Serbian - Remote
AI Training Contributor - Serbian - Remote

LILT, Inc. • Washington, Northern (KY)

Hybrid
USD 39,000 - 55,000