Staff Fullstack Engineer – Internal Tools

LILT

United States

On-site

USD 130,000 - 170,000

Full time

3 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

LILT is seeking a senior software engineer to join the Internal Tools group. You will own architecture and tooling for the AI benchmarking platform used by research teams, delivery, and operations.

The role focuses on end‑to‑end systems engineering, scaling infrastructure, and building self‑serve automation to accelerate annotator workflows. You will collaborate with delivery and engineering to set technical direction, extend pipelines, and raise the bar for quality and reliability across

Qualifications

  • 5+ years of professional full‑stack software engineering with end‑to‑end ownership.
  • Strong Python backend experience (FastAPI) + async SQLAlchemy with Postgres.
  • Solid React/TypeScript frontend and modern data fetching stacks.
  • Experience building and operating background job systems with idempotency.
  • Experience integrating APIs (GitHub, OAuth/SSO, LLMs) and webhook sync.
  • Ability to build Slack bot integrations for internal workflows.
  • CI/CD, containerization (Docker), and infrastructure‑as‑code (Terraform).
  • Ability to optimize memory usage in long‑running Python workers.

Responsibilities

  • Translate new benchmark and data‑quality requirements into technical designs.
  • Own internal workforce‑management tools for AI delivery and accounting flows.
  • Own architecture and long‑term direction across multiple services.
  • Extend audio/QA pipelines and QC tooling for data quality scoring.
  • Design and ship new modules to plug workflows into existing lifecycle.
  • Own API key provisioning and budget governance system.
  • Build ChatOps automation to accelerate annotator workflows.
  • Harden background workers and job processing infra.
  • Enforce testing, CI/CD, and IaC across codebases.

Skills

Full-stack experience
Python backend (FastAPI)
Async SQLAlchemy/Postgres
React/TypeScript
Background job systems
APIs & OAuth
Slack bot integrations
CI/CD & IaC
Memory management (numpy/ML env)
Internal platform tooling

Tools

Docker
Kubernetes
Helm
ArgoCD
GitOps
Terraform
PostgreSQL

Job description

About LILT

AI is changing how the world communicates — and LILT is leading that transformation.

We’re on a mission to make the world’s information accessible to everyone , regardless of the language they speak. We use cutting‑edge AI, machine translation, and human‑in‑the‑loop expertise to translate content faster, more accurately, and more cost‑effectively without compromising on brand, voice, or quality.

At LILT, we empower our teammates with leading tools, global collaboration, and growth opportunities to do their best work. Our company virtues—Work together, win together; Find a way or make one; Dance in the customer’s shoes; Quicker than they expect; Quality is Job 1 —guide everything we do. We are trusted by Intel Corporation, Canva, the United States Department of Defense, the United States Air Force, ASICS, and hundreds of global Enterprises. Backed by Sequoia, Intel Capital, and Redpoint, we’re building a category‑defining company in a $50B+ global translation market being redefined by AI.

As part of LILT’s Internal Tools group, you will build and run the tools that LILT’s own delivery, contributor‑ops, and engineering teams rely on to operate LILT’s Applied AI (AAI) benchmarking business — which builds and delivers multilingual benchmarks and evaluation data to frontier AI labs using LILT’s global network of subject‑matter experts. As the AAI business grows, expect the scope of these platforms to grow in parallel. This platform sits directly on the critical path of paid customer deliverables with real SLAs, maintained by a small, high‑leverage engineering team that produces reliable, production‑ready tooling. You will own significant surface area across both end‑to‑end: architecture, data‑pipeline design, and the hands‑on engineering that keeps this infrastructure reliable. This is a high‑impact, high‑visibility role: the reliability and craft you bring directly enables on‑time, high‑quality delivery for some of the most prominent AI labs in the world.

In this role, you will drive the long‑term technical strategy for internal platforms while working directly with the delivery, ops, and engineering teams across the business who depend on them daily. This team ships cross‑service automation in careful stages — advisory first, then assistive, only later decision‑relevant, always with a human fallback — and you’ll be expected to hold that same bar as you extend these systems. You will partner with a small existing team to raise the engineering bar across two very different runtimes, set technical direction, and make the calls that determine how this infrastructure scales as the business grows.

What You’ll Do

  • Partner with the benchmarking business’s researchers and TPMs to translate new benchmark and data‑quality requirements into scoped technical designs

  • Own the internal workforce‑management tools on our internal platform for AI delivery for hundreds of external contributors: vetting flows, candidate assessment, QC, payment/delivery tracking, and roster/reporting exports

  • Own architecture and long‑term technical direction across multiple services and the platform

  • Extend IAA and audio‑QA pipelines: annotator outlier detection, ASR sidecar enhancements, LLM‑based QC, and DNSMOS/librosa audio‑quality scoring

  • Design and ship new modules on the platform that plug new benchmark and vetting workflows into the existing multi‑stage review lifecycle

  • Own and extend the platform’s API key provisioning and budget‑governance system

  • Build self‑serve ChatOps‑style automation for the internal engineering org and contributor base to enable accelerated annotator workflows and query resolution

  • Harden background worker and job‑processing infrastructure

  • Set and enforce testing, CI/CD, and deployment practices across both codebases — from unit/integration testing through infrastructure‑as‑code and release automation

  • Raise the technical bar for a small, high‑leverage team through code review, design docs, and mentoring as the surface area grows

What We’re Looking For

Required
  • 5+ years of professional full‑stack software engineering experience, with a track record of owning production systems end‑to‑end across more than one runtime/language

  • Deep experience with a Python backend framework (FastAPI or comparable) plus async SQLAlchemy/Postgres, alongside production experience in at least one statically‑typed backend language (Go, Java, or similar)

  • Strong React/TypeScript frontend experience — component architecture, state management (Zustand, Redux, or similar), and a modern data‑fetching layer (TanStack Query or comparable)

  • Experience building and operating background job/worker systems (queue‑driven or polling‑based) with failure tolerance and idempotency in mind

  • Experience integrating with third‑party and platform APIs — including the GitHub API, OAuth/OIDC SSO, and at least one LLM API (Gemini, OpenAI, or similar) — handling auth, rate limits, and webhook‑driven sync

  • Experience building Slack (or comparable chat‑platform) bot integrations that automate internal workflows — resource provisioning, approvals, budget/TTL enforcement — with real operational guardrails, not just CRUD features

  • Comfort reading and extending applied‑statistics or ML‑adjacent code (agreement metrics, audio‑quality scoring, or comparable data‑quality tooling)

  • Solid grasp of CI/CD, containerized deployment (Docker, Helm, ArgoCD/GitOps or comparable), and infrastructure‑as‑code (Terraform or comparable)

  • Experience debugging and optimizing native‑library (numpy/scipy/onnxruntime‑class) memory growth in long‑running Python worker processes via safe, boundary‑aware process recycling — not just raising memory limits

  • Experience building and owning internal platforms/tools that increase leverage for a non‑engineering team (research, operations, support, data/workforce management, or similar) — not solely external‑customer‑facing product work

Strong Plus
  • Experience with audio/speech pipelines: ASR (Whisper or similar) or audio‑quality metrics (DNSMOS, librosa)

  • Experience building internal tools for managing a data‑labeling, annotation, or crowdsourced‑contributor workforce (vetting, QC, payments)

  • Experience with inter‑annotator agreement or statistical agreement metrics

  • Notification and delivery systems experience — Slack bot integrations, transactional email, and idempotent delivery guarantees

  • Experience designing abstractions over heterogeneous data sources with different consistency guarantees — e.g. a fully‑replayable event history vs. an observe‑only current‑state API requiring synthesized diffing — behind one common interface

  • Experience building ChatOps‑style automation — Slack or GitHub PR‑comment bot commands that trigger backend workflows or CI/CD runs

  • Experience implementing short‑lived, rotatable service‑to‑service JWT auth (key‑ID‑based rotation, replay‑protected tokens, fail‑fast config validation) alongside a separate human‑facing SSO flow in a paired service

  • Comfort owning both sides of a system with genuinely different runtimes without a large team to lean on

  • Prior experience as the primary or sole engineer on a small, high‑leverage internal platform

Bonus
  • Experience with LLM‑as‑judge or LLM‑based QA/review pipelines

  • Familiarity with OpenTelemetry or comparable observability instrumentation in Go services

  • Familiarity with LLM provider gateway/routing services (OpenRouter or comparable) — model aliasing, rate‑limit and timeout handling, and budget enforcement

  • Experience with data export/reporting tools (Excel generation, CSV pipelines, or BI‑style dashboards)

Our Story

Our founders, Spence and John met at Google working on Google Translate. As researchers at Stanford and Berkeley, they both worked on language technology to make information accessible to everyone. While together at Google, they were amazed to learn that Google Translate wasn’t used for enterprise products and services inside the company.The quality just wasn’t there. So they set out to build something better. LILT was born.

LILT has been a machine learning company since its founding in 2015. At the time, machine translation didn’t meet the quality standard for enterprise translations, so LILT assembled a cutting‑edge research team tasked with closing that gap. While meeting customer demand for translation services, LILT has prioritized investments in Large Language Models, human‑in‑the‑loop systems, and now agentic AI.

With AI innovation accelerating and enterprise demand growing, the next phase of LILT’s journey is just beginning.

Our Tech

What sets our platform apart:

  • Brand‑aware AI that learns your voice, tone, and terminology to ensure every translation is accurate and consistent

  • Agentic AI workflows that automate the entire translation process from content ingestion to quality review to publishing

  • 100+ native integrations with systems like Adobe Experience Manager, Webflow, Salesforce, GitHub, and Google Drive to simplify content translation

  • Human‑in‑the‑loop reviews via our global network of professional linguists, for high‑impact content that requires expert review

LILT in the News
  • Featured in ++The Software Report’s++ Top 100 Software Companies!

  • LILT makes it onto the ++Inc. 5000 List++.

  • LILT’s continues to be an intellectual powerhouse, holding ++numerous patents++ that help power the most efficient and sophisticated AI and language models in the industry.

  • Check out ++all our news on our website++.

Information collected and processed as part of your application process, including any job applications you choose to submit, is subject to LILT’s Privacy Policy at https://lilt.com/legal/privacy.

At LILT, we are committed to a fair, inclusive, and transparent hiring process. As part of our recruitment efforts, we may use artificial intelligence (AI) and automated tools to assist in the evaluation of applications, including résumé screening, assessment scoring, and interview analysis. These tools are designed to support human decision-making and help us identify qualified candidates efficiently and objectively. All final hiring decisions are made by people. If you have any concerns, require accommodations, or would like to opt‑out of the use of AI in our hiring process, please let us know at recruiting@lilt.com.

LILT is an equal opportunity employer. We extend equal opportunity to all individuals without regard to an individual’s race, religion, color, national origin, ancestry, sex, sexual orientation, gender identity, age, physical or mental disability, medical condition, genetic characteristics, veteran or marital status, pregnancy, or any other classification protected by applicable local, state or federal laws. We are committed to the principles of fair employment and the elimination of all discriminatory practices.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff Fullstack Engineer - Internal Tools
Staff Fullstack Engineer - Internal Tools

Lilt---Lega-Italiana-Per-La-Lotta-Contro-I-Tumori-1 • United States

On-site
USD 180,000 - 240,000
Technical Project Manager
Technical Project Manager

LILT AI • United States

On-site
USD 110,000 - 180,000
Machine Learning Engineer (Real-Time Speech Translation)
Machine Learning Engineer (Real-Time Speech Translation)

lilt-corporate • Washington (IN)

On-site
USD 140,000 - 210,000
Machine Learning Engineer (Real-Time Speech Translation)
Machine Learning Engineer (Real-Time Speech Translation)

LILT • Washington

On-site
USD 140,000 - 200,000
Enterprise Account Executive
Enterprise Account Executive

LILT • Boston (MA)

Remote
USD 120,000 - 210,000
401(k) matching
Flexible time off
Equity
Enterprise Account Executive
Enterprise Account Executive

Lilt---Lega-Italiana-Per-La-Lotta-Contro-I-Tumori-1 • Washington

Remote
USD 120,000 - 240,000
Market salary with OTE
Meaningful equity
401(k) matching
+3
Project Manager, Applied AI
Project Manager, Applied AI

lilt-corporate • Concord (NH)

On-site
USD 90,000 - 125,000
Senior Full Stack Engineer
Senior Full Stack Engineer

Lilt---Lega-Italiana-Per-La-Lotta-Contro-I-Tumori-1 • Boston (MA)

On-site
USD 180,000 - 240,000
Technical Project Manager, Applied AI (Contract)
Technical Project Manager, Applied AI (Contract)

Remote Worker LTD. • United States

Remote
USD 40,000 - 75,000
Enterprise Account Executive
Enterprise Account Executive

Lilt---Lega-Italiana-Per-La-Lotta-Contro-I-Tumori-1 • San Francisco (CA)

Remote
USD 150,000 - 250,000
Market salary with OTE
Equity
401(k) matching
+7