(Founding) AI Engineer

Rainmaker

Greater London

On-site

GBP 90,000 - 140,000

Full time

7 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Equity
Founding team

Job summary

Rainmaker is seeking its first dedicated AI Engineer to own the intelligence layer of the product, including retrieval over legal news, deals, and firm data. You will shape how we evaluate, what we run in-house, and how models behave reliably in production.

This founding hire reports to the CTO and joins a hands-on multi-disciplinary engineering team with meaningful equity. You will build and refine retrieval, grounding, and agentic features, ensuring precision and traceability while avoiding

Qualifications

  • 5–8 years in software engineering, with at least two spent building LLM-backed systems that real users depend on.
  • Deep retrieval experience: built RAG systems that worked and can explain why first versions failed.
  • Rigour about evaluation with dedicated eval infrastructure.
  • Strong grounding in Python and TypeScript; comfortable with AWS in production environments.
  • Fluency across model landscapes and related tooling to pick reliable options.
  • Comfort with ambiguity and owning a domain independently for a while.
  • Directness: raise issues early and base decisions on data.

Responsibilities

  • Own end-to-end retrieval: chunking, embedding, indexing, hybrid and re-ranked search over legal news and deal data.
  • Build entity resolution across sources so entities mean the same thing everywhere.
  • Ensure grounding and citation are intrinsic, traceable to sources.
  • Define where retrieval ends and structured queries begin; avoid unnecessary model use.
  • Develop LLM-backed features and guided-question loops for the BD Centre; safeguard against fabrications.
  • Design safe agent/tool-use flows and maintain versioned, testable prompt architecture.
  • Build the eval harness and measure quality metrics; monitor hallucinations, latency and cost per interaction.
  • Work with data pipelines to ingest legal content and turn it into structured records; improve precision/recall over time.
  • Ship production code in a TS/Python stack on AWS; own infra costs, latency and reliability.
  • Provide non-engineers with tools to inspect and improve model output without coding.
  • Advise product on feasibility, cost and research scope; set engineering standards for AI features.

Skills

Software engineering
LLM-backed systems
Retrieval / RAG
Evaluation infrastructure
Python
TypeScript
AWS production systems
Model landscape fluency
Directness

Job description

Rainmaker is built to give lawyers business intelligence for BD: legal news and deal data, a BD tool for building and managing approaches to prospective clients, plus a live rankings system that lets lawyers credential their work against their peers. The product launches later this year. It is AI-first in a market that has never had a product like it, and you will be shaping it from the start.

ABOUT THE ROLE

We are looking for our first dedicated AI Engineer. You will own the intelligence layer of the product: retrieval over legal news, deals and firm data, the agentic and guided-question flows inside the BD Centre, the extraction that turns unstructured coverage into structured records, and the evaluation infrastructure that tells us whether any of it is actually working.

This is a founding hire. The AI features that exist today were built alongside everything else. Your job is to take them from working to defensible, then build what comes next. You will decide how we do retrieval, how we evaluate, what we run in-house and what we buy, and you will be the person the rest of the engineering team asks when a feature depends on a model behaving predictably.

You will report to the CTO and work closely with our Head of Product and the wider engineering team. You will have meaningful equity.

We care about output rather than theory. Our users are lawyers, and lawyers do not forgive a confident wrong answer. Precision, traceability and knowing when the system should decline to answer matter more here than they do in most consumer products.

RESPONSIBILITIES
Retrieval and knowledge

- Own retrieval end to end: chunking, embedding, indexing, hybrid and re-ranked search over legal news, deal records, firm and lawyer profiles. - Build entity resolution that holds up across sources, so a firm, a lawyer and a deal mean the same thing wherever they appear. - Make grounding and citation a property of the system rather than a prompt instruction. Every claim the product shows a user should be traceable to a source. - Decide where retrieval ends and structured query begins, and stop us reaching for a model where a database would do.

Agentic and generative features

- Build the LLM-backed features in the product, including the guided-question loops in the BD Centre and the generative surfaces around dossiers, feeds and alerts. - Design agent and tool-use flows that fail safely, degrade to something useful and never invent a client, a deal or a quote. - Own prompt architecture as engineering rather than as text: versioned, tested, reviewable and cheap to change.

Evaluation and quality

- Build the eval harness. Define what good looks like for each AI surface, build the datasets, and make regression visible before a release rather than after it. - Instrument quality in production: hallucination rate, retrieval hit rate, refusal behaviour, latency and cost per interaction. - Run the experiments that decide model choice, and be willing to conclude that a smaller or cheaper model is the right answer.

Data and extraction

- Work with the data pipeline that ingests legal news and deal coverage, and own the extraction that turns it into structured, queryable records. - Improve precision and recall on that extraction over time, and know the current numbers for both. - Feed clean data into the rankings engine, and understand enough of its methodology to spot when the inputs are wrong.

Engineering and platform

- Ship production code into a TypeScript and Python stack running on AWS, and take responsibility for it in production. - Own cost, latency and reliability of the AI layer, including caching, batching, fallbacks and rate limits. - Build the internal tooling that lets non-engineers inspect, correct and improve model output without asking you.

Judgement and influence

- Tell Product what is feasible, what is expensive and what is a research project rather than a sprint. - Push back where a feature is being specified as an AI feature when it should not be one. - Set the standard and the practices that the next AI hires will work to.

WHAT WE ARE LOOKING FOR
  • 5-8 years in software engineering, with at least two spent building LLM-backed systems that real users depend on. Production experience, not prototypes.
  • Deep applied retrieval experience: you have built RAG systems that worked, and you can explain in detail why the first version did not.
  • Rigour about evaluation. You have built eval infrastructure and you treat an unmeasured AI feature as an unfinished one.
  • Strong engineering fundamentals in Python and TypeScript, comfortable in AWS and in production systems rather than notebooks.
  • Fluency across the current model landscape and the tooling around it, with the judgement to pick the boring option when it wins.
  • Comfort with ambiguity and with owning a domain alone. You will not have a team to delegate to for some time.
  • Directness. You raise problems early, you argue your case on evidence, and you change your mind when the data says so.
PREFERRED
  • Experience in legal, financial or another domain where accuracy is a professional obligation rather than a nice-to-have.
  • Experience with entity resolution, knowledge graphs or document-heavy extraction at scale.
  • Fine-tuning, distillation or model serving where it earned its keep against a hosted API.
  • Early-stage or pre-launch experience, and a working understanding of what unit economics mean for an AI product.
  • An open-source record, a paper, a side project, or anything else that shows us what you build when nobody is assigning it.

Rainmaker is an equal opportunity employer. We are committed to diversity and inclusivity.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Graduate AI Engineer
Graduate AI Engineer

Essential Counsel, LLC. • Greater London

On-site
GBP 70,000 - 110,000
Founding AI Engineer — Legal BD Platform (Equity)
Founding AI Engineer — Legal BD Platform (Equity)

Rainmaker • Greater London

On-site
GBP 90,000 - 140,000
Equity
Founding team
AI Engineer - Wexler
AI Engineer - Wexler

Wexler • Greater London

On-site
GBP 70,000 - 90,000
Competitive salary and equity
Budget for learning and growth
Bi-annual team retreats
AI Engineer
AI Engineer

Norton Rose Fulbright • Greater London

On-site
GBP 90,000 - 120,000
AI Engineer
AI Engineer

Norton Rose Fulbright • City of Westminster

Hybrid
GBP 85,000 - 125,000
AI Engineer
AI Engineer

Solve Intelligence • Greater London

On-site
GBP 65,000 - 85,000
Competitive Salary + Significant Equity
Full visa sponsorship
Private medical insurance
+2
AI Engineer - Wexler
AI Engineer - Wexler

Praxis, Inc. • Greater London

On-site
GBP 70,000 - 90,000
Competitive salary & equity
Autonomy & ownership
Learning & development budget
+1
Software Engineer - Wexler
Software Engineer - Wexler

Praxis, Inc. • Greater London

On-site
GBP 90,000 - 130,000
Competitive salary and significant 2x%
Autonomy to design core systems
Budget for learning and professional成长
+2
Senior AI Engineer - Legal Workflows
Senior AI Engineer - Legal Workflows

Harnham • Greater London

Hybrid
GBP 90,000 - 130,000
Hybrid work (2 days in London)
Full Stack Engineer
Full Stack Engineer

Praxis, Inc. • Greater London

On-site
GBP 90,000 - 120,000
Equity
Learning budget
Bi-annual team retreats
+1