Data Scientist, Agent

Lovable

Stockholms kommun

On-site

SEK 750,000 - 1,100,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Lovable, a Stockholm-based company, seeks a Data Scientist to own AI agent quality metrics and build eval systems. You will design experiments, trace agent behavior, and drive measurable improvements in success, completion, and error rates.

You’ll collaborate closely with the agent engineering team to translate telemetry into fixes and continuously enhance agent performance. The role emphasizes strong SQL and Python skills, applied statistics, and the ability to design noisy A/B tests.

Qualifications

  • Proficient in SQL and Python with strong statistics background.
  • Experience with designing and running A/B tests for agent changes.
  • Ability to build eval systems and observe model behavior.
  • Comfortable turning telemetry into concrete fixes for the agent.

Responsibilities

  • Define and own metrics for agent quality (success, completion, error rates).
  • Build eval systems and experiment framework (A/B rollout) to catch regressions.
  • Turn traces and telemetry into concrete fixes with the agent engineering team.
  • Develop tooling to evaluate the evolving agent continuously.
  • Set the bar for agent behavior where there is no fixed key.

Skills

SQL
Python
Applied statistics
Experimentation
A/B testing
Data analysis
Analytical mindset
Curiosity

Tools

BigQuery
PubSub
OTEL tracing
Braintrust
Hex
Lovable Apps

Job description

TL;DR — You own how we measure and improve Lovable's AI agent. You build the eval systems and experiments that tell us whether a change makes the agent better or worse, and you turn agent telemetry into the fixes that raise success rates and cut errors.

At Lovable, data scientists are not isolated model-builders; they sit close to the product, experimenting continuously with how intelligence changes user behavior and product dynamics.

Why Lovable?

Lovable lets anyone and everyone build software with any language. From solopreneurs to Fortune 100 teams, millions of people use Lovable to transform raw ideas into real products, fast. We are at the forefront of a foundational shift in software creation, which means you have an unprecedented opportunity to change the way the digital world works. Lovable-built applications and websites are visited hundreds of millions of times a month, and our enterprise footprint is compounding fast. And we're just getting started.

We're a small, talent-dense team building a generation-defining company from Stockholm. We value extreme ownership, high velocity, and low-ego collaboration. We seek out people who care deeply, ship fast, and are eager to make a dent in the world.

What we're looking for
  • A data scientist who wants to make an AI agent measurably better, not just report on it. You own agent quality metrics and drive them up.

  • Experience or strong interest in LLM evaluation and observability: building evals, scoring outputs, tracing agent behavior, and catching regressions.

  • Strong SQL and Python, applied statistics, and experimentation. Comfortable designing A/B tests for agent changes where outcomes are noisy.

  • You build systems and agents that produce this insight continuously, rather than one-off analyses.

  • Instinct for what "good" looks like in agent behavior (success, error rates, task completion) and how to measure it when there is no clean answer key.

  • Entrepreneurial. Thrives in ambiguity, and works closely with the engineers building the agent.

What you'll do
  • Define and own the metrics for agent quality: success, completion, error rates, and the behaviors that drive them.

  • Build the eval systems and experiment framework that decide whether an agent change ships, like an A/B-tested rollout that catches a change increasing errors before it reaches everyone.

  • Turn agent traces and telemetry into concrete fixes, working directly with the agent engineering team.

  • Build the tooling and agents that produce these evaluations continuously as the agent evolves.

  • Set the bar for how we judge agent behavior where there is no answer key to check against.

Our tech stack

We're building with tools that both humans and AI love:

  • Languages: SQL and Python

  • LLM evaluation & observability: Braintrust, OTEL tracing, many LLM providers

  • Warehouse & events: BigQuery, PubSub

  • Analytics & product: Hex, Lovable Apps

  • Experimentation: A/B and growth testing

  • Cloud: GCP

And always on the lookout for what's next.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Scientist, Product
Data Scientist, Product

Lovable • Stockholms kommun

On-site
SEK 900,000 - 1,300,000
Product Manager (Agents)
Product Manager (Agents)

Lovable • Stockholms kommun

On-site
SEK 1,100,000 - 1,500,000
Data Scientist
Data Scientist

Lovable Labs Sweden AB • Stockholms kommun

On-site
SEK 485,000 - 702,000
AI Ops Engineer (FBOS)
AI Ops Engineer (FBOS)

Lovable • Stockholms kommun

On-site
SEK 1,100,000 - 1,600,000
Data Scientist
Data Scientist

Lovable • Stockholms kommun

On-site
SEK 600,000 - 800,000
Product Manager (Agents)
Product Manager (Agents)

AI Chopping Block, Inc. • Stockholms kommun

On-site
SEK 762,000 - 981,000
AI Ops Engineer (People Team)
AI Ops Engineer (People Team)

Lovable • Stockholms kommun

On-site
SEK 900,000 - 1,300,000
Data Scientist, Growth
Data Scientist, Growth

Lovable • Stockholms kommun

On-site
SEK 600,000 - 1,000,000
Data Scientist, Pricing
Data Scientist, Pricing

Lovable • Stockholms kommun

On-site
SEK 900,000 - 1,200,000
IT Engineer
IT Engineer

Lovable • Stockholms kommun

On-site
SEK 900,000 - 1,300,000