Data Scientist, Lab & Protein Data

Praxy

Lausanne

Sur place

CHF 120 000 - 190 000

Plein temps

Il y a 3 jours
Soyez parmi les premiers à postuler
Générateur de candidature

N’envoyez pas de CV générique — générez un CV et une lettre de motivation adaptés à ce poste précis.

Passez les filtres ATS

Résumé du poste

Adaptyv is building an automated lab where AI agents design and test biology experiments, turning messy outputs into clean, usable data for models and scientists. You will craft the data layer that ensures quality, traceability, and interoperability across sequencing, structure, and design data.

Join a fast-growing biotech team shipping scalable data pipelines and analytics tools to enable AI-driven wet-lab discovery and benchmarking of protein designs.

Qualifications

  • Strong data science / bioinformatics background.
  • Fluent in Python and the scientific stack.
  • Experience with real-world experimental data.
  • Ability to ship end-to-end pipelines.

Responsabilités

  • Own the scientific logic of data quality across the foundry: define good data for each assay type and automate checks.
  • Build anomaly detection and QC models to catch subtle data issues.
  • Specify, review, and improve automated data pipelines with engineers and ML teams.
  • Connect experimental results back to protein design data for integrated insights.
  • Turn foundry outputs into structured, benchmark-grade datasets for customers and model training.
  • Apply statistical rigor to multi-condition data across thousands of samples.

Connaissances

Python
Bioinformatics
Data analysis
Statistics
Anomaly detection

Outils

Pandas
NumPy
SQL

Description du poste

Adaptyv is building an automated lab that lets AI agents run biology experiments.

We're entering the era of agentic science where AI models can now design novel proteins, propose hypotheses, and iterate on experimental results. But they can't run the experiments themselves - that's still a manual, months-long process. We're building the infrastructure that gives AI agents access to the physical world.

We are one of the fastest growing biotech companies, trusted by leading biopharmas, frontier AI labs, and the techbio companies pushing the field forward. This is a rare chance to help advance some of the most important work happening in biotech today.

Our automated lab is powered by a deep software + hardware stack: lab instruments worth millions of USD reverse-engineered into API-controllable hardware, dozens of devices orchestrated through complex workflows, full observability on everything that happens in the lab, processing pipelines for messy physical-world data, and AI systems that troubleshoot production results and accelerate assay development.

We’re growing rapidly and are hiring for talented people to scale and support the massive demand for AI-driven wet lab experimentation.

ABOUT THE ROLE

You'll build out the data science layer of Adaptyv's foundry — the work that turns tens of thousands of raw, messy experimental readouts into clean, trustworthy, structured data that our customers, our models, and our own scientists can rely on. Binding (BLI/SPR), developability, biophysical, and functional assays all produce data at scale; your job is to make that data correct, comparable, and useful.

This sits at the intersection of three things: data quality (is this number real, or an artifact?), bioinformatics (linking experimental results back to sequence, structure, and protein design), and dataset building (turning foundry output into the kind of high-quality, benchmarkable data that frontier AI labs actually want). You'll work shoulder-to-shoulder with the lab scientists who run the assays, the software team who own the pipelines, and the customers who train models on what we produce. This is a hands‑on build role, not a management one.

WHAT YOU'LL DO
  • Own the scientific logic of data quality across the foundry: define what "good data" looks like for each assay type — expected signal ranges, control thresholds, failure modes, edge cases — and turn it into automated checks.
  • Build anomaly detection and QC models that catch bad data the eye would miss: assay drift, instrument variability, plate effects, false passes and false fails — and distinguish real signal from noise statistically.
  • Work with the software and ML teams to specify, review, and improve the automated data pipelines that process instrument outputs, feeding back precise requirements for what to flag, auto-reject, or route to human review.
  • Connect experimental results back to the protein side — sequence, structure, and design — so wet‑lab data and computational models reinforce each other.
  • Turn foundry output into structured, documented, benchmark-grade datasets that are a genuine asset for our customers and for training and evaluating protein‑design models.
  • Apply real statistical rigor to multi-condition data at scale — thousands of samples across hundreds of simultaneous experiments — and make the results interpretable and comparable across runs.
WHAT WE'RE LOOKING FOR
  • Strong data science / bioinformatics background — you're fluent in Python (pandas, numpy, the scientific stack) and comfortable owning messy, real‑world experimental data end to end.
  • Genuine biology grounding — you understand proteins, assays, and sequence/structure/function well enough to know what the data means, not just how to process it. You don't need to be a bench scientist, but you can't be biology‑blind.
  • Statistical maturity — process control, anomaly detection, handling variability and batch effects; you can tell drift from noise and defend the call.
  • Prolific builder with the receipts to prove it. You've shipped a lot — pipelines, tools, models, datasets — and can point to concrete things you built end to end and put into real use, not prototypes that died in a notebook. You move fast, systematize what works, and have no patience for babysitting a fixed dashboard.
  • AI‑native builder. It's 2026 — you build with coding agents like Claude Code as a default, and you have sharp judgment about what they produce.
  • Interdisciplinary by instinct. You're energized working across the lab bench, software, and ML, and you treat automation and data infrastructure as part of your job.
  • Bonus: experience with protein/sequence‑structure data (bioinformatics tooling, structural data), ML on experimental data, or building datasets for model training and benchmarking.
DETAILS
  • Type: Full time

We are reviewing applicants on a rolling basis.

Obtenez votre examen gratuit et confidentiel de votre CV.

ou faites glisser et déposez votre fichier ici.

Similar jobs

Postes similaires à comparer

Data Scientist, Lab & Protein Data
Data Scientist, Lab & Protein Data

Adaptyv • Lausanne

Sur place
CHF 80 000 - 110 000
Data Scientist, Lab & Protein Data
Data Scientist, Lab & Protein Data

Embodied AI • Lausanne

Sur place
CHF 120 000 - 170 000
Software Engineer, Backend
Software Engineer, Backend

Adaptyv • Lausanne

Sur place
CHF 120 000 - 180 000
Bioengineer / Molecular Biologist
Bioengineer / Molecular Biologist

Embodied AI • Lausanne

Sur place
CHF 100 000 - 160 000
Forward Deployed Scientist
Forward Deployed Scientist

Embodied AI • Lausanne

Sur place
CHF 90 000 - 140 000
Research Associate (Molecular Biology / Protein engineering)
Research Associate (Molecular Biology / Protein engineering)

Adaptyv • Lausanne

Sur place
CHF 70 000 - 90 000
Head of Software Engineering
Head of Software Engineering

Adaptyv • Lausanne

Sur place
CHF 120 000 - 160 000
Bioengineer / Molecular Biologist
Bioengineer / Molecular Biologist

Adaptyv • Lausanne

Sur place
CHF 70 000 - 90 000
Scientific Account Executive
Scientific Account Executive

Embodied AI • Suisse

Sur place
CHF 120 000 - 180 000
Head of Software Engineering
Head of Software Engineering

Embodied AI • Lausanne

Sur place
CHF 180 000 - 230 000