Mid-Senior Data Scientist (Life Sciences)

Capgemini Engineering

Lisboa

Híbrido

EUR 55 000 - 75 000

Tempo integral

14 dias+

Recebe mais respostas dos empregadores

Envia um currículo específico para a oferta em poucos minutos.

Vantagens oferecidas por esta oferta de emprego

Hybrid work
Health and life insurance
Training and certifications programs
Career growth programs
Referral program with bonuses

Resumo da oferta

Capgemini Engineering in Portugal is seeking a Mid‑Senior Data Scientist with a biomedical focus to design and govern data foundations for life sciences projects. You will bridge science, engineering, and product teams to deliver scalable, FAIR data solutions and robust ML-enabled insights.

You'll build schema‑driven pipelines in Python/SQL, drive knowledge graph development with RDF/OWL/SPARQL, and apply Gen AI tools to extract value from complex biomedical data, while upholding data governance

Qualificações

  • MSc or PhD in Bioinformatics, Biomedical Engineering, Molecular/Cell Biology, Neuroscience, Genetics, or related field.
  • Excellent stakeholder management — able to drive alignment across scientific, engineering, and product stakeholders.
  • Data modelling and harmonization experience across complex biomedical domains including source-to-canonical mapping, ontology alignment, persistent identifiers, and provenance.
  • Strong Python and SQL skills; comfortable building data pipelines and performing exploratory data analysis.
  • Applied data science / ML experience relevant to knowledge graph or scientific data work.
  • Strong communication skills with both technical and business stakeholders.
  • Critical thinking, intellectual curiosity, and impact‑driven mindset.
  • Ability to adapt and manage priorities in fast‑paced environments.
  • Fluent in Portuguese and English.

Responsabilidades

  • Designing and governing biomedical data models: leading source-to-canonical mapping, ontology alignment, and schema governance including versioning, changelogs, and downstream impact assessments to ensure data integrity and scientific accuracy across complex biomedical domains
  • Building and maintaining data pipelines: developing robust, schema-driven pipelines in Python and SQL, performing exploratory data analysis, and implementing validation frameworks that support high-quality, reproducible scientific workflows
  • Driving knowledge graph development: applying hands‑on experience with RDF, OWL, SPARQL, and property graph modelling tools such as Neo4j and GraphDB to build and enrich knowledge graphs that connect biomedical entities across diverse data sources
  • Applying machine learning and Gen AI: leveraging applied ML experience and familiarity with Gen AI tools (including text generation APIs, chatbots, and enterprise search solutions) to extract insight and value from scientific data at scale
  • Championing FAIR data principles: designing and delivering FAIR data products, leading harmonisation efforts across multiple source systems, and ensuring persistent identifiers and provenance are embedded into every data product
  • Aligning stakeholders across disciplines: driving alignment between scientific, engineering, and product teams through clear communication, structured documentation, and a solutions‑focused mindset that keeps complex projects moving forward
  • Working with biomedical ontologies and controlled vocabularies: applying deep knowledge of resources such as Ensembl, UniProt, and Gene Ontology, including judgement on when and how to extend or map them to real‑world data challenges

Conhecimentos

Python
SQL
Data modelling
Knowledge graphs
Gen AI
Stakeholder management
Communication

Formação académica

MSc/PhD in Bioinformatics or related field

Ferramentas

Neo4j
GraphDB
RDF
SPARQL
LinkML

Descrição da oferta de emprego

At Capgemini Engineering, the world leader in engineering services, we bring together a global team of engineers, scientists, and architects to help the world’s most innovative companies unleash their potential. From autonomous cars to life-saving robots, our digital and software technology experts think outside the box as they provide unique R&D and engineering services across all industries. Join us for a career full of opportunities. Where you can make a difference. Where no two days are the same.

YOUR ROLE

We are looking for a Mid-Senior Data Scientist with a strong background in biomedical sciences and data engineering to join our growing team in Portugal. In this role, you will work at the heart of cutting-edge life sciences projects, helping to design, build, and govern the data foundations that power scientific discovery and product development. You will act as a bridge between scientific, engineering, and product stakeholders - translating complex biological knowledge into robust, scalable, and FAIR data solutions. In this role you will play a key role in:

  • Designing and governing biomedical data models: leading source-to-canonical mapping, ontology alignment, and schema governance including versioning, changelogs, and downstream impact assessments to ensure data integrity and scientific accuracy across complex biomedical domains
  • Building and maintaining data pipelines: developing robust, schema-driven pipelines in Python and SQL, performing exploratory data analysis, and implementing validation frameworks that support high-quality, reproducible scientific workflows
  • Driving knowledge graph development: applying hands‑on experience with RDF, OWL, SPARQL, and property graph modelling tools such as Neo4j and GraphDB to build and enrich knowledge graphs that connect biomedical entities across diverse data sources
  • Applying machine learning and Gen AI: leveraging applied ML experience and familiarity with Gen AI tools (including text generation APIs, chatbots, and enterprise search solutions) to extract insight and value from scientific data at scale
  • Championing FAIR data principles: designing and delivering FAIR data products, leading harmonisation efforts across multiple source systems, and ensuring persistent identifiers and provenance are embedded into every data product
  • Aligning stakeholders across disciplines: driving alignment between scientific, engineering, and product teams through clear communication, structured documentation, and a solutions‑focused mindset that keeps complex projects moving forward
  • Working with biomedical ontologies and controlled vocabularies: applying deep knowledge of resources such as Ensembl, UniProt, and Gene Ontology, including judgement on when and how to extend or map them to real‑world data challenges
YOUR PROFILE
  • MSc or PhD in Bioinformatics, Biomedical Engineering, Molecular/Cell Biology, Neuroscience, Genetics, or related field
  • Excellent stakeholder management — able to drive alignment across scientific, engineering, and product stakeholders
  • Data modelling and harmonization experience across complex biomedical domains including source-to-canonical mapping, ontology alignment, persistent identifiers, and provenance.
  • Strong Python and SQL skills; comfortable building data pipelines and performing exploratory data analysis
  • Applied data science / ML experience relevant to knowledge graph or scientific data work
  • Strong communication skills with both technical and business stakeholders
  • Critical thinking, intellectual curiosity, and impact‑driven mindset
  • Ability to adapt and manage priorities in fast‑paced environments
  • Fluent in Portuguese and English
  • Nice‑to‑have:
    • Prior experience in a pharmaceutical or biotech organization
    • Experience with data catalogue, metadata registry, or schema registry tooling
    • Data engineering fundamentals: pipeline architecture, schema‑driven automation, validation frameworks
    • Hands‑on experience with ML frameworks and model lifecycle (build, deploy, monitor)
    • Hands‑on experience with Gen AI models and tools, such as text generation APIs, chatbots, and enterprise search solutions.
    • Track record of leading schema governance: versioning, changelogs, tagged releases, downstream impact assessment
    • Experience designing FAIR data products and leading data harmonisation efforts across multiple source systems
    • Knowledge graph experience: RDF, OWL, SPARQL, and property graph modelling (Neo4j/GraphDB)
    • Experience with LinkML or equivalent schema modelling frameworks (classes, slots, ranges, constraints, cardinality, ontology bindings)
    • Strong command of biomedical ontologies and controlled vocabularies (e.g. Ensembl, UniProt, Gene Ontology), including judgement on when/how to extend or map them
What You'll Love About Working Here
  • Join a multicultural and inclusive team environment.
  • Enjoy a supportive atmosphere promoting work‑life balance.
  • Engage in exciting national and international projects.
  • Hybrid work.
  • Your career growth is central to our mission. Our array of career growth programs and diverse professionals are crafted to support you in exploring a world of opportunities.
  • Training and certifications programs.
  • Health and life insurance.
  • Referral program with bonuses for talent recommendations.
  • Great office locations.
About Capgemini

Capgemini is an AI‑powered global business and technology transformation partner, delivering tangible business value. We imagine the future of organizations and make it real with AI, technology, and people. With our strong heritage of nearly 60 years, we are a responsible and diverse group of 420,000 team members in more than 50 countries. We deliver end‑to‑end services and solutions with our deep industry expertise and strong partner ecosystem, leveraging our capabilities across strategy, technology, design, engineering and business operations.

Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Mid-Senior Data Scientist (Life Sciences)
Mid-Senior Data Scientist (Life Sciences)

Capgemini Engineering • Fundão

Híbrido
EUR 55 000 - 85 000
Hybrid work
Health and life insurance
Training and certifications programs
+2
Mid-Senior Data Scientist (Life Sciences)
Mid-Senior Data Scientist (Life Sciences)

Capgemini Engineering • Porto

Híbrido
EUR 55 000 - 85 000
Hybrid work
Health and life insurance
Training and certifications programs
+2
Mid-Senior Data Scientist (Life Sciences)
Mid-Senior Data Scientist (Life Sciences)

Capgemini • Lisboa

Híbrido
EUR 75 000 - 95 000
Health insurance
Hybrid work model
Training & certifications
Hybrid Senior Data Scientist, Life Sciences & Biomedical Data
Hybrid Senior Data Scientist, Life Sciences & Biomedical Data

Capgemini • Lisboa

Híbrido
EUR 75 000 - 95 000
Health insurance
Hybrid work model
Training & certifications
Senior Data Engineer (Databricks Focus)
Senior Data Engineer (Databricks Focus)

Capgemini • Lisboa

Híbrido
EUR 45 000 - 65 000
Health insurance
Life insurance
Referral bonuses
Senior AI Technical Lead
Senior AI Technical Lead

Capgemini • Lisboa

Presencial
EUR 70 000 - 100 000
Flexible work environment
Health and Life insurance
Career Acceleration Programs
+1
Business Intelligence Manager
Business Intelligence Manager

Capgemini • Lisboa

Híbrido
EUR 70 000 - 95 000
Data Strategy and Governance Lead
Data Strategy and Governance Lead

Capgemini • Lisboa

Híbrido
EUR 40 000 - 60 000
Health and Life insurance
Career Acceleration Programs
Referral bonuses
+1
Biomedical Data Scientist: FAIR Data & Graphs (Hybrid)
Biomedical Data Scientist: FAIR Data & Graphs (Hybrid)

Capgemini Engineering • Fundão

Híbrido
EUR 55 000 - 85 000
Hybrid work
Health and life insurance
Training and certifications programs
+2
Senior SAS Data Engineer
Senior SAS Data Engineer

Capgemini • Lisboa

Presencial
EUR 30 000 - 40 000
Health and Life insurance
Flexible and dynamic work environment
Career Acceleration Programs
+2