Senior Data Scientist I

Elsevier

Amsterdam

On-site

EUR 54,000 - 90,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Elsevier is seeking a Senior Data Scientist to join the Data Science Life Sciences team. You will develop and deploy GenAI, NLP and ML solutions, building production-ready Python code and collaborating with biology and chemistry experts to validate output.

Responsibilities include data collection, model development, multilingual data processing, RAG pipeline optimization, testing and deployment. You will mentor juniors and lead small projects while staying current with AI advances.

Qualifications

  • Master’s or Ph.D. in Computer Science, Data Science, AI, or related field.
  • 5+ years of applied experience in data science focusing on Generative AI, NLP, and ML.
  • Proficiency in Python for data analysis, model development and deployment.
  • Strong experience with transformer models.
  • Experience with Generative AI technologies, LLMs via API, evaluation tools, and prompt engineering.
  • Knowledge of RAG pipelines and practical implementation.
  • Experience building Agentic RAG systems.
  • Experience with LangChain or similar tools.
  • Familiarity with cloud platforms for production pipelines (e.g., AWS, Azure).
  • Proficiency with data visualization and version control tools (GitHub/GitLab).

Responsibilities

  • Data collection, analysis, model development, and quality assessment for stakeholders.
  • Create production-ready Python packages for data pipelines and model inference with software engineers.
  • Optimize and customize RAG pipelines for multilingual content ingestion and retrieval.
  • Ingest, preprocess, and transform large multilingual datasets for high-quality inputs.
  • Build AI agentic models integrated with RAG pipelines.
  • Rigorous testing and evaluation to ensure performance and reliability.
  • Integrate data science components and perform end-to-end quality checks.
  • Maintain robustness against model drift and ensure output quality.
  • Establish reporting for pipeline performance and automate retraining strategies.
  • Collaborate with cross-functional teams to integrate AI into products and services.
  • Lead and manage projects, mentor junior data scientists, and share knowledge.

Skills

GenAI
NLP
ML
Python
Data analysis
Multilingual NLP

Education

Master's or Ph.D. in Computer Science/Data Science/AI

Tools

Python
OpenSearch
Databricks
LangChain or similar

Job description

Are you interested in working with data and analytics to solve problems? Are you interested in bringing your GenAI, ML and NLP expertise to projects? About our Team Data Science Life Sciences is a diverse team focusing on GenAI, ML, NLP. We mainly develop best-in-class enrichment pipelines for Elsevier's life science .com products such as Reaxys, Embase and Pharmapendium.

About the Role As a Senior Data Scientist, you will play a pivotal role in the development and deployment of cutting-edge Gen AI models and solutions. You will be responsible for building, testing, and maintaining our Gen AI, RAG and NLP solutions You will work throughout the whole life cycle of data science projects: design, implementation, production and beyond. You will deliver efficient and production-ready Python code. You will collaborate closely with developers to deploy and productionize our data science pipelines and with subject matter experts in biology and chemistry domains to validate the output. This role requires a strong foundation in Natural Language Processing (NLP), Machine Learning, Transformer models and Generative AI, as well as proficiency in Python.

Responsibilities
  • Data collection, data analysis, model development, defining quality metrics, quality assessment of models and regular presentations to stakeholders.
  • Creating production-ready Python packages for each component of data science pipelines (such as pre-processing and model inference) and their deployment together with software engineering team
  • Optimizing and customizing Retrieval Augmented Generation (RAG) pipelines to meet specific project requirements that involve content ingestion, machine translation, and contextualized information retrieval
  • Ingesting, preprocessing, and transforming large-scale multilingual data to ensure high-quality inputs for downstream models.
  • Building AI agentic models integrated with RAG pipelines.
  • Conducting rigorous testing and evaluation of AI models to ensure high performance and reliability.
  • Integrating data science components and performing end-to-end quality assessments.
  • Maintaining robustness of data science pipelines against model drift and ensuring consistent output quality.
  • Establishing reporting processes for pipeline performance and developing automated re-training strategies for existing pipelines.
  • Collaborating with cross-functional teams to integrate AI solutions into existing products and services.
  • Leading and managing projects with a team of data scientists and independently executing the entire small-scale projects
  • Mentoring junior data scientists and fostering a knowledge-sharing culture within the team.
  • Staying up-to-date with the latest advancements in AI, machine learning, and NLP technologies.
Requirements
  • Master’s or Ph.D. in Computer Science, Data Science, Artificial Intelligence, or a related field.
  • 5+ years of relevant applied experience in data science, with a focus on Generative AI, NLP, and machine learning.
  • Proficiency in Python for data analysis, model development, and deployment.
  • Strong experience with transformer models
  • Proficiency in Generative AI technologies, including utilizing LLMs via API access, LLM evaluation tools, and prompt engineering.
  • Knowledge of various RAG pipelines and their practical implementation.
  • Experience building Agentic RAG systems is strong requirement.
  • Experience with AI agent management frameworks such as LangChain, or similar tools.
  • Experience with advanced algorithms in deep learning, neural networks, reinforcement learning, and transfer learning.
  • Familiarity with traditional machine learning algorithms such as random forests, SVM, logistic regression, and Bayesian modelling for model building, validation, and testing.
  • Familiarity with cloud platforms (e.g., Bedrock, AWS, Azure) for model deployment and the creation of production-ready pipelines.
  • Proficiency in data visualization tools and techniques.
  • Experience with version control systems (e.g., GitLab or GitHub), Jira, and working in an Agile environment.
  • Proficient in using OpenSearch and Databricks.
  • Excellent problem-solving and analytical skills, with strong attention to detail.
  • Strong communication skills and the ability to work effectively in a team-oriented environment.
Work in a way that works for you

We promote a healthy work/life balance across the organization. We offer an appealing working prospect for our people. With numerous wellbeing initiatives, shared parental leave, study assistance and sabbaticals, we will help you meet your immediate responsibilities and your long-term goals. Flexible working hours - flexing the times when you work in the day to help you fit everything in and work when you are the most productive.

About the business

Elsevier is a global leader in advanced information and decision support for science and healthcare. We believe that by working together with the communities we serve, we can shape human progress to go further, happen faster, and benefit all. We support continuous discovery and uphold the highest standards of content integrity, reliability, and reproducibility so the communities we serve can advance their field of science, healthcare or innovation with confidence. By combining high-quality content with powerful analytics, we transform complexity into clarity and deliver mission-critical insights that help professionals make better decisions when it matters most. We deliver insights that help research institutions, governments, and funders achieve their goals. We help researchers discover and share knowledge, collaborate, and accelerate innovation. We help librarians provide verified, quality information to universities. We help innovators turn knowledge into new products. We help health professionals improve patient care and educators train the next generation of doctors and nurses. Connecting quality content and innovative technologies, we make progress go further and happen faster. And by championing inclusion and sustainability, we ensure progress benefits all. With 9,500 employees, over 2,300 technologists in 5 major tech hubs, and more than 60 locations across the globe, we are committed to supporting the scientific and healthcare communities around the world. We offer a diverse range of opportunities across technology, commercial, business, and early career jobs. If you are looking for a career that inspires progress in science, innovation and health, and allows you to grow every day, find your team at Elsevier. Elsevier is part of RELX Group. Let’s shape progress together. Join us. elsevier.com/about/careers

Primary Location

Base Pay Range: NLD Amsterdam (Radarweg) €53,800 - €89,900. This role is covered by the Collective Labor Agreement Publishing Industry. We know your well-being and happiness are key to a long and successful career. We are delighted to offer country specific benefits.

Accommodations

If you have a disability or other need that requires accommodation or adjustment, please let us know by completing our Applicant Request Support Form or please contact 1-855-833-5120.

We are an equal opportunity employer: qualified applicants are considered for and treated during employment without regard to race, color, creed, religion, sex, national origin, citizenship status, disability status, protected veteran status, age, marital status, sexual orientation, gender identity, genetic information, or any other characteristic protected by law. USA Job Seekers: EEO Know Your Rights.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Data Scientist II
Senior Data Scientist II

RELX INC • Amsterdam

On-site
EUR 59,000 - 99,000
Senior Data Scientist I
Senior Data Scientist I

RX Brasil • Amsterdam

Hybrid
EUR 54,000 - 90,000
Content Marketing Manager
Content Marketing Manager

Elsevier • Amsterdam

On-site
GBP 42,000 - 65,000
Health benefits
Wellbeing platform
Parental leave
Senior Product Manager
Senior Product Manager

LexisNexis Risk Solutions • Amsterdam

On-site
EUR 79,000 - 132,000
Wellbeing initiatives
Shared parental leave
Study assistance
+1
Senior Product Manager
Senior Product Manager

Elsevier • Amsterdam

On-site
EUR 79,000 - 132,000
Senior Product Manager
Senior Product Manager

RELX • Amsterdam

On-site
EUR 79,000 - 132,000
Wellbeing programs
Study assistance
Sabbaticals
Principal Product Mgr I
Principal Product Mgr I

LexisNexis Risk Solutions • Amsterdam

Hybrid
EUR 111,000 - 184,000
Generous holiday allowance
Private medical benefits
Wellbeing programs
Director Software Engineering
Director Software Engineering

Elsevier • Amsterdam

On-site
EUR 111,000 - 184,000
Principal Data Scientist I
Principal Data Scientist I

RX Brasil • Amsterdam

On-site
EUR 87,000 - 145,000
Director Software Engineering
Director Software Engineering

RELX • Amsterdam

On-site
EUR 111,000 - 184,000