Data Scientist III

relx

Mexico

On-site

PHP 6,270,000 - 11,285,000

Full time

7 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Study assistance
Sabbaticals
Parental leave

Job summary

RELX is seeking a Data Scientist to design and deploy ML, NLP, and generative AI solutions across scientific discovery and knowledge extraction. You will handle large-scale, diverse data sources, including publications and knowledge graphs, to enable researchers and clinicians to access actionable insights.

Responsibilities include building production-grade models, developing semantic search and content classification, and collaborating with engineering, product, UX, analytics, and domain

Qualifications

  • Experience in data science, machine learning, AI, NLP, statistics or related quantitative field.
  • Experience with frontier LLMs (GPTs, Claude, Gemini) and fine-tuning.

Responsibilities

  • Design and build ML, NLP, and generative AI systems for scientific discovery and knowledge extraction.
  • Work with large-scale, heterogeneous data including publications, datasets, and ontologies.
  • Apply suitable methods (classification, regression, clustering, ranking, embeddings, LLMs).
  • Develop semantic search, information retrieval, content classification, and summarization capabilities.
  • Build production-ready models, write clean Python code, and contribute data pipelines.

Skills

Python
NLP
Machine Learning
LLMs
Data handling
Pandas
PyTorch
TensorFlow

Tools

NumPy
SciPy
Scikit-learn
Matplotlib

Job description

Data Scientist, USA-Philadephila or Mexico-Mexico City
About our Team

Our global team support products education electronic health records that introduce students to digital charting and prepare them to document care in today's modern clinical environment. We have a very stable product that we've worked to get to and strive to maintain. Our team values trust, respect, collaboration, agility, and quality.

About the Role

In this role, you will design and build machine learning, NLP, and generative AI solutions that support scientific discovery, knowledge extraction, decision support, and intelligent content understanding. You will work with large-scale scientific content and data, applying the right techniques to solve complex problems and deliver reliable, production-ready systems. Working closely with cross-functional partners, you will help turn ambiguous challenges into measurable outcomes that improve how researchers discover and use knowledge.

Responsibilities
  • Design and build machine learning, NLP, and generative AI systems for scientific discovery, knowledge extraction, decision support, and intelligent content understanding.
  • Work with large-scale, complex, and heterogeneous data, including scientific publications, research datasets, knowledge graphs, ontologies, taxonomies, citations, metadata, and content from every scientific discipline.
  • Apply the right technique to each problem, using approaches such as classification, regression, clustering, ranking, feature engineering, deep learning, embeddings, LLMs, retrieval, and generative AI.
  • Develop capabilities for semantic search, information retrieval, entity extraction, content classification, recommendation, ranking, summarization, question answering, and evidence-grounded generation.
  • Build, evaluate, fine-tune, prompt, and integrate models into robust production systems, while continuously improving quality, relevance, reliability, and user value.
  • Write clean, tested, production-quality Python and contribute reusable data science components, packages, and scalable data pipelines for preprocessing, inference, experimentation, monitoring, and continuous improvement.
  • Support deployment, monitoring, model maintenance, drift detection, automated retraining, and ongoing optimization of data science systems.
  • Collaborate with engineering, product, UX, analytics, research, and domain experts, and communicate technical concepts, model behavior, insights, trade-offs, and recommendations clearly to technical and non-technical audiences.
Requirements
  • Experience in data science, machine learning, artificial intelligence, NLP, statistics, applied mathematics, computer science, or a related quantitative area.
  • Experience working with frontier LLMs such as OpenAI's GPTs, Anthropic's Claude, and Google's Gemini, including fine-tuning LLMs and/or SLMs.
  • Strong Python skills and a habit of writing clean, maintainable, well-tested code.
  • A solid grasp of machine learning fundamentals, including supervised and unsupervised learning, feature engineering, model evaluation, model selection, and performance measurement.
  • Experience working with structured, semi-structured, or unstructured data, especially large-scale text or content datasets.
  • Familiarity with common data science and machine learning tools such as Pandas, NumPy, SciPy, Scikit-learn, PyTorch, TensorFlow, or Matplotlib.
  • The ability to translate complex and ambiguous requirements into practical, measurable, data-driven solutions, with strong analytical thinking, problem-solving skills, and attention to quality.
  • Clear communication skills, a collaborative approach to working with engineering, product, and business stakeholders, and a genuine interest in building production-ready systems that deliver real user value.
Work in a Way That Works for You

We promote a healthy work/life balance across the organisation. We offer an appealing working prospect for our people. With numerous wellbeing initiatives, shared parental leave, study assistance, and sabbaticals, we will help you meet your immediate responsibilities and your long-term goals.

Working Pattern

Working flexible hours - flexing the times when you work in the day to help you fit everything in and work when you are the most productive

About the Business

A global leader in information and analytics, we help researchers and healthcare professionals advance science and improve health outcomes for the benefit of society. Building on our publishing heritage, we combine quality information and vast data sets with analytics to support visionary science and research, health education and interactive learning, as well as exceptional healthcare and clinical practice.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Scientist I
Senior Data Scientist I

relx • Mexico

On-site
PHP 7,235,000 - 12,056,000
Manager Data Science
Manager Data Science

relx • Mexico

On-site
PHP 14,984,000 - 27,829,000
Flexible hours
Wellbeing initiatives
Parental leave
+3
Senior Data Engineer I
Senior Data Engineer I

LexisNexis Risk Solutions • Manila

On-site
PHP 900,000 - 1,300,000
Data Scientist – Junior / Mid / Senior
Data Scientist – Junior / Mid / Senior

Globetec Solutions • Taguig

Hybrid
PHP 700,000 - 1,100,000
Junior Data Scientist
Junior Data Scientist

KAVI Philippines • Philippines

On-site
PHP 600,000 - 900,000
Team-building activities
Employee recognition programs
Career development opportunities
+1
Senior Data Scientist
Senior Data Scientist

London Stock Exchange Group • Philippines

On-site
PHP 1,500,000 - 2,100,000
Data Scientist
Data Scientist

ERNI • Manila

Hybrid
PHP 900,000 - 1,500,000
Baby Basket
Fruit Basket
Free snacks and coffee in the office
+3
Data Engineer Mle
Data Engineer Mle

Insud Pharma • España

On-site
PHP 3,991,000 - 6,531,000
Flexible start time
Permanent contract
Life & accident insurance
+6
Data Analyst - Customer Intelligence
Data Analyst - Customer Intelligence

LexisNexis Risk Solutions • Manila

Hybrid
PHP 600,000 - 1,200,000
Senior Agentic Data Management Advisor
Senior Agentic Data Management Advisor

Lingaro Group Philippines • Manila

Hybrid
PHP 1,800,000 - 2,400,000
Office option
Work from anywhere
Udemy learning
+1