Data Scientist III

Elsevier

Philadelphia (Philadelphia County)

Hybrid

MXN 1,618,000 - 2,696,000

Full time

4 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Elsevier in Mexico City or Philadelphia seeks a Data Scientist to design and build ML, NLP, and generative AI solutions that accelerate scientific discovery and knowledge extraction.

You will work with large-scale scientific content, apply appropriate techniques, and deliver production-ready systems while collaborating with cross-functional teams to turn complex problems into measurable impact.

Qualifications

  • Experience in data science, machine learning, AI, NLP, statistics, applied mathematics, computer science, or a related quantitative area.
  • Experience with frontier LLMs such as OpenAI GPTs, Anthropic Claude, and Google Gemini, including fine-tuning LLMs and/or SLMs.
  • Strong Python skills and a habit of writing clean, maintainable, well-tested code.
  • A solid grasp of machine learning fundamentals, including supervised and unsupervised learning, feature engineering, model evaluation, model selection, and performance measurement.
  • Experience with structured, semi-structured, or unstructured data, especially large-scale text or content datasets.
  • Familiarity with common data science and machine learning tools such as Pandas, NumPy, SciPy, Scikit-learn, PyTorch, TensorFlow, or Matplotlib.

Responsibilities

  • Design and build machine learning, NLP, and generative AI systems for scientific discovery, knowledge extraction, decision support, and intelligent content understanding.
  • Work with large-scale, complex data including scientific publications, research datasets, knowledge graphs, ontologies, taxonomies, and metadata.
  • Apply the right technique to each problem, using approaches like classification, regression, clustering, ranking, feature engineering, deep learning, embeddings, LLMs, retrieval, and generative AI.
  • Develop capabilities for semantic search, information retrieval, entity extraction, content classification, recommendation, ranking, summarization, question answering, and evidence-grounded generation.
  • Build, evaluate, fine-tune, prompt, and integrate models into robust production systems, while improving quality, relevance, reliability, and user value.
  • Write clean, tested, production-quality Python and contribute reusable data science components, packages, and scalable data pipelines for preprocessing, inference, experimentation, monitoring, and continuous improvement.
  • Support deployment, monitoring, model maintenance, drift detection, automated retraining, and ongoing optimization of data science systems.
  • Collaborate with engineering, product, UX, analytics, research, and domain experts, and communicate technical concepts clearly to technical and non-technical audiences.

Skills

Python
NLP
Machine learning
Deep learning
Data analysis
Statistics
Data processing
Pandas
NumPy
SciPy
Scikit-learn
PyTorch
TensorFlow
Matplotlib

Tools

Pandas
NumPy
SciPy
Scikit-learn
PyTorch
TensorFlow
Matplotlib

Job description

Data Scientist, USA-Philadelphia or Mexico-Mexico City

Are you excited by the opportunity to use machine learning, NLP, and generative AI to help researchers discover knowledge faster and make better decisions?

Would you enjoy turning complex scientific and business challenges into practical, production-ready AI solutions that create real user value?

About our Team

Our global team support products education electronic health records that introduce students to digital charting and prepare them to document care in today’s modern clinical environment. We have a very stable product that we’ve worked to get to and strive to maintain. Our team values trust, respect, collaboration, agility, and quality.

About the Role

In this role, you will design and build machine learning, NLP, and generative AI solutions that support scientific discovery, knowledge extraction, decision support, and intelligent content understanding. You will work with large-scale scientific content and data, applying the right techniques to solve complex problems and deliver reliable, production-ready systems. Working closely with cross-functional partners, you will help turn ambiguous challenges into measurable outcomes that improve how researchers discover and use knowledge.

Responsibilities
  • Design and build machine learning, NLP, and generative AI systems for scientific discovery, knowledge extraction, decision support, and intelligent content understanding.
  • Work with large-scale, complex, and heterogeneous data, including scientific publications, research datasets, knowledge graphs, ontologies, taxonomies, citations, metadata, and content from every scientific discipline.
  • Apply the right technique to each problem, using approaches such as classification, regression, clustering, ranking, feature engineering, deep learning, embeddings, LLMs, retrieval, and generative AI.
  • Develop capabilities for semantic search, information retrieval, entity extraction, content classification, recommendation, ranking, summarization, question answering, and evidence-grounded generation.
  • Build, evaluate, fine-tune, prompt, and integrate models into robust production systems, while continuously improving quality, relevance, reliability, and user value.
  • Write clean, tested, production-quality Python and contribute reusable data science components, packages, and scalable data pipelines for preprocessing, inference, experimentation, monitoring, and continuous improvement.
  • Support deployment, monitoring, model maintenance, drift detection, automated retraining, and ongoing optimization of data science systems.
  • Collaborate with engineering, product, UX, analytics, research, and domain experts, and communicate technical concepts, model behavior, insights, trade-offs, and recommendations clearly to technical and non-technical audiences.
Requirements
  • Experience in data science, machine learning, artificial intelligence, NLP, statistics, applied mathematics, computer science, or a related quantitative area.
  • Experience working with frontier LLMs such as OpenAI’s GPTs, Anthropic’s Claude, and Google’s Gemini, including fine-tuning LLMs and/or SLMs.
  • Strong Python skills and a habit of writing clean, maintainable, well-tested code.
  • A solid grasp of machine learning fundamentals, including supervised and unsupervised learning, feature engineering, model evaluation, model selection, and performance measurement.
  • Experience working with structured, semi-structured, or unstructured data, especially large-scale text or content datasets.
  • Familiarity with common data science and machine learning tools such as Pandas, NumPy, SciPy, Scikit-learn, PyTorch, TensorFlow, or Matplotlib.
  • The ability to translate complex and ambiguous requirements into practical, measurable, data-driven solutions, with strong analytical thinking, problem-solving skills, and attention to quality.
  • Clear communication skills, a collaborative approach to working with engineering, product, and business stakeholders, and a genuine interest in building production-ready systems that deliver real user value.
Work in a Way That Works for You

We promote a healthy work/life balance across the organisation. We offer an appealing working prospect for our people. With numerous wellbeing initiatives, shared parental leave, study assistance, and sabbaticals, we will help you meet your immediate responsibilities and your long-term goals.

Working Pattern

Working flexible hours - flexing the times when you work in the day to help you fit everything in and work when you are the most productive

About the Business

A global leader in information and analytics, we help researchers and healthcare professionals advance science and improve health outcomes for the benefit of society. Building on our publishing heritage, we combine quality information and vast data sets with analytics to support visionary science and research, health education and interactive learning, as well as exceptional healthcare and clinical practice. At Elsevier, your work contributes to the world's grand challenges and a more sustainable future. We harness innovative technologies to support science and healthcare to partner for a better worl

U.S. National Base Pay Range: $95,300 - $158,800. Geographic differentials may apply in some locations to better reflect local market rates.

We know your well-being and happiness are key to a long and successful career. We are delighted to offer country specific benefits.

Please read our Candidate Privacy Policy.

We are an equal opportunity employer: qualified applicants are considered for and treated during employment without regard to race, color, creed, religion, sex, national origin, citizenship status, disability status, protected veteran status, age, marital status, sexual orientation, gender identity, genetic information, or any other characteristic protected by law.

USA Job Seekers: EEO Know Your Rights.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Scientist I
Senior Data Scientist I

LexisNexis Risk Solutions • Philadelphia

Hybrid
MXN 1,950,000 - 3,249,000
Wellbeing initiatives
Shared parental leave
Study assistance
+1
Senior Data Scientist I
Senior Data Scientist I

Elsevier • Philadelphia

On-site
MXN 1,959,000 - 3,264,000
Data Scientist III
Data Scientist III

LexisNexis Risk Solutions • Philadelphia

Hybrid
MXN 1,610,000 - 2,683,000
Wellbeing initiatives
Parental leave
Study assistance
+2
Data Scientist II
Data Scientist II

RXinsider LTD. • Philadelphia

On-site
USD 72,000 - 119,000
Senior Data Scientist
Senior Data Scientist

Elsevier • Philadelphia

Hybrid
USD 95,300 - 158,800
Annual incentive bonus
Work with a rich collection of scientific data
Senior Data Scientist
Senior Data Scientist

RELX INC • Annapolis (MD)

On-site
USD 100,000 - 167,000
Annual incentive bonus
Country-specific benefits
Principal Software Engineer/Principal AI Engineer
Principal Software Engineer/Principal AI Engineer

LexisNexis Risk Solutions • Philadelphia

Hybrid
USD 115,000 - 192,000
Comprehensive Pension Plan
Generous vacation entitlement
Sabbatical leave
+1
Manager Data Science
Manager Data Science

LexisNexis Risk Solutions • Philadelphia

On-site
USD 115,000 - 193,000
Annual incentive bonus
Senior Data Scientist
Senior Data Scientist

RELX • Philadelphia

On-site
USD 95,000 - 159,000
Senior Data Scientist
Senior Data Scientist

RXinsider LTD. • Philadelphia

On-site
USD 95,000 - 191,000