Senior Data Scientist - LLM Evaluation & Business Intelligence

Paradigm Health, Inc.

Town of Columbus (NY)

Hybrid

USD 120,000 - 180,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Paradigm Health, Inc. is seeking a rigorous Data Scientist to define how we measure and trust LLM-powered clinical data products and to build a reliable business intelligence layer across the org.

You'll work at the intersection of AI evaluation, analytics engineering, and clinical research, partnering with AI engineers, data engineers, clinicians, and business stakeholders to translate questions into well-defined data products.

Qualifications

  • Bachelor's degree or higher in data science or related field.
  • 4+ years as a data scientist or analytics role in life sciences or healthcare.
  • Experience evaluating LLM/NLP systems and performance metrics.
  • Strong SQL and dbt production data modeling experience.
  • Python (preferred) and PySpark proficiency.
  • Experience with Databricks and BI tools like Hex.

Responsibilities

  • Design and run evaluations of production LLM pipelines with accuracy metrics.
  • Evaluate Patient Trial Evaluation (PTE) pipeline for accuracy and failure roots.
  • Improve retrieval-augmented generation (RAG) systems end-to-end.
  • Inform cost/quality tradeoffs across trials, prompts, and models.
  • Create a library of best prompts for clinical concepts with evidence.
  • Collaborate with AI Engineering to bridge evaluation and production.
  • Build and maintain data models using SQL and dbt on Databricks.
  • Develop a semantic layer for consistent, trusted reporting.
  • Advance data governance across dbt, Databricks, and Hex.
  • Translate business needs into well-defined data products with non-technical stakeholders.
  • Support internal and external reporting with well-tested datasets.
  • Mentor and support junior data scientists for senior applicants.

Education

Bachelor's degree or higher in data science, statistics, biostatistics, computer science, mathematics, epidemiology, or a related field

Tools

Databricks
dbt
Hex
PySpark
SQL

Job description

Paradigm Health is rebuilding the clinical research ecosystem by enabling equitable access to trials for all patients. Our platform enhances trial efficiency and reduces the barriers to participation for healthcare providers. Incubated by ARCH Venture Partners and backed by leading healthcare and life sciences investors, Paradigm’s seamless infrastructure implemented at healthcare provider organizations, will bring potentially life-saving therapies to patients faster.

Our team hails from a broad range of disciplines and is committed to the company’s mission to create equitable access to clinical trials for any patient, anywhere. Join us, and bring your expertise, passion, creativity, and drive as we work together to realize this mission.

About the role:

As a Data Scientist on our team, you will help define how we measure and trust the LLM systems that power Paradigm's clinical data products and build the business intelligence layer that turns that data into reliable reporting across the org. You'll work at the intersection of applied AI evaluation, analytics engineering, and clinical research, partnering with AI engineers, data engineers, clinicians, and non-technical stakeholders.

In this role you'll be responsible for delivering reliable data that helps Paradigm measure its business and drive outcomes. You'll partner with teams across the organization to define requirements, then research and build the models and reporting that meet them. This is a deeply collaborative, internally-facing role: your partners range from highly technical teammates to non-technical business, clinical, and product stakeholders, and much of your impact comes from translating between them — turning ambiguous questions into well-defined data products and communicating results, and their caveats, clearly to every audience.

Your primary focus will be evaluating production LLM pipelines: designing accuracy metrics, measuring retrieval effectiveness, and making principled cost/quality tradeoffs for the prompts and models that extract clinical meaning from patient data. Alongside this, you'll help establish a trustworthy business intelligence foundation — production data models, a shared semantic layer, and governed pipelines that give the whole organization consistent, defensible numbers. It's a role for someone who wants to bring scientific rigor to a fast-moving AI system and see their work drive measurable impact for patients and providers.

What you'll do:
  • Design and run evaluations of production LLM pipelines, developing accuracy and quality metrics that tell us how well our systems perform on real clinical tasks.

  • Support evaluation of our Patient Trial Evaluation (PTE) pipeline — assessing prompts and trial-matching logic for accuracy in reporting, and surfacing where and why they fail.

  • Measure and improve the effectiveness of our retrieval-augmented generation (RAG) systems, from retrieval quality through final output.

  • Use accuracy metrics to inform cost/quality tradeoffs across trials, prompts, and models, giving the team a clear basis for which approaches to ship.

  • Establish and maintain a library of "best prompts" for recurring clinical concepts, backed by evidence rather than intuition.

  • Partner with AI Engineering to close the loop between evaluation findings and production improvements.

  • Build and maintain production-grade data models using SQL and dbt on Databricks, ensuring analytics logic is reliable, maintainable, and production-ready.

  • Help establish and grow a semantic layer that enables consistent, trusted reporting across the organization.

  • Advance data governance practices across dbt, Databricks, and Hex — documentation, testing, versioning, and clear ownership of metrics.

  • Collaborate with non-technical internal stakeholders to develop reporting logic, answer data questions, and translate business needs into well-defined data products.

  • Support both internal and external reporting needs with accurate, well-tested datasets.

  • For senior applicants: mentor and support junior data scientists — providing guidance on evaluation and analytics approaches, technical best practices, and professional growth.

Who you are

You are a rigorous, versatile data scientist who is energized by making AI systems measurable and trustworthy. You care about getting the number right and being able to defend it. You're equally comfortable designing an evaluation framework for an LLM pipeline and writing the SQL and dbt models that power reliable reporting.

You bring a strong interest in clinical research and a passion for advancing healthcare through data innovation. You're autonomous and driven — you can take an ambiguous problem and run with it, structuring the approach yourself rather than waiting for a fully specified spec. You're curious, collaborative, and hold yourself to data science best practices as a default, not an afterthought.

What you bring
  • Bachelor's degree or higher in data science, statistics, biostatistics, computer science, mathematics, epidemiology, or a related field.

  • 4+ years of experience as a data scientist, analytics engineer, ML/data professional, or other highly analytical role in life sciences, biotech, healthcare, or a related industry.

  • Solid understanding of how production LLM pipelines work, and hands-on experience evaluating LLM or NLP systems — accuracy metrics, error analysis, RAG/retrieval quality, prompt evaluation, or similar.

  • Strong proficiency in SQL, with experience building production data models in dbt (or a comparable transformation framework).

  • Proficiency in at least one programming language (Python preferred; PySpark a plus).

  • Experience with modern data platforms such as Databricks, and BI/analytics tools such as Hex, or similar.

  • Fluency with data science best practices: reusability, version control, documentation, testing, code review, and reproducibility.

  • Excellent written and verbal communication, with the ability to translate technical concepts into actionable insights for non-technical stakeholders and to develop reporting logic collaboratively with them.

  • Comfort with AI tooling and agentic workflows in day-to-day work.

  • Comfort with ambiguity and adaptability to a mission-driven, fast-paced startup environment.

Bonus
  • Master's degree or higher in a quantitative field.

  • Experience establishing or working within a semantic layer for organization-wide reporting.

  • Experience integrating LLM-based workflows into production systems, and working closely with ML/AI engineering teams.

  • Familiarity with clinical data elements (oncology a plus) and the clinical trial industry or research operations.

  • Strong statistical knowledge, including regression, classification, hypothesis testing, causal inference, or Bayesian methods (e.g., continuous toxicity monitoring for trial protocols).

  • Prior experience mentoring other data scientists or leading cross-functional projects (for senior candidates).

At Paradigm Health, we are committed to providing equal employment opportunities to all qualified individuals. We encourage and welcome candidates from all backgrounds and perspectives to apply for our open positions. We are interested in all qualified individuals and ensure that all employment decisions are based on job-related factors such as skills, experience, and qualifications.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Software Engineer, Data Platform Engineering
Staff Software Engineer, Data Platform Engineering

Paradigm Health, Inc. • New York (NY), Northern (KY)

Hybrid
USD 180,000 - 240,000
Staff Software Engineer, Data Platform Engineering at Paradigm Health
Staff Software Engineer, Data Platform Engineering at Paradigm Health

Matcha • Northern (KY)

Hybrid
USD 150,000 - 240,000
Staff Software Engineer, Data Platform Engineering
Staff Software Engineer, Data Platform Engineering

Paradigm • New York (NY)

Hybrid
USD 180,000 - 240,000
Senior Data Scientist: LLM Eval & BI for Trials
Senior Data Scientist: LLM Eval & BI for Trials

Paradigm Health, Inc. • Town of Columbus (NY)

Hybrid
USD 120,000 - 180,000
IT Data Engineer III
IT Data Engineer III

Paradigm • Lombard (IL)

On-site
USD 110,000 - 140,000
Health insurance
401(k) matching
Paid time off
+3
Staff Data Scientist
Staff Data Scientist

MD Ally • Cambridge (MA)

On-site
USD 200,000 - 325,000
Equitable and inclusive culture
Accommodations for application process
Staff Data Scientist
Staff Data Scientist

CloudDevs • United States

On-site
USD 90,000 - 130,000
Senior Data Platform Engineer
Senior Data Platform Engineer

Ellipsis Health • San Francisco (CA)

On-site
USD 150,000 - 170,000
401(k) matching
Health insurance
Vision & Dental insurance
+1
Senior AI/ML Engineer
Senior AI/ML Engineer

Prudentia Sciences • Boston (MA)

On-site
USD 206,000 - 240,000
Remote-friendly
Competitive compensation
Equity
Head of Data Science
Head of Data Science

Stealth Startup • United States

On-site
USD 180,000 - 260,000
Health benefits
Dental benefits