Data Scientist

TruLegal (formerly TRU Staffing)

United States

Hybrid

USD 165,000 - 225,000

Full time

3 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Hybrid work arrangement

Job summary

TruLegal (formerly TRU Staffing) seeks a Data Scientist to join its AI & Data Analytics team. You will deliver AI solutions across active matters and firm-wide initiatives, embedding with case teams and building production-ready data pipelines. This role blends development, consulting, and deployment in a hybrid setup near a U.S.

or London office. You will work with Python, SQL, LLMs, and NLP to drive retrieval-augmented generation and knowledge retrieval, while collaborating with IT and

Qualifications

  • Bachelor's degree or higher in a quantitative field; Master’s preferred.
  • Experience in data science, ML, or advanced analytics; familiarity with confidential environments.
  • Proficiency in Python for analysis, modeling, and prototyping; advanced SQL.
  • NLP experience: text classification, entity recognition, semantic search, and parsing.
  • Hands-on with LLM platforms (Claude, OpenAI, Hugging Face) and RAG architectures.
  • Experience building data pipelines for structured and unstructured data.
  • Strong visualization skills (Power BI/Tableau) and Excel.
  • Ability to measure model outputs with defined metrics and communicate results clearly.

Responsibilities

  • Partner with trial teams to scope data problems and deliver executable solutions under deadlines.
  • Build workflows for document-heavy matters: LLM-based extraction, entity recognition, classification.
  • Support early case assessment and dispute intelligence by structuring case materials.
  • Design data-modeling projects across discovery datasets and third-party data sources.
  • Support trial and witness preparation workflows and evolve case intelligence.
  • Produce clear visualizations and analyses for attorneys to act on.

Skills

Python for data analysis
Advanced SQL
NLP & ML
LLMs & RAG
Data pipelines
Visualization tools
Model evaluation
Communication
Version control

Education

Bachelor's degree in Data Science, Computer Science, Statistics, Mathematics, or related quantitative field
Master's degree preferred

Tools

Claude
OpenAI
Hugging Face
LangChain
Vector databases
Power BI
Tableau

Job description

Our client, a leading global litigation law firm, is seeking a Data Scientist to join its AI & Data Analytics team and help deliver sophisticated AI solutions across active legal matters and firmwide initiatives. This hands-on, highly consultative role will partner directly with case teams to manage AI projects, deploy proprietary technology, design scalable AI capabilities, and evaluate emerging tools to determine the right build-versus-buy approach. The ideal candidate is a technically strong, flexible generalist with hands-on experience across Python, SQL, LLMs, RAG, NLP, data pipelines, and AI evaluation who can move comfortably between technical development, consulting, and practical deployment. This is a hybrid position based near one of the firm's U.S. or London offices.

About The Role

We are looking for a Data Scientist to join the AI & Data Analytics Group. This is a hands-on, matter-facing role at the intersection of data science, artificial intelligence, and trial practice.

You will build and deploy tools and models that change how our litigators find, analyze, and act on information - from large-scale document review and information extraction to case intelligence and knowledge retrieval. Roughly half your time will be spent embedded directly with case teams on live matters, at times including client-facing work; the rest will go to firm-wide tooling, evaluation, and capability building.

This is a role for someone who wants to solve hard problems with substantial real-world impact, on a compressed litigation clock.

What You'll Do
Matter-embedded data science
  • Partner with trial teams to scope data problems on active matters and turn ambiguous asks into structured, executable solutions, often under deadline pressure set by the court.
  • Build custom workflows for document-heavy matters: LLM-based information extraction, semantic chunking, entity recognition, and classification.
  • Support early case assessment and dispute intelligence — turning raw case materials into structured knowledge that lets attorneys identify patterns and test case theories.
  • Design and execute data-modeling projects across discovery datasets, court filings, deposition transcripts, and third-party legal data sources.
  • Support trial preparation and witness preparation workflows, and maintain case intelligence that evolves as a matter progresses.
  • Produce clear visualizations and analyses that attorneys can act on directly, and that hold up under scrutiny from opposing counsel.
AI solution development
  • Prototype, build, and deploy retrieval-augmented generation (RAG) pipelines against firm and matter document corpora.
  • Develop and maintain the data pipelines that feed them, working with both structured data (timekeeping, matter metadata, docket data) and unstructured document sets.
  • Move promising prototypes into production in partnership with IT and Information Security.
Evaluation and verification
  • Define and run evaluation frameworks for AI outputs — precision, recall, extraction accuracy, hallucination rates, and robustness.
  • Design and maintain the human verification protocols that sit between model output and attorney work product.
  • Monitor deployed model performance over time and flag drift or degradation.
  • Support the firm's AI governance program on responsible use, data privacy, confidentiality, privilege, and compliance with firm policies and client requirements.
Capability building
  • Serve as a technical resource on AI for the group and the broader firm.
  • Contribute to firm-wide enablement, including AI Office Hours and attorney training sessions.
  • Evaluate emerging tools, models, and vendor offerings, and give clear recommendations on what is worth adopting.
  • Document your work so others can maintain and extend it.
What You'll Bring
Required
  • Experience: Demonstrated experience in data science, machine learning, or advanced analytics. Experience in a law firm, legal services, professional services, or other confidentiality-sensitive environment is strongly preferred.
  • Education: Bachelor's degree in Data Science, Computer Science, Statistics, Mathematics, or a related quantitative field. Master's preferred.
  • Programming: Strong Python for data analysis, modeling, and prototyping (pandas, scikit-learn, and at least one deep learning framework). Advanced SQL. Claude Skills.
  • NLP: Practical experience with text classification, named entity recognition, semantic search, summarization, and document parsing.
  • LLMs: Hands-on work with modern LLM platforms and tooling (e.g., Anthropic Claude, OpenAI, Hugging Face, LangChain), including RAG architectures and vector databases.
  • Pipelines: Demonstrated experience building data pipelines and preparing both structured and unstructured datasets for production use.
  • Visualization: Proficiency with Power BI, Tableau, or equivalent, plus strong Excel.
  • Evaluation: Experience measuring model and LLM output quality with defined metrics.
  • Judgment under pressure: Comfort working to litigation deadlines, and the discipline to state plainly what a model does and does not establish.
  • Communication: Ability to explain technical concepts to attorneys and clients in plain language, and to push back constructively when a proposed use of AI is a poor fit.
  • Version control: Git or equivalent.
Preferred
  • Familiarity with eDiscovery platforms and workflows, particularly Relativity, including structured analytics and custom development.
  • Working knowledge of legal data types — matter metadata, timekeeping data, docket and court filing data, billing data.
  • Exposure to the litigation lifecycle and the practical constraints of privilege, confidentiality, work product, and protective orders.
  • Experience supporting expert work, damages analysis, or other quantitative analysis offered in a contested proceeding.
  • Cloud experience (Azure, AWS, or GCP).
  • Front-end skills (JavaScript, HTML, CSS) sufficient to build lightweight internal tools and dashboards.
  • Familiarity with legal technology platforms: document management, docketing, and case management systems.

Expected salary for this role is $165,000 - $225,000, commensurate with experience, training, skills, qualifications, and other market factors.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data Scientist (Litigation)
Data Scientist (Litigation)

TruLegal (formerly TRU Staffing) • United States

On-site
USD 133,000 - 191,000
Sr. Data Scientist
Sr. Data Scientist

Insight Global • Raleigh (NC)

On-site
USD 140,000 - 180,000
Innovations and AI Solutions Engineer
Innovations and AI Solutions Engineer

Wilson Sonsini Goodrich & Rosati • United States

Hybrid
USD 120,000 - 158,000
Senior Data Scientist II
Senior Data Scientist II

RELX • California (MO)

On-site
USD 105,000 - 175,000
Annual incentive bonus
Legal AI Engineer
Legal AI Engineer

Epiq • New York (NY)

On-site
USD 140,000 - 190,000
Litigation Paralegal
Litigation Paralegal

Perplexity • San Francisco (CA)

On-site
USD 120,000 - 195,000
Legal Data Engineer Lead
Legal Data Engineer Lead

Greenberg Traurig, LLP • Miami (FL)

On-site
USD 120,000 - 190,000
Legal Engineer- Litigation
Legal Engineer- Litigation

Cozen O'Connor Corporation • Philadelphia

On-site
USD 120,000 - 160,000
Legal Data Engineer Lead
Legal Data Engineer Lead

Greenberg Traurig, LLP • Charlotte (NC)

On-site
USD 120,000 - 160,000
Legal AI Systems Analyst
Legal AI Systems Analyst

TBG | The Bachrach Group • New York (NY)

On-site
USD 90,000 - 130,000