Data Scientist

M Science

United States

On-site

USD 90,000 - 175,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Annual discretionary incentive bonus
Medical coverage
Dental coverage
Vision coverage
401(k)
Life insurance
Disability insurance
Wellness programs
Employee discount program
Paid time off packages

Job summary

M Science, a data-driven research and analytics firm in New York, seeks a Data Scientist to design and develop pipelines and AI/ML models across 50+ alternative data panels. You will build production‑grade code in Python, SQL, PySpark, and cloud environments, testing new data assets and contributing to the analytics library.

The role emphasizes statistical modeling, data quality, and scalable data pipelines with Databricks, Airflow, and Spark, plus strong problem-solving and testing practices.

Qualifications

  • Advanced Python for data processing, scripting, and automation.
  • Fluency in PySpark/Spark or SQL.
  • Bachelor's or higher degree in Statistics, Mathematics, Computer Science, Information Science, or a similar quantitative discipline.
  • Excellent knowledge of multivariate statistical analysis and panel methods (e.g., OLS, PCA, factor analysis, LDA).
  • Experience with gen-AI tools for workflow improvement.

Responsibilities

  • Solve complex data science problems using appropriate statistical and/or ML models; present results to stakeholders.
  • Process, cleanse, and verify the integrity of data used for analysis.
  • Create automated alerting and notification systems for deviations in data quality, validation failures, or unusual patterns.
  • Evaluate new datasets for the firm.
  • Contribute to the firm’s documented, unit-tested analytics library.
  • Design, develop, and optimize scalable and fault-tolerant data pipelines using Databricks, Airflow, Python, and Spark.
  • Build resilient data pipelines that handle vendor-related issues such as delayed deliveries and schema changes.
  • Utilize agentic workflows to automate redundant work.

Skills

Advanced Python
PySpark/Spark or SQL
Troubleshooting data pipelines
Gen-AI workflow tools

Education

Bachelor's or higher in a quantitative field

Tools

Databricks
Airflow
Spark
Python
SQL

Job description

Title

Data Scientist

Location

New York, NY

About M Science

M Science is a data-driven research and analytics firm, uncovering new insights for leading financial institutions and corporations. M Science is revolutionizing research, discovering new data sets, and pioneering methodologies to provide actionable intelligence. Our research teams have decades of experience working with massive amounts of unstructured data in near real-time to discern critical insights that help clients make smarter, more informed decisions. We combine the best of finance, data, and technology to create a truly unique value proposition for both financial services firms and major corporations.

Job Overview

We are seeking a highly skilled Data Scientist to design and develop pipelines and AI/ML models and workflows on our 50+ alternative data panels. The ideal candidate will have deep expertise in mathematical and statistical modeling and will have experience building data pipelines and models using Python, SQL, and PySpark. This person will test new data assets for the firm, solve complex data problems, contribute to the firm’s analytics library, and use traditional machine learning and statistical methods to improve panel data. M Science expects its data scientists to implement production code, so the ideal candidate will have experience writing well tested, performant, object-oriented code.

Responsibilities
  • Solve complex data science problems using appropriate statistical and/or ML models; present results to the stakeholders
  • Process, cleanse, and verify the integrity of data used for analysis
  • Create automated alerting and notification systems for deviations in data quality, validation failures, or unusual patterns
  • Evaluate new datasets for the firm
  • Contribute to the firm’s documented, unit-tested analytics library
  • Design, develop, and optimize scalable and fault-tolerant data pipelines using Databricks, Airflow, Python, and Spark
  • Build resilient data pipelines that handle vendor-related issues such as delayed deliveries, schema changes, incomplete records, and data corruption
  • Able to utilize agentic workflows to automate redundant work
Qualifications
  • Advanced Python for data processing, scripting, and automation
  • Fluency in PySpark/Spark or SQL
  • Bachelor's or higher degree, or significant experience in Statistics, Mathematics, Computer Science, Information Science, or a similar quantitative discipline
  • Excellent knowledge of multivariate statistical analysis, including but not limited to ordinary least squares, principal component analysis, factor analysis, LDA, and panel methods
  • Excellent knowledge of other ML methods including additive modeling and ensemble modeling
  • Experience using gen-AI tools for workflow improvement
  • Experience with named entity resolution methods a strong plus
  • Familiarity with cloud data platforms (AWS) and cloud-based storage solutions
  • Strong troubleshooting skills to diagnose and resolve performance bottlenecks in data pipelines
Primary Location

New York, NY

Salary Range

$90,000-$175,000 USD/Annual

For eligible employees, the following benefits are offered:

  • Annual discretionary incentive bonus
  • Medical coverage
  • Dental coverage
  • Vision coverage
  • 401(k)
  • Life insurance
  • Accident insurance
  • Disability insurance
  • Wellness programs including discounted and flexible gym memberships
  • Robust employee discount program
  • Paid time off packages that include planned time off (vacation), unplanned time off (sick leave), paid holidays and paid parental leave
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Scientist
Data Scientist

M Science • New York (NY)

On-site
USD 90,000 - 175,000
Annual discretionary incentive bonus
Medical, dental & vision coverage
401(k)
+3
Data Scientist (US)
Data Scientist (US)

AVP VIGILANT TECHNOLOGY PVT LTD • New York (NY)

On-site
USD 110,000 - 190,000
Competitive salary
Growth opportunities
Comprehensive benefits
+2
Data Scientist
Data Scientist

The Wilson Group KW23 • Fairbanks (AK)

Remote
USD 90,000 - 150,000
Flexible work hours
Competitive pay
Professional development opportunities
+1
Data Scientist
Data Scientist

Atrosphere Technologies • New York (NY)

Hybrid
USD 80,000 - 120,000
Competitive salary and equity options
Comprehensive health and dental insurance
Flexible work arrangements
+3
Senior Data Scientist
Senior Data Scientist

United States Digital Space LLC • Seattle (WA)

Hybrid
USD 192,000 - 288,000
Equity
Company bonus
401(k) plan
+2
Data Scientist
Data Scientist

myBridge Corporation • Washington

On-site
USD 120,000 - 130,000
Medical Insurance
Dental Insurance
Vision Insurance
+2
Data Scientist Opportunity
Data Scientist Opportunity

Bridge Technologies and Solutions • Boston (MA)

On-site
USD 100,000 - 120,000
Data Scientist
Data Scientist

Peraton • United States

Hybrid
USD 80,000 - 128,000
Data Scientist
Data Scientist

Franchise World Headquarters, LLC • Miami (FL)

On-site
USD 102,000 - 129,000
Insurance Plans (Medical, Life)
Pension/401K/RSP
Competitive Bonus
+4
Data Scientist
Data Scientist

Hunter Bond • Town of Texas (WI)

On-site
USD 130,000
Extremely competitive bonus tied to performance
Structured learning budget for courses and certifications
Clear progression path into senior data roles