Data Scientist

Jobtailor

Raleigh (NC)

On-site

USD 110,000 - 170,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Jobtailor is seeking a data science professional in Raleigh to design and deploy statistical and machine learning models for time-series forecasting, anomaly detection, and asset health scoring across utility networks.

You will build end-to-end ML pipelines on Databricks from feature engineering to deployment, validate models, monitor performance, and ensure regulatory compliance with stakeholders.

Qualifications

  • Bachelor's degree or equivalent in a quantitative field.
  • 3+ years of applied statistical modeling and ML with deployed production models.
  • Proficient Python (PySpark, pandas, NumPy, scikit-learn) and R.
  • Experience with time-series modeling on high-volume data.
  • Familiar with Databricks ML ecosystem and MLflow.
  • Strong communication and ability to explain complex concepts to non-technical stakeholders.

Responsibilities

  • Design and implement statistical and ML models for forecasting and anomaly detection.
  • Build and maintain end-to-end ML pipelines on Databricks.
  • Translate business problems into modeling problems and select appropriate methods.
  • Ensure model interpretability and explainability for regulatory compliance.
  • Implement monitoring, drift detection, and retraining workflows.

Skills

Statistical Modeling
Machine Learning
Python Programming
R Programming
Time-Series Modeling
SQL
Databricks ML Ecosystem
Model Deployment
Model Monitoring
Communication
Problem-Solving
Explainability

Education

Bachelor’s degree or equivalent in Statistics, Applied Mathematics, Physics, Engineering, Data Science, or a related quantitative field

Tools

Databricks
PySpark
pandas
NumPy
scikit-learn
statsmodels
XGBoost
MLflow
Git

Job description

  • Design and implement statistical and machine learning models for time-series forecasting, anomaly detection, and asset health scoring across utility networks.
  • Build and maintain end-to-end ML pipelines on Databricks from feature engineering and model training to validation, deployment, and monitoring in production.
  • Apply classical statistical methods (GLMs, GAMs, mixed-effects models, Bayesian inference) alongside modern ML techniques (ensemble approaches, network analysis, neural networks) to solve grid operations problems.
  • Develop predictive maintenance and degradation models for utility infrastructure using telemetry and SCADA data at scale.
  • Translate ambiguous business problems into well-defined modeling problems with appropriate statistical frameworks - e.g., knowing when a LM/GLM is sufficient and when gradient boosting or deep learning is warranted.
  • Implement model monitoring, drift detection, and automated retraining workflows to maintain model performance over time.
  • Contribute to load forecasting, demand response optimization, and outage prediction systems.
  • Ensure model interpretability and explainability for utility stakeholders and regulatory compliance.
  • Contribute to internal knowledge-sharing on statistical best practices.
Requirements
  • Bachelor’s degree or equivalent in Statistics, Applied Mathematics, Physics, Engineering, Data Science, or a related quantitative field.
  • 3+ years of experience (with Bachelor’s), 2+ years of experience (with Masters), or 1+ years (with PhD) in applied statistical modeling and machine learning, with a track record of deployed production models.
  • Strong programming skills in Python (PySpark, pandas, NumPy, scikit-learn, statsmodels, XGBoost), R (tidyverse, lme4, glmmTMB, glmnet, mgcv), and SQL for large-scale data analysis.
  • Experience with time-series modeling (ARIMA, state-space models, LSTM, Darts, or similar) on high-volume meter data.
  • Exposure to Databricks ML ecosystem (Feature Store, Experiment Track, Model Serving, Mosaic AI) and MLflow.
  • Familiarity with distributed computing concepts - PySpark, Optuna/Ray, Spark SQL, partitioning strategies, and medallion architecture.
  • Understanding of software engineering principles - version control (Git), testing, CI/CD for ML systems.
  • Ability to communicate complex statistical/ML concepts to non-technical stakeholders.
Core Competencies

Demonstrates expertise in statistical modeling and machine learning for time-series forecasting and anomaly detection, with a strong focus on building and maintaining ML pipelines and ensuring model performance and interpretability. Proficient in translating business problems into statistical frameworks and communicating complex concepts to stakeholders.

Highest-signal resume keywords
  • Statistical Modeling
  • Machine Learning
  • Python Programming
  • Time-Series Modeling
  • Databricks ML Ecosystem
ATS Optimization Keywords
Hard Skills
  • Statistical Methods
  • Machine Learning Techniques
  • Feature Engineering
  • Model Training
  • Model Validation
  • Model Deployment
  • Model Monitoring
  • Anomaly Detection
  • Predictive Maintenance
  • Data Analysis
Soft Skills
  • Communication
  • Problem-Solving
Industry Keywords
  • Utility Networks
  • Telemetry Data
  • SCADA Data
  • Regulatory Compliance
  • Grid Operations
Tools & Technologies
  • Databricks
  • PySpark
  • SQL
  • MLflow
  • Git
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Scientist
Data Scientist

Inizio Partners Corp • New York (NY)

On-site
USD 85,000 - 120,000
Senior Data Engineer
Senior Data Engineer

Jobtailor • Utah

On-site
USD 120,000 - 180,000
Lead Data Scientist
Lead Data Scientist

BayOne Solutions • Oakland (CA)

On-site
USD 140,000 - 190,000
Principal Data Scientist
Principal Data Scientist

The Corporate • Oakland (CA)

On-site
USD 150,000 - 230,000
Data Scientist
Data Scientist

Jobtailor • Washington

On-site
USD 110,000 - 170,000
Machine Learning Engineer 3 4P/392
Machine Learning Engineer 3 4P/392

4P Consulting Inc. • Atlanta (GA)

On-site
USD 110,000 - 130,000
Data Science Manager
Data Science Manager

Jobtailor • Hoboken (NJ)

On-site
USD 180,000 - 240,000
Data Scientist
Data Scientist

Inizio Partners Corp • Houston (TX)

On-site
USD 80,000 - 110,000
Staff Data Scientist
Staff Data Scientist

Jobtailor • New York (NY)

On-site
USD 180,000 - 240,000
Principal Associate, Data Scientist
Principal Associate, Data Scientist

Jobtailor • Illinois

On-site
USD 130,000 - 170,000