Data Scientist III

Kaiser Permanente

Northern (KY)

Hybrid

USD 95,000 - 125,000

Full time

10 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Kaiser Permanente in the United States seeks an individual contributor to design and develop data pipelines and automation for acquiring and ingesting raw data from multiple sources, under the guidance of senior data scientists. The role focuses on transforming data into usable features for machine learning, training models, and deploying them into production.

You will work with internal and external stakeholders across domains to deliver statistically driven outcomes, verify model performance,

Qualifications

  • Minimum two years of experience with Exploratory Data Analysis (EDA) and visualization methods.
  • Minimum one year machine learning and/or algorithmic experience.
  • Minimum two years statistical analysis and modeling experience.
  • Minimum two years programming experience.
  • Bachelor's degree in Mathematics, Statistics, Computer Science, Engineering, Economics, Public Health, or related field AND Minimum three years experience in data science or a directly related field. Advanced degrees may be substituted for the work experience requirements.

Responsibilities

  • Participates in the design and development of data pipelines and automation for data acquisition and ingestion of raw data from multiple data sources and data formats.
  • Writes and optimizes diverse SQL queries and demonstrates knowledge of database fundamentals.
  • Analyses data sets, summarizes key characteristics, and visualizes patterns to test hypotheses.
  • Develops features used in machine learning and trains models under guidance of senior data scientists.
  • Deploys and maintains models in production and verifies performance.
  • Collaborates with internal and external stakeholders across domains to deliver statistical driven outcomes.

Skills

Ambiguity/Uncertainty Management
Attention to Detail
Communication
Critical Thinking
Problem Solving
Teamwork
Data Visualization
Decision Making
Learning Agility
Cross-Group Collaboration
Organizational Savvy

Education

Bachelor's degree in Mathematics, Statistics, Computer Science, Engineering, Economics, Public Health, or related field

Tools

Open Source Languages & Tools
Relational Database Management
Microsoft Excel
Data Visualization Tools

Job description

Job Summary:

This individual contributor is primarily responsible for participating in the design and development of data pipelines and automation for data acquisition and ingestion of raw data from multiple data sources and data formats under the guidance of more senior data scientists. This role is also responsible for developing detailed problem statements outlining hypotheses and their effect on target clients/customers, analyzing and investigating data sets and summarizing key characteristics, selecting, manipulating and transforming data into features used in machine learning algorithms, training statistical models under the guidance of more senior data scientists, deploying and maintaining reliable and efficient models through production, verifying model performance, and working with internal and external stakeholders across domains to develop and deliver statistical driven outcomes.

Essential Responsibilities:
  • Pursues effective relationships with others by proactively providing resources, information, advice, and expertise with coworkers and members. Listens to, seeks, and addresses performance feedback; provides mentoring to team members. Pursues self-development; creates plans and takes action to capitalize on strengths and develop weaknesses; influences others through technical explanations and examples. Adapts to and learns from change, challenges, and feedback; demonstrates flexibility in approaches to work; helps others adapt to new tasks and processes. Supports and responds to the needs of others to support a business outcome.
  • Completes work assignments autonomously by applying up-to-date expertise in subject area to generate creative solutions; ensures all procedures and policies are followed; leverages an understanding of data and resources to support projects or initiatives. Collaborates cross-functionally to solve business problems; escalates issues or risks as appropriate; communicates progress and information. Supports, identifies, and monitors priorities, deadlines, and expectations. Identifies, speaks up, and implements ways to address improvement opportunities for team.
  • Develops detailed problem statements outlining hypotheses and their effect on target clients/customers by defining scope, objectives, outcome statements and metrics.
  • Participates in the design and development of data pipelines and automation for data acquisition and ingestion of raw data from multiple data sources and data formats under the guidance of more senior data scientists by transforming, cleansing, and storing data for consumption by downstream processes; writing and optimizing diverse SQL queries; and demonstrating a working knowledge of database fundamentals.
  • Analyzes and investigates data sets and summarizes key characteristics by employing data visualization methods; and determining how best to manipulate data sources to discover patterns, spot anomalies, test hypotheses, and/or check assumptions.
  • Selects, manipulates, and transforms data into features used in machine learning algorithms by leveraging techniques to conduct dimensionality reduction, feature importance, and feature selection.
  • Trains statistical models under the guidance of more senior data scientists by using algorithms and data mining techniques; testing models with various algorithms to assess the input dataset and related features; and applying techniques to prevent overfitting such as cross-validation.
  • Deploys and maintains reliable and efficient models through production.
  • Verifies model performance by demonstrating a working knowledge of a variety of model validation techniques to assess and discriminate the goodness of model fit; and leveraging feedback and output to manage and strengthen model performance.
  • Works with internal and external stakeholders across domains to develop and deliver statistical driven outcomes by delivering insights and values from heterogeneous data to investigate problems for multiple use cases; driving informed decision-making; and presenting findings to both technical and non-technical audiences.
Knowledge, Skills and Abilities: (Core)
  • Ambiguity/Uncertainty Management
  • Attention to Detail
  • Business Knowledge
  • Communication
  • Critical Thinking
  • Cross-Group Collaboration
  • Decision Making
  • Dependability
  • Diversity, Equity, and Inclusion Support
  • Drives Results
  • Facilitation Skills
  • Health Care Industry
  • Influencing Others
  • Integrity
  • Learning Agility
  • Organizational Savvy
  • Problem Solving
  • Short- and Long-term Learning & Recall
  • Teamwork
  • Topic-Specific Communication
Knowledge, Skills and Abilities: (Functional)
  • Advanced Quantitative Data Modeling
  • Algorithms
  • Applied Data Analysis
  • Business Intelligence Tools
  • Data Ensemble Techniques
  • Data Extraction
  • Data Manipulation/Wrangling
  • Data Visualization Tools
  • Design Thinking
  • Feature Analysis/Engineering
  • Machine Learning
  • Microsoft Excel
  • Model Optimization
  • Open Source Languages & Tools
  • Relational Database Management
Minimum Qualifications:
  • Minimum two (2) years experience working with Exploratory Data Analysis (EDA) and visualization methods.
  • Minimum one (1) year machine learning and/or algorithmic experience.
  • Minimum two (2) years statistical analysis and modeling experience.
  • Minimum two (2) years programming experience.
  • Bachelors degree in Mathematics, Statistics, Computer Science, Engineering, Economics, Public Health, or related field AND Minimum three (3) years experience in data science or a directly related field. Additional equivalent work experience in a directly related field may be substituted for the degree requirement. Advanced degrees may be substituted for the work experience requirements.
Preferred Qualifications:
  • One (1) year experience in a leadership role with or without direct reports.
  • One (1) year healthcare experience.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Scientist III
Data Scientist III

Kaiser Permanente • Fulton (MD)

On-site
USD 110,000 - 170,000
Data Scientist
Data Scientist

The Home Depot • Atlanta (GA)

On-site
USD 110,000 - 160,000
Senior Data Scientist
Senior Data Scientist

Jobtailor • Town of Florida (NY)

On-site
USD 90,000 - 130,000
Data Scientist - Delivery Data Science
Data Scientist - Delivery Data Science

The Home Depot • Atlanta (GA)

On-site
USD 120,000 - 170,000
Data Scientist
Data Scientist

Compunnel, Inc. • Camden (NJ)

On-site
USD 100,000 - 130,000
Data Scientist
Data Scientist

IRB USA Inspire Resources • Atlanta (GA)

On-site
USD 85,000 - 115,000
Data Scientist 2 4P/187
Data Scientist 2 4P/187

4P Consulting Inc. • Atlanta (GA)

On-site
USD 80,000 - 120,000
Data Scientist
Data Scientist

Paladin Consulting • The Woodlands (TX)

On-site
USD 120,000 - 160,000
Data Scientist - People Analytics
Data Scientist - People Analytics

The Home Depot • Atlanta (GA)

On-site
USD 90,000 - 120,000
Senior Data Scientist, Delivery
Senior Data Scientist, Delivery

The Home Depot • Atlanta (GA)

On-site
USD 130,000 - 170,000