Data Scientist

RitePros

Portland (ME)

On-site

USD 100,000 - 150,000

Full time

2 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

RitePros seeks a Data Scientist with a Bachelor's degree in Computer Science or related field to collaborate across product, data, and business teams. You will translate requirements into analytical tasks, develop ML models, and build data pipelines.

You will create dashboards, perform statistical analyses, and contribute to Retrieval-Augmented Generation projects with embeddings and vector databases. Remote-friendly? No, Portland-based with client travel.

Qualifications

  • Bachelor's degree in Computer Science, CIS, IT or equivalent.
  • Experience translating business requirements into analytical tasks and success metrics.
  • Proficiency in Python/SQL with data processing libraries.
  • Familiarity with ML models, evaluation metrics, and dashboards.

Responsibilities

  • Collaborate with product owners and stakeholders to understand data and analytics requirements.
  • Translate business requirements into analytical tasks, specs, success metrics, and steps.
  • Collect, clean, profile, and analyze structured and unstructured data to identify opportunities.
  • Perform exploratory Data Analysis and statistical tests.
  • Develop, train, tune, and evaluate ML models for various use cases.
  • Prepare dashboards and visualizations to communicate results.
  • Develop Retrieval-Augmented Generation pipelines using embeddings and vector databases.
  • Apply prompt-engineering and guardrails for Generative AI applications.
  • Develop Python-based components and reusable APIs for data processing and AI.
  • Contribute to data models and analytical schemas across databases and warehouses.
  • Support ETL/ELT pipelines and data-quality checks.
  • Coordinate with cross-functional teams across locations.

Skills

Python
SQL
Pandas
NumPy
PySpark
Statistical Analysis
Data Visualization
Dashboarding
Communication
Problem Solving

Education

Bachelor’s degree in Computer Science / CIS / IT or equivalent

Tools

MongoDB
PostgreSQL
MySQL
SQL Server
BigQuery
Snowflake
Databricks
FastAPI
Flask
Power BI
Tableau

Job description

Contact with us through our representative or submit a business inquiry online.

Data Scientist with Bachelor’s degree in Computer Science, Computer Information Systems, Information Technology, or a combination of education and experience equating to the U.S. equivalent of a Bachelor’s degree in one of the aforementioned subjects.
Job Duties and Responsibilities:
  • Collaborate with product owners, data scientists, data engineers, analysts, architects, and business stakeholders to understand data, artificial intelligence, and analytics requirements.
  • Contribute to translating business requirements into analytical tasks, technical specifications, success metrics, acceptance criteria, and implementation steps.
  • Collect, clean, profile, analyze, and interpret structured, semi-structured, and unstructured datasets to identify trends, patterns, anomalies, and business opportunities.
  • Perform exploratory data analysis, statistical analysis, hypothesis testing, correlation analysis, segmentation, and feature engineering using established analytical methods.
  • Develop, train, tune, test, and evaluate machine-learning models for regression, classification, clustering, forecasting, recommendation, and anomaly-detection use cases.
  • Compare model performance using metrics such as precision, recall, F1 score, AUC, RMSE, and MAE, and prepare dashboards, reports, and visualizations to communicate the results.
  • Develop and support Retrieval-Augmented Generation pipelines using enterprise data, embeddings, semantic search, hybrid search, reranking, vector databases, and large language models.
  • Apply prompt-engineering, grounding, citation-validation, structured-output, and response-evaluation techniques to improve the accuracy and reliability of Generative AI applications.
  • Develop, test, and support single-agent and multi-agent workflows using LangChain, LangGraph, LlamaIndex, or comparable Agentic AI frameworks.
  • Support human-in-the-loop approvals, confidence thresholds, fallback mechanisms, retry policies, and testing for hallucinations, bias, prompt injection, sensitive-data exposure, and unauthorized tool usage.
  • Develop Python-based components and reusable APIs for data processing, machine learning, Generative AI, and Agentic AI applications.
  • Contribute to the development and maintenance of data models and analytical schemas using MongoDB, relational databases, cloud data warehouses, and data lakes.
  • Develop and support ETL and ELT pipelines for data ingestion, cleansing, transformation, validation, normalization, enrichment, and analytical preparation.
  • Support batch, streaming, Change Data Capture, and event-driven data pipelines, including data-quality checks, schema validation, logging, metadata management, and monitoring.
  • Participate in testing, deployment, versioning, monitoring, troubleshooting, and documentation of machine-learning models, AI applications, and data pipelines while coordinating with cross-functional, onshore, and offshore teams.
Technologies Involved / Skills required for the position:
  • Effective communication, collaboration, analytical, problem-solving, and documentation skills for working with cross-functional technical and business teams.
  • Ability to understand business requirements and translate them into analytical tasks, technical specifications, success metrics, acceptance criteria, and implementation steps.
  • Proficiency in Python and SQL, with experience using Pandas, NumPy, PySpark, or comparable technologies to collect, clean, profile, transform, and analyze structured, semi-structured, and unstructured data.
  • Working knowledge of exploratory data analysis, statistical analysis, hypothesis testing, correlation analysis, segmentation, and feature-engineering techniques.
  • Experience developing and evaluating regression, classification, clustering, forecasting, recommendation, and anomaly-detection models using Scikit-learn, XGBoost, LightGBM, TensorFlow, PyTorch, or comparable frameworks.
  • Knowledge of model-evaluation metrics, including precision, recall, F1 score, AUC, RMSE, and MAE, with experience creating dashboards, reports, and visualizations using Power BI, Tableau, Looker, Omni, or comparable tools.
  • Experience developing Retrieval-Augmented Generation solutions using embeddings, semantic search, hybrid search, reranking, vector databases, large language models, and enterprise data.
  • Working knowledge of prompt engineering, grounding, citation validation, structured outputs, response evaluation, and guardrails for Generative AI applications.
  • Experience developing and testing single-agent and multi-agent workflows using LangChain, LangGraph, LlamaIndex, or comparable Agentic AI frameworks.
  • Knowledge of human-in-the-loop processes, confidence thresholds, fallback mechanisms, retry policies, hallucination evaluation, bias detection, prompt-injection prevention, and sensitive-data protection.
  • Proficiency in developing Python-based data-processing, machine-learning, Generative AI, and Agentic AI components, with knowledge of reusable APIs using FastAPI, Flask, or comparable frameworks.
  • Experience working with data models and analytical schemas using MongoDB, PostgreSQL, MySQL, SQL Server, BigQuery, Snowflake, Databricks, or comparable database and cloud data technologies.
  • Experience developing and supporting ETL and ELT pipelines for data ingestion, cleansing, transformation, validation, normalization, enrichment, and analytical preparation.
  • Working knowledge of batch, streaming, Change Data Capture, and event-driven pipelines using technologies such as Apache Airflow, Kafka, Google Cloud Pub/Sub, AWS SQS, SNS, or EventBridge.
  • Understanding of testing, deployment, versioning, monitoring, troubleshooting, and documentation practices for machine-learning models, AI applications, and data pipelines, including familiarity with Docker, CI/CD, MLOps, and LLMOps.

Work location is Portland, ME with required travel to client locations throughout USA.

Rite Pros is an equal opportunity employer (EOE).

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Scientist
Senior Data Scientist

Jobtailor • Minnesota

On-site
USD 150,000 - 230,000
Senior Data Scientist GenAI / RAG
Senior Data Scientist GenAI / RAG

Compunnel, Inc. • Houston (TX)

On-site
USD 120,000 - 150,000
GenAI Consultant/AI Engineer/Data Scientist
GenAI Consultant/AI Engineer/Data Scientist

NeerInfo Solutions • Austin (TX)

On-site
USD 120,000 - 150,000
Senior Data Scientist
Senior Data Scientist

HRB • United States

Remote
USD 170,000 - 210,000
Remote work
Data Scientist, Level 2
Data Scientist, Level 2

Spectraforce Technologies • Newark (NJ)

Hybrid
USD 100,000 - 140,000
Data Scientist
Data Scientist

Compunnel, Inc. • Manassas (VA)

On-site
USD 90,000 - 120,000
Data Scientist
Data Scientist

Agentic Dream • United States

Remote
USD 120,000 - 170,000
Senior Data Scientist
Senior Data Scientist

Unisys • McLean (VA)

On-site
USD 130,000 - 160,000
Senior Full Stack AI and Data Engineer
Senior Full Stack AI and Data Engineer

RBA, Inc. • Minneapolis (MN)

Hybrid
USD 100,000 - 130,000
Flexible work schedule
Collaborative work environment
Career development opportunities
Data Scientist
Data Scientist

Creative Solutions Services, LLC • New Jersey

Hybrid
USD 90,000 - 103,000