Senior Data Engineer

Jobtailor

Utah

On-site

USD 120,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Jobtailor is seeking a data & ML platform engineer to design and maintain scalable data ingestion, transformation, and modeling pipelines. You will collaborate with data scientists, ML engineers, and business stakeholders to deliver robust data assets and AI-enabled solutions.

The role emphasizes cloud-based infrastructure, Databricks, PySpark, and modern tooling for end-to-end ML lifecycles, including feature stores and model observability. Utah location with on-site work expectations.

Qualifications

  • Experience designing and building scalable data ingestion and transformation pipelines.
  • Strong SQL development and query optimization skills.
  • Experience with Databricks and large-scale data processing.
  • Knowledge of data lake, lakehouse, data warehouse, and data mart architectures.
  • Experience with ML feature stores and data quality testing.
  • Ability to translate business processes into data/ML solutions.

Responsibilities

  • Collaborate with cross-functional teams to enable effective use of core data assets.
  • Design, develop, and maintain scalable data ingestion and transformation pipelines.
  • Build and optimize data lake, lakehouse, warehouse, and data mart architectures.
  • Develop and maintain data models including facts, dimensions, feature datasets, and domain-specific data products.
  • Translate business requirements into design docs and ML feature pipelines.
  • Design and manage cloud-based data and ML infrastructure (Databricks preferred).
  • Design, build, and operationalize ML pipelines for training, validation, deployment, and observability.
  • Support ML model lifecycle management, including versioning and lineage.
  • Develop and maintain ML feature stores and reusable feature pipelines.
  • Build AI-powered applications and agentic workflows.
  • Implement data pipelines for AI systems including unstructured data.
  • Develop tests (unit/integration/data quality) and participate in code reviews.

Skills

Python Development
SQL Development
Data Modeling
Data Pipeline Development
MLOps
AI Prompt Frameworks
Cross-Functional Collaboration
Requirement Gathering
Adaptability

Education

Bachelor's Degree in Computer Science
Bachelor's Degree in Information Systems
Quantitative Field

Tools

Databricks
AWS
PySpark
Dbt
Git

Job description

  • Collaborate with data analysts, data scientists, ML engineers, software engineers, and business stakeholders to enable effective use of core data assets
  • Design, develop, and maintain scalable data ingestion and transformation pipelines using Python, SQL, and modern data tooling
  • Build and optimize data lake, lakehouse, warehouse, and data mart architectures
  • Develop and maintain data models including facts, dimensions, feature datasets, and domain-specific data products
  • Translate business requirements into design documents (e.g., ERDs, data flow diagrams) data models and ML feature pipelines
  • Design and manage cloud-based data and ML infrastructure (Databricks preferred), including development, staging, and production environments
  • Design, build, and operationalize machine learning pipelines for training, validation, deployment, and observability (e.g., performance, drift, reliability)
  • Support ML model lifecycle management, including versioning, reproducibility, and lineage
  • Develop and maintain ML feature stores and reusable feature pipelines for ML models
  • Build and integrate AI-powered applications and agentic workflows (e.g., LLM-based agents, retrieval-augmented generation systems, workflow automation agents)
  • Design and implement data pipelines for AI systems, including unstructured data (text, logs, embeddings, vector stores)
  • Develop and maintain unit, integration, and data quality tests
  • Participate in peer code reviews, pull requests, and team coding standards
  • Document data pipelines, ML pipelines, models, infrastructure, and standard operating procedures
  • Define infrastructure as code and support CI/CD pipelines for data and ML systems
  • Ensure data privacy, security, and access control best practices (including AI data governance considerations)
  • Identify and implement improvements in efficiency, scalability, resilience, and performance
  • Contribute to evolving data, ML, and AI platform architecture, tools, and best practices
Requirements
  • Ability to gather requirements and translate business processes into data, ML, and AI solutions
  • Comfortable working cross-functionally with both technical and non-technical stakeholders
  • Ability to quickly learn new domains and technologies
  • Strong Python development experience
  • Advanced SQL development and query optimization skills
  • Understanding of Databricks and large-scale data processing
  • Experience building and scaling data pipelines using Databricks and PySpark
  • Deep understanding of data lake, lakehouse, data warehouse, and data mart architectures
  • Experience with data modeling across a variety of business domains
  • Experience with modern data tooling (e.g., dbt or similar transformation frameworks)
  • Knowledge of data formats, data patterns, and modeling best practices
  • Experience with cloud platforms (AWS preferred)
  • Experience with CI/CD pipelines in a data engineering environment
  • Git-based development workflows
  • Hands‑on experience with AI prompt and agent frameworks (e.g., Claude Code, Cursor, Windsurfer, or similar)
  • Experience building AI agents and agentic workflows
  • Exposure to LLMs, embeddings, vector databases, or generative AI systems
  • Familiarity with handling structured and unstructured data (e.g., text, logs, embeddings)
  • Experience building or supporting machine learning pipelines in production
  • Familiarity with AI and MLOps in Databricks
  • Experience with ML feature engineering and feature stores
  • Understanding of ML model lifecycle management, monitoring, and evaluation
  • Bachelor's degree in computer science, information systems, a quantitative field, or equivalent practical experience
Core Competencies

Demonstrates expertise in designing and developing scalable data ingestion and transformation pipelines using Python and SQL, while effectively collaborating with cross‑functional teams to translate business requirements into actionable data and ML solutions. Proficient in managing cloud‑based data infrastructure and operationalizing machine learning pipelines, ensuring data privacy and governance best practices.

Highest-signal resume keywords
  • Python Development
  • SQL Development
  • Databricks Experience
  • Data Pipeline Development
  • Machine Learning Lifecycle Management
ATS Optimization Keywords
Hard Skills
  • Data Ingestion
  • Data Transformation
  • Data Modeling
  • Feature Engineering
  • CI/CD Pipelines
  • Unit Testing
  • Integration Testing
  • Data Quality Testing
  • AI Prompt Frameworks
  • MLOps
Soft Skills
  • Cross‑Functional Collaboration
  • Requirement Gathering
  • Adaptability
Certifications & Qualifications
  • Bachelor's Degree in Computer Science
  • Information Systems
  • Quantitative Field
Industry Keywords
  • Data Lake
  • Lakehouse
  • Data Warehouse
  • Data Mart
  • AI Systems
  • Unstructured Data
  • Generative AI
  • Data Governance
  • ML Feature Stores
  • Data Privacy
Tools & Technologies
  • Databricks
  • AWS
  • PySpark
  • Dbt
  • Git
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer
Data Engineer

Jobtailor • Denver (CO)

On-site
USD 140,000 - 190,000
Staff Engineer – Data Engineering
Staff Engineer – Data Engineering

Jobtailor • Arizona

On-site
USD 140,000 - 190,000
Senior Data Engineer
Senior Data Engineer

Jobtailor • Chicago (IL)

On-site
USD 130,000 - 180,000
Data Engineering Team Lead
Data Engineering Team Lead

Jobtailor • New York (NY)

On-site
USD 170,000 - 230,000
AI Data Scientist – Enterprise AI
AI Data Scientist – Enterprise AI

Jobtailor • Austin (TX)

On-site
USD 140,000 - 190,000
Senior Associate – Agentic AI and Machine Learning Developer
Senior Associate – Agentic AI and Machine Learning Developer

Jobtailor • Town of Florida (NY)

On-site
USD 110,000 - 150,000
Head of AI and Data Platform Engineering – Specialty Distribution
Head of AI and Data Platform Engineering – Specialty Distribution

Jobtailor • Town of Florida (NY)

On-site
USD 150,000 - 240,000
Data Engineer
Data Engineer

Jobtailor • Alabama

On-site
USD 95,000 - 130,000
Data Engineering Manager
Data Engineering Manager

Jobtailor • Town of Florida (NY)

On-site
USD 180,000 - 240,000
Distinguished Data Scientist
Distinguished Data Scientist

Jobtailor • California (MO)

On-site
USD 180,000 - 260,000