Sr. Data Engineer - AI

Dairy Challenge

Kansas City, Northern (KS, KY)

Hybrid

USD 99,000 - 139,000

Full time

24 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Dairy Challenge is seeking an experienced Data Engineer to design, build, and optimize scalable ETL/ELT pipelines supporting AI/ML model training and inference. You will ensure data accuracy, security, and reliability through cleansing, validation, and continuous updates, while troubleshooting complex integration issues and promoting best practices.

Ideal candidates will have 8+ years in data engineering, strong Python/SQL/PySpark skills, and experience with Azure Data Factory, Kafka, Snowflake,

Qualifications

  • Undergraduate degree in Computer Science, Data Engineering, Information Systems, or related field
  • 8+ years data engineering, software development, or related experience including data pipelines for AI/ML systems
  • Proficiency with Python, SQL, PySpark, Snowflake, Databricks, and Apache Airflow
  • Strong knowledge of relational and NoSQL databases, data warehouses or data lakes, and streaming tech like Kafka
  • Experience with at least one major cloud data ecosystem, preferably Azure, and familiarity with infrastructure-as-code and automated deployment practices
  • Certifications in cloud data engineering or data privacy/security and experience with MLOpspreferred

Responsibilities

  • Design, build, and optimize scalable ETL/ELT pipelines for AI/ML model training and inference
  • Ensure data quality, accuracy, and security through cleansing, normalization, validation, automation, and monitoring
  • Troubleshoot complex integration and performance issues and remove bottlenecks
  • Own key AI data infrastructure components with monitoring, logging, data lineage, and quality metrics
  • Apply enterprise architecture, AI governance, privacy, security, and CI/CD standards to data solutions
  • Collaborate with software engineers, data architects, domain experts, product and platform teams to align data solutions with model requirements and business goals
  • Deliver assigned work independently, influence cross-team decisions, promote data engineering best practices
  • Note: duties may vary and other responsibilities may be assigned

Skills

Python
SQL
PySpark
Snowflake
Databricks
Apache Airflow
Kafka
Azure

Education

Undergraduate degree in Computer Science, Data Engineering, Information Systems, or related field

Tools

Azure Data Factory
CI/CD
IaC (infrastructure-as-code)

Job description

Job Duties and Responsibilities:

  • Design, build, and optimize scalable ETL/ELT pipelines that ingest, transform, and integrate structured and unstructured enterprise data for AI/ML model training and inference
  • Ensure data is accurate, current, reliable, and secure through cleansing, normalization, validation, automation, continuous updates, monitoring, and error handling
  • Troubleshoot complex integration and performance issues, remove bottlenecks, and apply appropriate techniques such as caching, distributed processing, and specialized ETL patterns
  • Own key AI data infrastructure components and ensure they are scalable, maintainable, and supported by effective monitoring, logging, documentation, data lineage, and quality metrics
  • Apply enterprise architecture, AI governance, privacy, security, coding, and CI/CD standards to data solutions and communicate risks or limitations to leadership
  • Partner with Software engineers, data architects, domain experts, product teams, and platform teams to align data solutions with model requirements, business goals, and enterprise standards
  • Independently deliver assigned work, influence cross-team decisions, promote data engineering best practices, and evaluate technologies that improve AI data delivery
  • The requirements herein describe the general nature and level of work performed by the employee but are not a complete list of responsibilities, duties, and skills required. Other duties may be assigned as needed
Requirements

Education and Experience

  • Undergraduate degree in Computer Science, Data Engineering, Information Systems, or related field
  • 8+ years data engineering, software development, or related experience, including work on data pipelines for AI/ML systems
  • Proficiency with data engineering languages, platforms, and orchestration tools such as Python, SQL, PySpark, Snowflake, Databricks, and Apache Airflow
  • Strong knowledge of relational and NoSQL databases, data warehouses or data lakes, integration platforms such as Azure Data Factory, and streaming technologies such as Kafka
  • Experience with at least one major cloud data ecosystem, preferably Azure, and familiarity with infrastructure-as-code and automated deployment practices
  • Proven ability to design and implement ETL/ELT pipelines, manage databases or data lakes, and integrate large-scale data systems
  • Certifications in cloud data engineering or in data privacy/security and experience with MLOps or AI/ML lifecycle management preferred

Knowledge, Skills and Abilities

  • Must be willing to travel 15-25% (1-2 times per quarter)
  • Deep knowledge of data engineering and AI/ML pipeline practices, including the ability to design and optimize scalable, reliable, and efficient data solutions
  • Strong understanding of data quality, governance, privacy, security, and regulatory requirements
  • Able to independently analyze and troubleshoot complex data and integration issues and apply creative, data-driven solutions when standard approaches are insufficient
  • Able to align technical solutions with business priorities, including speed, cost, risk, reliability, and desired outcomes, and exercise sound judgment in situations with limited precedent
  • Able to communicate complex technical concepts clearly, collaborate across teams, and build stakeholder trust through transparency and reliable delivery
  • Able to lead technical initiatives, document approaches clearly, influence peers, and promote adoption of standards and best practices
  • Must be able to read, write, and speak English

An Equal Opportunity Employer including Disabled/Veterans

Pay Range

$99,000-$139,000/Annually

An Equal Opportunity Employer including Disabled/Veterans

We endeavor to make this site accessible to any and all users. If you would like to contact us regarding the accessibility of our website or need assistance completing the application process, please contact us: careers@dfamilk.com or 877-215-8701.

EEO is the Law

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr. Data Engineer - AI
Sr. Data Engineer - AI

Dairy Farmers of America, Inc. • Kansas City (KS), Northern (KY)

Hybrid
USD 99,000 - 139,000
Data & AI Engineer
Data & AI Engineer

SolomonEdwards • Charlotte (NC)

On-site
USD 120,000 - 180,000
Data Engineer
Data Engineer

Southern Arkansas University • Warner Robins (GA)

Remote
Flexible hours
Weekly bonus of $500–$1000 USD
Work from anywhere
AI Data Engineer
AI Data Engineer

TechDigital Group • Town of Florida (NY)

On-site
USD 180,000 - 240,000
Senior Data Engineer - AI Infrastructure Integration, High Performance Compute
Senior Data Engineer - AI Infrastructure Integration, High Performance Compute

United States Digital Space LLC • Town of Charlotte (NY)

On-site
USD 128,000 - 182,000
Benefits eligible
Discretionary incentive
Data Engineer (AI Pipelines)
Data Engineer (AI Pipelines)

DeWinter Group • Campbell (CA)

Remote
USD 68,880 - 241,080
Sr. Systems Engineer - AI
Sr. Systems Engineer - AI

Dairy-Farmers-of-America,-Inc. • Kansas City (KS)

On-site
USD 98,000 - 139,000
Data Engineer
Data Engineer

Real Chemistry • United States

On-site
USD 90,000 - 120,000
Comprehensive medical, dental, and vision plans
Paid time off
Mental wellness support
+1
Data Engineer - AI
Data Engineer - AI

Compunnel, Inc. • Pennsylvania

On-site
USD 100,000 - 130,000
AI Data Engineer
AI Data Engineer

Green Key Resources • New York (NY)

On-site
USD 150,000 - 190,000