Data Engineer

Baker Group

Ankeny (IA)

On-site

USD 90,000 - 150,000

Full time

2 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Baker Group is seeking a Data Engineer in Ankeny, IA to design, build, and maintain data pipelines powering its Microsoft Fabric data warehouse. You will own ingestion, transformation, and orchestration from ERP, HRIS, MRP and other sources into governed, reliable data products for analysts and developers.

You'll implement data models, governance, and scalable architectures, collaborating with Data Scientists, Analysts, and business system owners to enable reporting, BI, and AI/ML initiatives

Qualifications

  • Bachelor's degree in a quantitative field required.
  • 3–5 years of data engineering, ETL/ELT experience.
  • Proficiency with SQL and data transformation.
  • Experience with Microsoft Fabric and Azure Data Factory.
  • Experience with medallion architecture and modern data warehousing.
  • Knowledge of data governance and metadata management; AI/ML data prep a plus.

Responsibilities

  • Designs, builds, and maintains ETL/ELT pipelines ingesting data from enterprise systems into Microsoft Fabric.
  • Architects the Fabric medallion Lakehouse structure as the single source of truth.
  • Owns pipeline orchestration, scheduling, and monitoring for data availability.
  • Maintains core datasets across employee, finance, project, service, and manufacturing domains.
  • Establishes data quality, validation, and reconciliation across pipelines.
  • Defines data ontologies and canonical definitions for consistency.
  • Prepares data for AI/ML use cases, including vector embeddings and RAG pipelines.
  • Collaborates with Data Analysts, Data Scientists, and business system owners.

Skills

SQL
Python
PySpark
T-SQL
Data Modeling
Data Governance
Microsoft Fabric
Azure Data Factory

Education

Bachelor's degree in Computer Science, Data Engineering, Information Systems, or other relevant quantitative field

Tools

Microsoft Fabric
Azure Data Factory

Job description

If you are unable to complete this application due to a disability, contact this employer to ask for an accommodation or an alternative application process.

Full Time Professionals Ankeny, IA, US

PURPOSE

The Data Engineer is responsible for designing, building, and maintaining the data pipelines and infrastructure that power Baker Group's Microsoft Fabric data warehouse, serving as the organization's single source of truth. This role owns the ingestion, transformation, and orchestration of data from disparate internal systems (ERP, HRIS, MRP and other structured data sources) into governed, reliable data products used by Data Analysts and developers to deliver insights to executive and operational teams and ensures that data is structured to support both traditional reporting and emerging AI and machine learning use cases. The Data Engineer curates and maintains core datasets spanning employees, finance, construction and manufacturing projects, and service, and partners with the Data Scientist, Data Analyst, and Software Development roles to ensure data is trustworthy, well-structured, and fit for downstream use.

ESSENTIAL FUNCTIONS AND RESPONSIBILITIES

The following duties are typical for this job. These are not to be constructed as exclusive or all inclusive. Other duties may be required and assigned.

  • Designs, builds, and maintains ETL/ELT pipelines that ingest data from enterprise systems into Microsoft Fabric.
  • Architects and maintains the Fabric medallion Lakehouse structure (bronze, silver, gold layers) as Baker Group's single source of truth.
  • Develop and implement best practices for the data infrastructure and environment (e.g. Development/Test/Production environments, Git for version control).
  • Owns pipeline orchestration, scheduling, and monitoring to ensure reliable, timely, and accurate data availability.
  • Curates and maintains core datasets across employee, finance, project, service, and manufacturing domains.
  • Establishes and enforces data quality, validation, and reconciliation processes across all pipelines.
  • Designs and manages data models, schemas, and semantic layers that support Data Analyst reporting and Data Scientist modeling work.
  • Defines and maintains data ontologies and canonical business definitions (for example, what constitutes a "project," "employee," or "cost code") to ensure consistent meaning across systems and consumers.
  • Prepares and structures data to support AI and machine learning use cases, including feature‑ready datasets, retrieval‑augmented generation (RAG) pipelines, and vector embedding storage.
  • Manages Fabric capacity planning, workspace organization, and performance optimization.
  • Implements data governance practices, including access controls, lineage tracking, and metadata management, consistent with Baker Group's data classification standards.
  • Partners with business system owners (ERP, HRIS, MRP, etc.) to understand upstream data structures and manage change impacts.
  • Collaborates with the Data Scientist to ensure pipeline outputs support analytical and machine learning use cases.
  • Collaborates with Data Analysts to ensure data products support paginated reporting, dashboards, and self‑service BI needs.
  • Collaborates with Software Development and DevOps Teams to ensure data products support application development needs.
  • Coordinates with 3rd party consultants when necessary to deliver data engineering projects and augment capacity for demanding business needs.
  • Develops and maintains documentation for pipelines, schemas, and integration logic.
  • Troubleshoots and resolves pipeline failures, latency issues, and data quality incidents.
  • Monitors and maintains data‑specific infrastructure, including Fabric capacity, pipeline orchestration tools, and monitoring/alerting systems.
  • Evaluates and recommends new data engineering tools, patterns, and best practices.
  • Stays current on emerging trends in data engineering, cloud data platforms, and integration techniques.
MINIMUM EDUCATION and EXPERIENCE REQUIRED TO PERFORM ESSENTIAL FUNCTIONS
  • Bachelor's degree in Computer Science, Data Engineering, Information Systems, or other relevant quantitative field
  • Three to five years of experience in data engineering, ETL/ELT development, or a related field
  • Proficiency with SQL and database technologies for data extraction, transformation, and loading
  • Experience with Microsoft Fabric, Azure Data Factory, or similar cloud ETL/orchestration tools
  • Experience with medallion architecture and modern data warehousing patterns
  • Experience with a programming language such as Python, PySpark, or T‑SQL for data transformation
  • Familiarity with data modeling techniques (dimensional modeling, star schema)
  • Understanding of data governance, data quality, and metadata management practices
  • Experience preparing data for AI/ML consumption (e.g., vector embeddings, RAG architectures) is a plus
  • Business acumen and understanding of construction or related industries is a plus
CERTIFICATES, LICENSES, REGISTRATIONS
  • No specific requirements; however, relevant certifications such as Microsoft Certified: Fabric Data Engineer Associate, Azure Data Engineer Associate, or similar cloud platform certifications are a plus
MENTAL AND PHYSICAL COMPETENCIES REQUIRED TO PERFORM ESSENTIAL FUNCTIONS
  • Strong analytical and troubleshooting skills with the ability to diagnose and resolve complex pipeline and data quality issues
  • Excellent time and project management skills with the ability to prioritize across multiple pipeline and infrastructure projects
  • Current with industry trends in data engineering, cloud platforms, and integration best practices
  • Strong communication skills with the ability to translate technical data structures for non‑technical stakeholders
  • Team player with strong collaboration skills, particularly with the Data Scientist, Data Analysts, and business system owners
  • Must be able to focus on complex technical problems and work independently with minimal supervision
  • Ability to work in a fast‑paced environment and adapt to changing business priorities
  • Meticulous attention to detail and commitment to producing reliable, well‑documented data infrastructure
ENVIRONMENTAL ADAPTABILITY
  • Prolonged periods of sitting at a desk and working on a computer
  • Must be able to lift 10 pounds occasionally
  • May have occasional visits to a job site which would require periods of standing, walking and/or climbing stairs
EQUIPMENT/TOOLS
  • Use a computer for 8 hours a day

Baker Group is an Equal Opportunity Employer. In compliance with the Americans with Disabilities Act, Baker Group will consider reasonable accommodations for qualified individuals with disabilities and encourage prospective employees and incumbents to discuss potential accommodations with the Employer.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer
Data Engineer

Baker Group • Des Moines (IA)

On-site
USD 95,000 - 135,000
Data Engineer
Data Engineer

Sakataornamentals • Burlington (WA)

On-site
USD 110,000 - 125,000
Health insurance
401(k) Program + Company Match
Holiday Bonus
+2
Data Engineer
Data Engineer

Sakata Seed America, Inc. • Woodland (CA)

On-site
USD 110,000 - 155,000
Medical, Dental & Vision Insurance
401(k) with Company Match
Paid Vacation & Holidays
Data Engineer
Data Engineer

Ohio Cat • Broadview Heights (OH)

On-site
USD 90,000 - 130,000
401(k) Match
Health Insurance (HSA)
Dental & Vision
+5
Analytics Engineer
Analytics Engineer

Baker Botts • Austin (TX)

Hybrid
USD 120,000 - 180,000
Data Engineer/Data Analyst
Data Engineer/Data Analyst

Midland Industries • Kansas City (MO)

On-site
USD 90,000 - 130,000
Analytics Engineer
Analytics Engineer

Baker Botts LLP • Austin (TX), Northern (KY)

Hybrid
USD 110,000 - 150,000
Hybrid work arrangement
Fabric Data Engineer
Fabric Data Engineer

Lasting Change Inc • Fort Wayne (IN)

Hybrid
USD 90,000 - 130,000
Sr. Data Engineer
Sr. Data Engineer

Save-A-Lot, Ltd. • Missouri

On-site
USD 100,000 - 140,000
401K match up to 4%
Paid Time Off
Medical Insurance options including FS
+5
Data Engineer: Build AI-Ready Data Pipelines & Lakehouse
Data Engineer: Build AI-Ready Data Pipelines & Lakehouse

Baker Group • Ankeny (IA)

On-site
USD 90,000 - 150,000