Senior Data Engineer - AI & Analytics Infrastructure

IBM

New York (NY)

On-site

USD 120,000 - 190,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

IBM Consulting is seeking an experienced Data Engineer to support design and scaling of data pipelines and infrastructure for a high-priority Agentic AI engagement. You will work with AI architects to ensure the right data reaches the right systems in the right form, leveraging modern platforms for reliable pipelines and governance.

Responsibilities include building scalable data architectures, implementing data quality checks, and collaborating with AI teams to support analytics and AI/ML use

Qualifications

  • 7+ years of experience designing, developing, and maintaining scalable data pipelines across Azure and AWS.
  • Build and optimize enterprise data platforms using Azure Data Factory, Azure Data Lake, AWS S3, AWS Glue, Databricks, and Snowflake.
  • Develop robust ETL/ELT frameworks for analytics, reporting, and AI/ML use cases in cloud/hybrid environments.
  • Implement scalable ingestion and transformation pipelines for structured, semi-structured, and unstructured data sources.
  • Support data industrialization through reusable pipelines, observability, CI/CD, and governance practices.

Responsibilities

  • Data Pipeline Design & Development: design, build, and maintain robust pipelines for ingestion, transformation, and delivery of high-quality data across the platform.
  • Data Infrastructure & Architecture: architect and maintain data lakehouse patterns and layered data models for AI and analytics.
  • Data Quality & Governance: implement quality checks, lineage, cataloging, and access controls to ensure trustworthy data outputs.

Skills

Azure Data Factory
Databricks
Snowflake
ETL/ELT design
CI/CD
Data governance

Education

Master’s Degree

Tools

Azure Data Lake
AWS S3
Azure Synapse Analytics
Databricks

Job description

Introduction

A career in IBM Consulting is built on long-term client relationships and close collaboration worldwide. You’ll work with leading companies across industries, helping them shape their hybrid cloud and AI journeys. With support from our strategic partners, robust IBM technology, and Red Hat, you’ll have the tools to drive meaningful change and accelerate client impact. At IBM Consulting, curiosity fuels success. You’ll be encouraged to challenge the norm, explore new ideas, and create innovative solutions that deliver real results. Our culture of growth and empathy focuses on your long-term career development while valuing your unique skills and experiences.

Your Role And Responsibilities

We are seeking an experienced Data Engineer to support the design and scaling of data pipelines and infrastructure for a high-priority Agentic AI engagement. This role is central to the success of the program - the quality, accessibility, and governance of data directly enables the AI and analytics use cases being built.

You will work alongside AI architects and engineers to ensure that the right data reaches the right systems in the right form. The client is looking for someone with strong hands-on experience across modern data platforms who can operate with confidence and deliver at pace.

What You’ll Do
Data Pipeline Design & Development
  • Design, build, and maintain robust data pipelines that ingest, transform, and deliver high-quality data across the platform
  • Develop scalable architectures using Microsoft Fabric, Databricks, and/or Azure Synapse Analytics
  • Ensure pipelines are performant, reliable, and built to handle the scale and variability of enterprise data
  • Implement data transformation and orchestration workflows that feed AI models and analytics dashboards
Data Infrastructure & Architecture
  • Architect and maintain the underlying data infrastructure that supports AI and analytics use cases
  • Define and implement data lakehouse patterns, medallion architecture, and layered data models
  • Collaborate with AI engineers and architects to ensure data outputs are structured and accessible for model consumption
  • Manage and optimize data storage, compute, and processing environments for cost and performance
Data Quality & Governance
  • Implement data quality checks, validation frameworks, and monitoring to ensure trustworthy data outputs
  • Establish and enforce data governance standards including lineage tracking, cataloging, and access controls
  • Partner with stakeholders to document data assets and ensure discoverability across the platform
Preferred Education

Master’s Degree

Required Technical And Professional Expertise
  • 7+ years of experience designing, developing, and maintaining scalable batch and real-time data pipelines across Azure and AWS.
  • Build and optimize enterprise data platforms leveraging services such as Azure Data Factory, Azure Data Lake, AWS S3, AWS Glue, Databricks, and Snowflake.
  • Develop robust ETL/ELT frameworks supporting analytics, reporting, operational, and AI/ML use cases across cloud and hybrid ecosystems.
  • Implement scalable ingestion and transformation pipelines for structured, semi-structured, and unstructured enterprise data sources.
  • Support data industrialization efforts through reusable pipeline frameworks, standardized engineering practices, observability, monitoring, automated testing, and CI/CD deployment patterns.
  • Enable trusted enterprise data foundations by implementing data quality controls, metadata management, lineage, cataloging, and governance capabilities.
  • Optimize data models, distributed processing workloads, storage strategies, and query performance within Databricks and Snowflake environments.
  • Integrate enterprise applications, APIs, ERP systems, CRM platforms, and event-driven architectures into centralized cloud data platforms.
  • Collaborate with AI engineers, architects, analysts, and business stakeholders to support analytics, AI, and generative AI initiatives.
  • Support Infrastructure-as-Code, cloud-native deployment practices, and secure enterprise data operations across Azure and AWS platforms.
Preferred Skills
  • Familiarity with Azure Data Factory, Event Hubs, or other Azure data integration services
  • Experience implementing data governance frameworks and working with data cataloging tools
  • Knowledge of MLOps data pipelines and feature engineering for AI model consumption
  • Background supporting Agentic AI or generative AI programs where data quality is mission-critical
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer - AI & Analytics Infrastructure
Senior Data Engineer - AI & Analytics Infrastructure

IBM • Chicago (IL)

On-site
USD 120,000 - 180,000
Senior Data Engineer - AI & Analytics Infrastructure
Senior Data Engineer - AI & Analytics Infrastructure

IBM • Dallas (TX)

On-site
USD 110,000 - 160,000
Senior Data Engineer: AI-Driven Data Pipelines
Senior Data Engineer: AI-Driven Data Pipelines

IBM • New York (NY)

On-site
USD 120,000 - 190,000
Senior Data Engineer - AI-Ready Data Pipelines
Senior Data Engineer - AI-Ready Data Pipelines

IBM • Dallas (TX)

On-site
USD 110,000 - 160,000
Senior Data Engineer - AI-Ready Data Pipelines & Lakehouse
Senior Data Engineer - AI-Ready Data Pipelines & Lakehouse

IBM • Chicago (IL)

On-site
USD 120,000 - 180,000
Associate Data Engineer 2026 – Data Services
Associate Data Engineer 2026 – Data Services

IBM • Atlanta (GA)

On-site
USD 70,000 - 90,000
Senior AI Solution Architect
Senior AI Solution Architect

IBM • New York (NY)

Remote
USD 180,000 - 240,000
Senior AI Architect – Azure & Cloud AI
Senior AI Architect – Azure & Cloud AI

IBM • Dallas (TX)

On-site
USD 180,000 - 240,000
Senior AI Architect – Azure & Cloud AI
Senior AI Architect – Azure & Cloud AI

IBM • Chicago (IL)

On-site
USD 150,000 - 210,000
Senior AI Architect – Azure & Cloud AI
Senior AI Architect – Azure & Cloud AI

IBM • New York (NY)

On-site
USD 140,000 - 190,000