Data Engineer

IBM Computing

Town of Yorktown (NY)

On-site

USD 120,000 - 160,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

IBM Quantum in the United States seeks a Data Engineer specializing in Data Integration to design and operate data pipelines powering insights into hardware performance, system reliability, and user workloads. You will collaborate with hardware, software, and analytics teams to build production-ready ETL/ELT workflows and support Lakehouse initiatives with modern tooling.

Ideal candidates have strong SQL, Python, Airflow, Kafka, and Presto experience, plus a track record of delivering reliable

Qualifications

  • Experience building scalable data pipelines for analytics and hardware insights.
  • Experience with Lakehouse architectures and IBM watsonx.data preferred.
  • Familiarity with data governance concepts and lineage is a plus.

Responsibilities

  • Design, build, and maintain scalable data pipelines for analytics and hardware insights.
  • Contribute to IBM Quantum Lakehouse by implementing scalable data connectors.
  • Develop and operate ETL/ELT workflows with focus on data quality and timeliness.
  • Collaborate with cross-functional teams to productionize data solutions.

Skills

ETL development
SQL proficiency
Apache Airflow
Python data tooling
PostgreSQL
Presto/Trino
Apache Kafka
Lakehouse
IBM watsonx.data
Data modeling
Cloud / distributed envs
Git version control

Tools

Apache Airflow
Kafka
Presto/Trino
Spark
IBM watsonx.data

Job description

Introduction

IBM Quantum is building the world’s leading quantum computing systems, software, and cloud services. The Data Engineer in this role will design and operate the data pipelines that power insight into quantum hardware performance, system reliability, user workloads, and platform operations. You will work closely with quantum hardware, firmware, cloud, and product teams to turn diverse technical datasets into trusted analytics assets that guide decision-making across IBM Quantum’s roadmap.

Your role and responsibilities

As a Data Engineer specializing in Data Integration, you will design and build solutions to transfer data from operational and external environments to the business intelligence environment. Your expertise will ensure the seamless flow of data throughout the business intelligence solution’s lifecycle. Your primary responsibilities will include:

  • Design Data Integration Solutions: Create and implement Extract, Transform, and Load (ETL) processes to facilitate data transfer between environments
  • Develop ETL Processes: Build and maintain efficient ETL processes to ensure accurate and timely data flow, adhering to best practices and industry standards.
  • Ensure Seamless Data Flow: Monitor and troubleshoot data integration issues, collaborating with stakeholders to resolve problems and optimize data flow.
  • Optimize Data Integration Solutions: Continuously evaluate and improve data integration solutions, identifying opportunities for process improvements and efficiency gains.
Required technical and professional expertise
  • Design, build, and maintain scalable, reliable data pipelines supporting analytics, operational dashboards, and hardware performance insights for IBM Quantum systems.
  • Contribute towards building IBM Quantum’s Lakehouse by implementing scalable data connectors.
  • Develop and operate ETL/ELT workflows and tooling with a focus on data quality, accuracy, timeliness, and continuous improvement.
  • Apply advanced SQL skills using PostgreSQL and Presto to support analytical workloads, including complex queries and performance tuning.
  • Build and operate orchestration workflows in Apache Airflow, including dependency management, retries, backfills, monitoring, and operational reliability.
  • Implement data transformations and validations using Python (e.g., pandas and related libraries).
  • Support large-scale batch processing for high-volume, heterogeneous datasets, including system telemetry, experiment metadata, cloud operations data, and device performance metrics.
  • Work with streaming platforms such as Apache Kafka or IBM Event Streams to consume event-driven data from distributed quantum systems and services.
  • Apply streaming architecture concepts including topics, partitions, consumer groups, and schema evolution.
  • Integrate multiple technical data sources—quantum hardware telemetry, calibration data, experiment logs, job execution data, user activity, system health metrics—into trusted analytical datasets.
  • Collaborate with quantum hardware, software, product, SRE, and analytics teams to translate requirements into robust, production-ready data solutions.
  • Use Git-based version control, contribute via code reviews, and follow industry-standard software engineering best practices.
Preferred technical and professional experience
  • Experience with Lakehouse solutions and architectures, including IBM watsonx.data
  • Experience with distributed analytics engines such as Presto/Trino, or Apache Spark
  • Familiarity with data modeling techniques for analytical and reliability engineering use cases.
  • Exposure to data governance concepts such as access control, dataset ownership, lineage, and lifecycle management.
  • Experience operating data pipelines in cloud-based or distributed environments (e.g., hybrid cloud, containerized systems).
  • Experience working with hardware telemetry, infrastructure monitoring data, or high-volume operational datasets.
  • Interest in or exposure to quantum computing, advanced hardware systems, cryogenics, or other deep-technology platforms.

IBM is committed to creating a diverse environment and is proud to be an equal-opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, gender, gender identity or expression, sexual orientation, national origin, caste, genetics, pregnancy, disability, neurodivergence, age, veteran status, or other characteristics. IBM is also committed to compliance with all fair employment practices regarding citizenship and immigration status.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Advisory Data Scientist - IBM Quantum
Advisory Data Scientist - IBM Quantum

IBM Computing • Town of Yorktown (NY)

On-site
USD 120,000 - 190,000
Advisory Data Scientist - IBM Quantum
Advisory Data Scientist - IBM Quantum

IBM • Town of Yorktown (NY)

On-site
USD 120,000 - 170,000
Quantum Data Engineer: ETL & Lakehouse Architect
Quantum Data Engineer: ETL & Lakehouse Architect

IBM Computing • Town of Yorktown (NY)

On-site
USD 120,000 - 160,000
Advisory Data Scientist - IBM Quantum
Advisory Data Scientist - IBM Quantum

IBM • San Diego (CA)

On-site
USD 120,000 - 180,000
Advisory Data Scientist - IBM Quantum
Advisory Data Scientist - IBM Quantum

IBM • New York (NY)

On-site
USD 120,000 - 180,000
Quantum Data Analyst Intern 2027
Quantum Data Analyst Intern 2027

IBM • Town of Yorktown (NY)

On-site
USD 52,000 - 70,000
Quantum Data Analyst Intern 2027
Quantum Data Analyst Intern 2027

IBM • San Jose (CA)

On-site
USD 20,000 - 30,000
Quantum Data Analyst Intern 2027
Quantum Data Analyst Intern 2027

IBM Computing • Town of Yorktown (NY)

On-site
USD 30,000 - 47,000
Quantum Algorithm Engineer — AMER
Quantum Algorithm Engineer — AMER

IBM • Town of Yorktown (NY)

On-site
USD 140,000 - 210,000
Research Quantum Computing Systems Engineer Professional Yorktown Heights, US
Research Quantum Computing Systems Engineer Professional Yorktown Heights, US

IBM • Town of Yorktown (NY)

On-site
USD 100,000 - 130,000
Healthcare benefits
401(k) and pension plan
Paid time off
+2