Data Warehouse Specialist

Tata Consultancy Services

Naperville (IL)

On-site

USD 115,000 - 125,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Discretionary Annual Incentive
Comprehensive Medical Coverage

Job summary

Tata Consultancy Services in Naperville, IL, is seeking a Senior Data Engineer to architect and optimize enterprise-grade data pipelines across cloud, on-premises, and hybrid environments. You will leverage Snowflake, Qlik Replicate, DBT Cloud, and Astronomer Airflow to build scalable pipelines, implement schema governance, and drive CI/CD automation with a focus on security and governance.

This role demands strong engineering discipline and a proactive approach to performance tuning and

Qualifications

  • Experience designing scalable data pipelines on Snowflake and cloud platforms.
  • Proficiency with Python-based ingestion frameworks and DBT modeling.
  • Experience with CI/CD workflows and metadata-driven design.

Responsibilities

  • Data Pipeline Architecture & Development
  • Design and implement scalable, resilient data pipelines using Snowflake features including Snowpipe, Tasks, Streams, Dynamic Tables, and advanced SQL.
  • Build and maintain DBT models with strong testing, documentation, and lineage.
  • Develop Python ingestion frameworks for files and APIs, including schema validation, retries, and metadata capture.
  • Engineer ingestion for CSV, fixed width multi record layouts, JSON, XML, Excel, and semi structured formats.
  • Design Mainframe VSAM data ingestion pattern for complex EBCDIC data formats.
  • Schema Drift & Schema Evolution
  • Detect, analyze, and manage schema drift across file, API, and replicated database sources.
  • Implement metadata driven schema evolution strategies to ensure downstream stability.
  • Coordinate schema changes through controlled CI/CD workflows.
  • Database Replication & CDC
  • Configure and manage Qlik Replicate tasks for CDC and full load replication from Oracle, SQL Server, and DB2.
  • Ensure idempotent, auditable, and recoverable replication pipelines with strong monitoring and reconciliation.
  • Data Governance, Security & Tokenization
  • Implement and maintain Snowflake Data Masking policies, including dynamic masking, conditional masking, and role based masking rules.
  • Apply Protegrity tokenization for sensitive data fields across ingestion and transformation
  • Enforce RBAC, data access controls, and governance standards across Snowflake and supporting systems.
  • Orchestration & Automation
  • Build and schedule workflows using Astronomer Airflow, ensuring dependency management, retries, SLAs, and observability.
  • Integrate pipelines with enterprise DevOps processes using GitLab and Azure DevOps for CI/CD automation. Version Control & Code Quality
  • Manage code repositories using GitLab, including branching strategies, merge requests, code reviews, and approvals.
  • Monitoring, Alerting & Performance Optimization
  • Implement monitoring and alerting for ingestion pipelines, schema drift, replication, and transformation workloads.
  • Optimize Snowflake compute, storage, and query performance; scale ingestion pipelines to meet evolving data volume and latency requirements.

Skills

Data architecture
Python scripting
SQL optimization

Tools

Snowflake
Qlik Replicate
DBT Cloud
Astronomer Airflow
Python
Pyspark
GitLab
CI/CD

Job description

Job Description

Must Have Technical/Functional Skills

Job description: We are seeking a highly skilled Senior Data Engineer to architect, build, and optimize enterprise grade data pipelines across cloud, on prem, and hybrid environments. This role requires deep expertise in Qlik Replicate, Snowflake, DBT Cloud, Astronomer Airflow, and Python based ingestion frameworks, with strong engineering discipline around schema governance, DevOps CI/CD, monitoring, and performance optimization. The ideal candidate thrives in complex data ecosystems and brings a strong mindset around automation, metadata driven design, and secure, governed ingestion.

Responsibilities
  • Data Pipeline Architecture & Development
  • Design and implement scalable, resilient data pipelines using Snowflake features including Snowpipe, Tasks, Streams, Dynamic Tables, and advanced SQL.
  • Build and maintain DBT models with strong testing, documentation, and lineage.
  • Develop Python ingestion frameworks for files and APIs, including schema validation, retries, and metadata capture.
  • Engineer ingestion for CSV, fixed width multi record layouts, JSON, XML, Excel, and semi structured formats.
  • Design Mainframe VSAM data ingestion pattern for complex EBCDIC data formats.
  • Schema Drift & Schema Evolution
  • Detect, analyze, and manage schema drift across file, API, and replicated database sources.
  • Implement metadata driven schema evolution strategies to ensure downstream stability.
  • Coordinate schema changes through controlled CI/CD workflows.
  • Database Replication & CDC
  • Configure and manage Qlik Replicate tasks for CDC and full load replication from Oracle, SQL Server, and DB2.
  • Ensure idempotent, auditable, and recoverable replication pipelines with strong monitoring and reconciliation.
  • Data Governance, Security & Tokenization
  • Implement and maintain Snowflake Data Masking policies, including dynamic masking, conditional masking, and role based masking rules.
  • Apply Protegrity tokenization for sensitive data fields across ingestion and transformation

layers.

  • Enforce RBAC, data access controls, and governance standards across Snowflake and supporting systems.

Orchestration & Automation

  • Build and schedule workflows using Astronomer Airflow, ensuring dependency management, retries, SLAs, and observability.
  • Integrate pipelines with enterprise DevOps processes using GitLab and Azure DevOps for CI/CD automation. Version Control & Code Quality
  • Manage code repositories using GitLab, including branching strategies, merge requests, code reviews, and approvals.
  • Monitoring, Alerting & Performance Optimization
  • Implement monitoring and alerting for ingestion pipelines, schema drift, replication, and

transformation workloads.

  • Optimize Snowflake compute, storage, and query performance; scale ingestion pipelines to meet evolving data volume and latency requirements.
Required Skills & Experience
  • Deep expertise with Snowflake, including data masking policies, RBAC, performance tuning, and advanced SQL.
  • Strong experience with Qlik Replicate for CDC and database replication.
  • Excellent proficiency in Python and Pyspark for ingestion frameworks and automation.
  • Hands on experience with DBT Cloud and Astronomer Airflow.
  • Experience with schema drift detection and schema evolution patterns.
  • Experience with GitLab and CI/CD pipelines.
  • Familiarity with Protegrity or similar data protection platforms.
TCS Employee Benefits Summary
  • Discretionary Annual Incentive.
  • Comprehensive Medical Coverage: Medical & Health, Dental & Vision, Disability Planning & Insurance, Pet Insurance Plans.
  • Family Support: Maternal & Parental Leaves.
  • Insurance Options: Auto & Home Insurance, Identity Theft Protection.
  • Convenience & Professional Growth: Commuter Benefits & Certification & Training Reimbursement.
  • Time Off: Vacation, Time Off, Sick Leave & Holidays.
  • Legal & Financial Assistance: Legal Assistance, 401K Plan, Performance Bonus, College Fund, Student Loan Refinancing.

Salary Range: $115,000-$125,000 a year

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer
Senior Data Engineer

Tata Consultancy Services • Raleigh (NC)

On-site
USD 100,000 - 120,000
Data Engineer
Data Engineer

Siri InfoSolutions Inc • Raleigh (NC)

On-site
USD 120,000 - 160,000
Engineer Data Integration
Engineer Data Integration

Tata Consultancy Services • San Antonio (TX)

On-site
USD 64,000 - 85,000
Comprehensive Medical Coverage
401K Plan
Certification & Training Reimbursement
+1
Senior SnowFlake Developer
Senior SnowFlake Developer

Tata Consultancy Services • Charlotte (NC)

On-site
USD 120,000 - 135,000
Discretionary Annual Incentive.
Comprehensive Medical Coverage: Health
Family Support: Parental Leaves
+4
Architect, Business Analytics and Data Intelligence
Architect, Business Analytics and Data Intelligence

PDS Health • Irvine (CA)

On-site
USD 150,000 - 194,000
Medical/Dental/Vision insurance
Paid time off
Tuition reimbursement
+2
Sr. Snowflake Data Engineer
Sr. Snowflake Data Engineer

BravoTECH • Plano (TX)

On-site
USD 120,000 - 150,000
Snowflake Engineer
Snowflake Engineer

Blue Cross of Idaho • Carmel (IN)

Hybrid
USD 96,183 - 144,275
Paid time off
401(k) matching
Health insurance
Senior Snowflake Data Engineer - Talent Pipeline
Senior Snowflake Data Engineer - Talent Pipeline

BlueCloud • United States

On-site
USD 150,000 - 210,000
Senior Data Engineer
Senior Data Engineer

Snowflake • Menlo Park (CA)

Hybrid
USD 110,000 - 140,000
Comprehensive health insurance plans
Health savings accounts
Robust retirement plans
+2
BI / Data Architect
BI / Data Architect

Tata Consultancy Services • Marlborough (MA)

On-site
USD 150,000 - 160,000
Discretionary annual incentive
Comprehensive medical coverage
Parental leave
+4