Senior Databricks Data Engineer: Streaming & Delta Lake

HR POD - Hiring Talent Globally

Lahore

On-site

PKR 300,000 - 600,000

Full time

5 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

HR POD - Hiring Talent Globally seeks a data engineer to own end-to-end ingestion pipelines using Databricks, Delta Lake and Medallion Architecture. You will work with Spark SQL, PySpark, and production-grade streaming data from Azure Event Hubs, Kafka, or Kinesis.

Responsibilities include handling schema drift, data quality, and coordinating with a US-based team to resolve incidents and ensure data correctness across Bronze-Silver-Gold layers.

Qualifications

  • 4+ years of experience in data engineering with production Databricks experience.
  • Strong expertise in Spark SQL, PySpark, Delta Lake, Medallion Architecture and Delta Live Tables (DLT).
  • Hands-on with streaming ingestion using Azure Event Hubs, Kafka or Kinesis.
  • Experience debugging data discrepancies across source-to-warehouse pipelines and handling schema drift.

Responsibilities

  • Build and maintain resilient ingestion pipelines for third-party vendor REST APIs.
  • Work with voice-AI observability and telephony data platforms.
  • Handle multiple pagination schemes and manage rate limits with time-windowing.
  • Implement robust schema-drift handling and alert when fields are renamed or removed.
  • Own streaming ingestion from Azure Event Hubs into Databricks with Structured Streaming/Auto Loader.
  • Manage checkpoints, offsets, watermarking and at-least-once deduplication.
  • Develop Delta Lake pipelines using Medallion Architecture (Bronze-Silver-Gold).
  • Build entity-resolution pipelines for dirty, free-text data.

Skills

SQL
Python
Data pipelines
Data quality

Tools

Databricks
Spark SQL
PySpark
Delta Lake
Delta Live Tables
Azure Event Hubs
Kafka
Kinesis
Holistics AML/AQL
dbt Metrics

Job description

HR POD - Hiring Talent Globally seeks a data engineer to own end-to-end ingestion pipelines using Databricks, Delta Lake and Medallion Architecture. You will work with Spark SQL, PySpark, and production-grade streaming data from Azure Event Hubs, Kafka, or Kinesis.

Responsibilities include handling schema drift, data quality, and coordinating with a US-based team to resolve incidents and ensure data correctness across Bronze-Silver-Gold layers.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer
Senior Data Engineer

Systems Limited • Punjab

On-site
PKR 2,400,000 - 3,600,000
Senior Databricks Engineer
Senior Databricks Engineer

Digifloat • Islamabad

On-site
PKR 2,600,000 - 3,400,000
Senior Data Engineer - Databricks & Streaming - Healthcare AI (Onsite, Evening Shift, Lahore, PKR Salary)
Senior Data Engineer - Databricks & Streaming - Healthcare AI (Onsite, Evening Shift, Lahore, PKR Salary)

HR POD - Hiring Talent Globally • Lahore

On-site
PKR 300,000 - 600,000
Senior Data Engineer — Real-Time Data Platform Lead
Senior Data Engineer — Real-Time Data Platform Lead

Systems Limited • Punjab

On-site
PKR 2,400,000 - 3,600,000
Lead Data Engineer - Real-Time Pipelines & Lakehouse
Lead Data Engineer - Real-Time Pipelines & Lakehouse

Systems Limited • Punjab

On-site
PKR 3,200,000 - 5,200,000
Data Engineer
Data Engineer

Zorba Consulting • Hyderabad City Taluka

On-site
INR 1,200,000 - 2,400,000
Senior Data Engineer (Pyspark, Databricks)
Senior Data Engineer (Pyspark, Databricks)

Strategic Systems International • Lahore

On-site
PKR 1,800,000 - 3,200,000
Data Engineer
Data Engineer

Archisurance • Lahore

On-site
PKR 1,200,000 - 1,800,000
Principal Data Engineer
Principal Data Engineer

Systems Limited • Punjab

On-site
PKR 3,200,000 - 5,200,000
Principal Data Engineer
Principal Data Engineer

Tkxel LLC • Lahore

On-site
PKR 27,777,000 - 38,889,000