Senior Data Engineer(Databricks)

Recruitzz

Lahore

On-site

PKR 3,200,000 - 5,600,000

Full time

3 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Recruitzz in Lahore, Pakistan, is seeking a Senior Data Engineer to own end-to-end data pipelines using Databricks and Python, on-site near FC College Lahore.

The role requires 5+ years in data engineering with production experience in Spark SQL, PySpark, Delta Lake, and medallion architecture; you will work on streaming data ingestion, API integration, and data matching, collaborating with a US team.

Experience with a semantic layer and strong English communication is essential.

Qualifications

  • 5+ years in data engineering with substantial production Databricks experience: Spark SQL, PySpark, Delta Lake, medallion architecture, and DLT.
  • Structured Streaming or equivalent streaming ingestion in production (Event Hubs, Kafka, or Kinesis sources), including checkpoint recovery, watermarking, and dedup strategies.
  • Demonstrated experience debugging source-vs-warehouse discrepancies — you can walk through a real incident where counts didn't match and explain how you found the cause.
  • Production integration with third-party REST APIs, including pagination edge cases, rate/row limits, retries, and schema-drift handling.
  • Entity resolution or data-matching work on messy real-world text data.
  • Experience with a metrics/semantic layer (Holistics AML/AQL, dbt metrics, or LookML) and a working understanding of why non-additive measures can't be computed from pre-aggregated rollups.
  • Strong SQL and Python; comfortable owning pipelines end-to-end without heavy oversight.
  • Written and spoken English strong enough for async collaboration with a US team.

Responsibilities

Skills

Databricks
Spark SQL
PySpark
Delta Lake
Medallion architecture
DLT
Structured Streaming
Streaming ingestion
Event Hubs
Kafka
Kinesis
Checkpointing
Watermarking
Deduplication
Source-vs-Warehouse
REST APIs
Pagination
Rate limits
Retries
Schema drift
Entity resolution
Data matching
Holistics AML/AQL
dbt metrics
LookML
SQL
Python
End-to-end pipelines
English

Tools

dbt
LookML

Job description

Location: Near FC College Lahore (On-site)

Must-have experience

  • 5+ years in data engineering with substantial production Databricks experience: Spark SQL, PySpark, Delta Lake, medallion architecture, and DLT.
  • Structured Streaming or equivalent streaming ingestion in production (Event Hubs, Kafka, or Kinesis sources), including checkpoint recovery, watermarking, and dedup strategies.
  • Demonstrated experience debugging source-vs-warehouse discrepancies — you can walk through a real incident where counts didn't match and explain how you found the cause.
  • Production integration with third-party REST APIs, including pagination edge cases, rate/row limits, retries, and schema-drift handling.
  • Entity resolution or data-matching work on messy real-world text data.
  • Experience with a metrics/semantic layer (Holistics AML/AQL, dbt metrics, or LookML) and a working understanding of why non-additive measures can't be computed from pre-aggregated rollups.
  • Strong SQL and Python; comfortable owning pipelines end-to-end without heavy oversight.
  • Written and spoken English strong enough for async collaboration with a US team.

Nice to have

  • Healthcare data experience: referrals, payer taxonomy, claims/eligibility, or other PHI-adjacent datasets; familiarity with HIPAA handling expectations.
  • Voice-agent, call-center, or telephony/conversation data (call transcripts, containment/outcome metrics).
  • Holistics specifically (AML/AQL modeling).
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Data Engineer - Databricks & Streaming - Healthcare AI (Onsite, Evening Shift, Lahore, PKR Salary)
Senior Data Engineer - Databricks & Streaming - Healthcare AI (Onsite, Evening Shift, Lahore, PKR Salary)

HR POD - Hiring Talent Globally • Lahore

On-site
PKR 300,000 - 600,000
Senior Databricks Engineer
Senior Databricks Engineer

Digifloat • Islamabad

On-site
PKR 2,600,000 - 3,400,000
Senior Data Engineer - Real-Time Spark/PySpark Pipelines
Senior Data Engineer - Real-Time Spark/PySpark Pipelines

Recruitzz • Lahore

On-site
PKR 3,200,000 - 5,600,000
Principal Senior Data Engineer Databricks
Principal Senior Data Engineer Databricks

Crescendo Global Leadership Hiring India • Hyderabad City Taluka

Hybrid
PKR 2,500,000 - 4,500,000
Data Engineer
Data Engineer

Zorba Consulting • Hyderabad City Taluka

On-site
INR 1,200,000 - 2,400,000
Data Architect - Microsoft Azure Data Services, DataLake, Databricks
Data Architect - Microsoft Azure Data Services, DataLake, Databricks

HireOn • Pakistan

Hybrid
PKR 3,000,000 - 5,500,000
Data Management & Analytics Specialist
Data Management & Analytics Specialist

NorthBay Solutions • Islamabad

On-site
PKR 1,800,000 - 2,800,000
Hybrid work model
Data Management & Analytics Specialist
Data Management & Analytics Specialist

NorthBay Solutions • Karachi Division

On-site
PKR 2,200,000 - 3,200,000
Hybrid work model
International client exposure
Data Management & Analytics Specialist
Data Management & Analytics Specialist

NorthBay Solutions LLC • Lahore

On-site
PKR 2,400,000 - 4,000,000
Data Management & Analytics Specialist
Data Management & Analytics Specialist

NorthBay Solutions LLC • Pakistan

On-site
PKR 2,500,000 - 4,700,000