Lead Data Engineer

Talentrabbit

Hyderabad

On-site

INR 1,800,000 - 2,500,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Talentrabbit is seeking an experienced data engineering leader to design and maintain scalable data pipelines using Databricks, PySpark, Delta Lake, and Kafka. You will drive real‑time and batch ingestion, data transformations for digital twins, and governance via Unity Catalog.

The role requires 7+ years of hands‑on experience in data engineering, collaboration with ML teams, and production support across distributed environments. Hyderabad, India based with on‑site work.

Qualifications

  • 7+ years of hands‑on data engineering experience
  • Track record of building and maintaining production‑grade data pipelines
  • Experience with Delta Live Tables for declarative pipeline development
  • Experience working in agile, cross‑functional teams
  • Familiarity with time‑series data patterns and operational data modelling

Responsibilities

  • Design, develop, and maintain scalable data pipelines using Databricks, PySpark, and Delta Lake
  • Build real-time and batch data ingestion pipelines from diverse operational systems using high-performance Kafka data pipelines
  • Implement data transformations that serve digital twin platforms and operational analytics
  • Integrate Kafka event streams with Databricks for real-time operational state updates
  • Implement data quality checks using Delta Live Tables expectations
  • Ensure data governance compliance through Unity Catalog (lineage, access control, metadata)
  • Collaborate with ML engineers to deliver feature-engineered datasets
  • Support production data systems through monitoring, troubleshooting, and incident resolution
  • Build business data warehouse solutions using Terradata for business intelligence

Skills

Databricks
Kafka
PySpark
Delta Lake
Data governance
Delta Live Tables

Tools

Databricks
Kafka
Delta Live Tables
Unity Catalog
Terradata
Databricks SQL
Structured Streaming
Delta Lake

Job description

Role & responsibilities
  • Design, develop, and maintain scalable data pipelines using Databricks, PySpark, and Delta Lake
  • Build real-time and batch data ingestion pipelines from diverse operational systems using high-performance Kafka data pipelines.
  • Implement data transformations that serve digital twin platforms and operational analytics
  • 2+years of Technical Leader ship Experience
  • Integrate Kafka event streams with Databricks for real-time operational state updates
  • Implement data quality checks using Delta Live Tables expectations
  • Ensure data governance compliance through Unity Catalog (lineage, access control, metadata)
  • Optimize pipeline performance, reliability, and cost efficiency
  • Write clean, well-documented, and testable code following engineering best practices
  • Collaborate with ML engineers to deliver feature-engineered datasets
  • Participate in code reviews, knowledge sharing, and continuous improvement initiatives
  • Support production data systems through monitoring, troubleshooting, and incident resolution.
  • Build business data warehouse solutions using Terradata for business intelligence.
Preferred candidate profile
Our core data platform stack includes:

Data Platform & Lakehouse

  • Databricks as the single point of truth for all data
  • Realtime Data Pipelines implemented using Kafka for data ingestion.
  • Databricks SQL for analytical queries
  • Unity Catalog for metadata management and governance
  • Terradata for data warehouse and business intelligence.

Stream & Event Processing

  • Apache Kafka for real-time event ingestion
  • Structured Streaming for continuous data processing
  • Delta Live Tables for declarative, quality-enforced pipelines

Data Quality

  • Delta Live Tables expectations for data validation
  • Data profiling and anomaly detection
Preferred Qualifications
  • 7+ years of hands‑on data engineering experience
  • Track record of building and maintaining production‑grade data pipelines
  • Experience with Delta Live Tables for declarative pipeline development
  • Experience working in agile, cross‑functional teams
  • Familiarity with time‑series data patterns and operational data modelling
Highly Desirable
  • Experience building data pipelines for digital twin or simulation platforms
  • Familiarity with operational state modeling for real‑time systems
  • Exposure to physics‑informed or time‑series ML feature engineering
  • Experience working with distributed, multidisciplinary teams
  • Exposure to industrial domains such as Manufacturing, Logistics, or Transportation is a plus
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Data Engineer, Databricks
Lead Data Engineer, Databricks

Jobtailor • Bengaluru

On-site
INR 900,000 - 1,500,000
Lead Data Engineer
Lead Data Engineer

Gradera • Hyderabad

On-site
INR 3,000,000 - 6,000,000
Lead Data Engineer / Engineering Manager
Lead Data Engineer / Engineering Manager

Gurgaon Hire • Mhalunge

Hybrid
INR 4,000,000 - 6,500,000
Senior Data Engineer - Databricks, PySpark & Lakehouse
Senior Data Engineer - Databricks, PySpark & Lakehouse

Tata Consultancy Services • Bengaluru, Pune District, Chennai District

On-site
INR 1,500,000 - 3,000,000
Lead Software Engineer
Lead Software Engineer

Impetus • Bengaluru

On-site
INR 1,000,000 - 2,000,000
Lead Data Engineer (Databricks & PySpark)
Lead Data Engineer (Databricks & PySpark)

Experis • Pune District

Hybrid
INR 4,000,000 - 7,000,000
Senior Data Engineer
Senior Data Engineer

KSB • Pune District

On-site
INR 800,000 - 1,200,000
Data Engineer
Data Engineer

Tekskills • Chennai District

On-site
INR 2,000,000 - 4,000,000
Senior Engineer - Data Engineering
Senior Engineer - Data Engineering

KSB Company • Maharashtra

On-site
INR 600,000 - 1,000,000
Lead Data Engineer
Lead Data Engineer

PocketFM • Bengaluru

On-site
INR 3,000,000 - 5,400,000
Health insurance
Paid time off
Remote learning budget