Lead Engineer - Data Engineering

Pine Labs

Uttar Pradesh

On-site

INR 3,000,000 - 5,200,000

Full time

7 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Pine Labs in Noida seeks a seasoned Data Engineering Leader with 5–10 years of experience to build and scale data pipelines, data lake platforms, and analytics-ready data for strategic decision-making. You will own the roadmap, mentor engineers, and collaborate with ML, product, and business teams.

You will champion real-time streaming, data modeling standards, and cloud-native deployments while ensuring governance and data quality across the stack.

Qualifications

  • Proven expertise in Kafka/MSK, Debezium, and real-time event-driven architectures.
  • Hands-on experience with Pinot, Redshift, RocksDB or NoSQL DB.
  • Strong background in custom tooling using Java and Python.
  • Experience with Apache Airflow, Superset, Iceberg, Athena, and Glue.
  • Strong AWS ecosystem knowledge (IAM, S3, Lambda, Glue, ECS/EKS, CloudWatch).
  • Deep understanding of data lake architecture, streaming vs batch processing, CDC concepts.
  • Familiarity with modern data formats (Parquet, Avro).

Responsibilities

  • Own the data engineering roadmap – from streaming ingestion to analytics-ready layers.
  • Lead and evolve real-time data pipelines using Kafka/MSK, Debezium CDC, and Pinot.
  • Design and maintain custom-built Java/Python-based frameworks for data ingestion, validation, transformation, and replication.
  • Design and manage data lakes using Iceberg, Glue, and Athena over S3.
  • Define and enforce data modelling standards across Pinot, Redshift and warehouse layers.
  • Lead implementation and scaling of open-source tools and orchestration platforms (Airflow).
  • Drive cloud-native deployment strategies using ECS/EKS, Terraform/CDK, Docker.
  • Collaborate with ML, product, and business teams to support advanced analytics and AI use cases.
  • Mentor data engineers and evangelize modern architecture and engineering best practices.
  • Ensure observability, lineage, data quality, and governance across the data stack.

Skills

Kafka/MSK
Debezium
Real-time systems
Pinot
Redshift
NoSQL/RocksDB
Java/Python
Airflow
Superset
Iceberg
Athena
Glue
AWS ecosystem
Data lake
CDC concepts
Parquet/Avro

Education

Bachelor's or Master's degree in CS/Engineering/related

Tools

Java
Python

Job description

Role Purpose:

We are looking for a highly skilled and motivated Data engineering leader with 5-10 years of experience to join our growing team. You will create data pipelines to move data from source to data lake, Maintain the data lake, real time transform the data to uncover actionable insights and contribute to strategic decision-making across various departments. The ideal candidate will have strong analytical and programming skills, a solid understanding of data pipeline and programming logic, and the ability to collaborate effectively with cross-functional teams. You will be helping all application teams by providing data interface, also will be ensuring that data analytics team is getting required data for their initiatives

The responsibilities we entrust you
  • Own the data engineering roadmap – from streaming ingestion to analytics-ready layers.
  • Lead and evolve our real-time data pipelines using Apache Kafka/MSK, Debezium CDC, and Apache Pinot.
  • Design and maintain custom-built Java/Python-based frameworks for data ingestion, validation, transformation, and replication.
  • Design and manage data lakes using Apache Iceberg, Glue, and Athena over S3.
  • Define and enforce data modelling standards across Pinot, Redshift and warehouse layers.
  • Lead implementation and scaling of open-source tools and orchestration platforms (Airflow).
  • Drive cloud-native deployment strategies using ECS/EKS, Terraform/CDK, Docker.
  • Collaborate with ML, product, and business teams to support advanced analytics and AI use cases.
  • Mentor data engineers and evangelize modern architecture and engineering best practices.
  • Ensure observability, lineage, data quality, and governance across the data stack.
Wha mattters in this role:
Must-Have:
  • Proven expertise in Kafka/MSK, Debezium, and real-time event-driven architectures
  • Hands-on experience with Pinot, Redshift, RocksDB or noSQL DB
  • Strong background in custom tooling using Java and Python
  • Experience with Apache Airflow, Superset, Iceberg, Athena, and Glue
  • Strong AWS ecosystem knowledge (IAM, S3, Lambda, Glue, ECS/EKS, CloudWatch, etc.)
  • Deep understanding of data lake architecture, streaming vs batch processing, CDC concepts
  • Familiarity with modern data formats (Parquet, Avro) and storage abstractions
Nice-to-Have:
  • Exposure to dbt, Trino/Presto, ClickHouse, or Druid
  • Familiarity with data security practices, encryption at rest/in-transit, GDPR/PCI compliance
  • Experience with DevOps practices, GitHub Actions, CI/CD, and Terraform/CloudFormation
Education
  • Bachelor's or Master's degree in Computer Science, Engineering, or a related field.
Communication & Ownership:
  • Strong communication, interpersonal, and conflict resolutio
  • n skillsAble to work both independently and as part of team
Location: Sector 62, Noida
Things you should be comfortable with:
  • Working from office: 5 days a week
  • Pushing the boundaries: Have a big idea? See something that you feel we should do but haven’t done? We will hustle hard to make it happen. We encourage out of the box thinking, and if you bring that with you, we will make sure you get a bag that fits all the energy you bring along.
What we value inour people:
  • You take the shot: You decide fast and deliver right.
  • You are the CEO of what you do: You show ownership and make things happen.
  • You sing your work like an artist: You seek to learn and take pride in thework you do.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Data Engineer
Lead Data Engineer

Terrantic Inc. • India

On-site
INR 7,662,000 - 11,495,000
Competitive salary
Meaningful equity
Flexible remote work
Lead Data Engineer
Lead Data Engineer

MathCo • Bengaluru

On-site
INR 1,500,000 - 2,500,000
Principal Engineer - Data Engineer
Principal Engineer - Data Engineer

Staples India • Chennai District

On-site
INR 3,000,000 - 6,000,000
Lead Data Engineer
Lead Data Engineer

Smart Ims • Bengaluru

Hybrid
INR 1,000,000 - 1,500,000
Opportunity to work with cutting-edge technologies
Collaborate with top-tier engineers
Data Engineering Manager
Data Engineering Manager

Good co India • India

On-site
INR 2,400,000 - 5,400,000
Engineering Manager- Data Platform
Engineering Manager- Data Platform

Meesho • Bengaluru

On-site
INR 2,000,000 - 3,000,000
Opportunity to shape technology and processes
Collaborate with talented engineers
Impact millions of users
Data Engineering Manager
Data Engineering Manager

AiQMEN • Dadri

On-site
INR 600,000 - 1,200,000
Comprehensive benefits
Opportunities for professional development
Team events and social activities
Data Engineer
Data Engineer

xtsworld • Pune District

On-site
INR 1,500,000 - 2,000,000
Free meals
Collaborative work culture
Health benefits
Delivery Lead/Senior Data Engineer 3
Delivery Lead/Senior Data Engineer 3

1203 Barclays Global Serv. Cent • Pune District

On-site
INR 2,500,000 - 3,500,000
Data Architect
Data Architect

Infosys • Bengaluru Urban

On-site
INR 1,200,000 - 2,000,000
Cutting-edge cloud and data technologies
Collaborative culture focusing on innovation
Cross-functional project visibility