Data Engineering Operations Lead

TransUnion

Pune District

On-site

INR 3,500,000 - 7,000,000

Full time

34 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

TransUnion in Pune invites a Data Engineering Operations Lead to drive scalable data pipelines across cloud (GCP) and on-prem environments. You will own ingestion workflows, ensure data quality, and lead a team of engineers to deliver robust data assets.

You will collaborate with data scientists, analysts and product partners to onboard new data sources, meet ingestion SLAs, and continually improve pipeline reliability and observability. This role is hands-on with leadership responsibilities.

Qualifications

  • Bachelor’s degree in CS/Engineering/Statistics or related field required.
  • 10 years data engineering experience with at least 3 years in lead roles.
  • 6+ years of Big Data technologies (Spark, Hive, Hadoop, Databricks).
  • Proven ability to lead, mentor and support junior team members.

Responsibilities

  • Own daily Data Engineering operations including data ingestion workflows.
  • Oversee pipeline health, triage issues, and ensure data quality across ingestion layers.
  • Organize daily workload, assign tasks, track progress, remove blockers.
  • Maintain and improve ingestion pipelines into data lakes, warehouses, and real-time systems.
  • Lead escalation for failures, SLA breaches, and data anomalies with RCA and fixes.
  • Define runbooks, alerting, on-call procedures, and incident management practices.
  • Coordinate with upstream providers and downstream consumers on dependencies.
  • Lead daily stand-ups and mentor junior engineers on best practices.
  • Collaborate with data scientists, analysts, and product teams to onboard data sources.
  • Drive observability, reliability, and efficiency improvements in pipelines.
  • Design and deploy scalable data solutions, including data lakes and warehouses.
  • Oversee end-to-end technical delivery from inception to product.

Skills

Data engineering leadership
Operational management
Mentoring
Cross-functional collaboration
Incident management

Education

Bachelor’s degree in Computer Science, Engineering, Statistics or related field
Google Cloud Professional Data Engineer (desirable)

Tools

Spark
Hive
Databricks
Airflow
Python
SQL (BigQuery)
GCP data stack (BigQuery, Dataflow, Pub/Sub)

Job description

We are looking for a Data Engineering Operations Lead to join our growing Data Engineering and Analytics Practice who will drive building next generation suite of products and platforms by designing, coding, building, and deploying highly scalable and robust data solutions. The team reports to the Head of Data Engineering & Analytics and is part of the wider Global Technology function.

Role Overview and Core Responsibilities

The role exists to own and manage the internal function responsible for development, support and operation of our ingestion pipelines as well as the assets created and owned by the team across our cloud (GCP) and on-premises environments. As Data Engineering Operations lead, you will be responsible to ensure the quality and robustness of the developed pipelines, operational resilience, SLA compliance for all of our main data assets created by our team. As a lead engineer most of your time will be spent on hands‑on engineering work with expected ~25% spent on operational activities.

Responsibilities include:
  • Own and manage daily Data Engineering operations, including data ingestion workflows, pipeline monitoring, and incident resolution.
  • Oversee end-to-end pipeline health - proactively monitor, triage, and resolve failures, bottlenecks, and data quality issues across all ingestion layers.
  • Organize and prioritize the engineering operations function’s daily workload - assign tasks, track progress, and remove blockers to ensure smooth operational delivery.
  • Maintain and improve ingestion pipelines from various data sources into data lakes, warehouses, and real‑time streaming systems.
  • Act as the first point of escalation for pipeline failures, SLA breaches, and data anomalies - driving root cause analysis and permanent fixes.
  • Define and enforce operational standards - runbooks, alerting thresholds, on‑call procedures, and incident management practices.
  • Coordinate with upstream data providers and downstream consumers to manage dependencies and communicate pipeline status.
  • Lead daily stand‑ups and work organization for the function under the wider Data Engineering practice.
  • Mentor and guide junior engineers on operational best practices, debugging, and pipeline development.
  • Collaborate with data scientists, analysts, and product partners to onboard new data sources and meet ingestion SLAs.
  • Drive continuous improvement in pipeline reliability, observability, and efficiency.
  • Implement best practices in data governance, data quality monitoring, and compliance across all pipelines.
  • Design, build, test, and deploy Data solutions at scale, including data lakes, data warehouses, and real‑time analytics.
  • Lead technical delivery on use cases, plan and delegate tasks to junior team members, and oversee work from inception to final product.
Required Knowledge and Experiences
  • Bachelor’s degree in Computer Science, Engineering, Statistics or a related field
  • 10 years of data engineering experience with at least 3 years in lead roles.
  • 6+ years of experience in Big Data technologies (e.g., Spark, Hive, Hadoop, Databricks).
  • Excellent knowledge of data engineering concepts and best practices.
  • Proven experience leading, mentoring and supporting junior team members.
  • Ability to lead technical deliverables autonomously and guide junior data engineers.
  • Ability to organize and manage daily engineering workload - task assignment, prioritization, and delivery tracking.
  • Strong attention to detail and adherence to best practices and defined policies.
Essential Technical Skills:
  • Advanced proficiency with Apache Spark (PySpark) including tuning and performance optimisation experience.
  • Proficiency in Python, Pandas (Scala/Java knowledge is desirable).
  • Working knowledge of Apache Hive.
  • Strong SQL knowledge and experience (T-SQL, working with SQL Server, SSMS, GCP BigQuery).
  • Expertise in designing and implementing scalable data pipelines and ETL processes using the GCP data stack, including BigQuery, Dataflow, Pub/Sub, Cloud Storage, Cloud Composer, Cloud Functions, Dataproc (Spark).
  • Experience with batch, real‑time streaming, and ETL processes, including incident resolution and pipeline recovery.
  • Experience building and managing ETL workflows using Apache Airflow, including DAG creation, scheduling, and error handling.
  • Source control with Git.
  • Knowledge of CI/CD concepts and experience designing CI/CD for data pipelines.
  • Knowledge of Delta Lake concepts and common data formats, Lakehouse architecture.
  • Software engineering principles including OOP, design patterns, SDLC, Agile, TDD, and performance optimization.
Desirable Technical Skills:
  • Experience designing logical data models and physical data models, including data warehouse and data mart designs.
  • Relevant certifications (e.g. Google Cloud Professional Data Engineer).
  • Experience with streaming services such as Kafka is a plus.
  • R & Sparklyr experience is a plus.
  • Knowledge of MLOps concepts, AI/ML lifecycle management, and MLflow.
  • Harness experience is a plus.

We’re also looking for the preferred skills below. Whether you are proficient or could use

some brushing up, we’re happy to support your career development and growth in:

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineering Manager
Data Engineering Manager

Good co India • India

On-site
INR 2,400,000 - 5,400,000
Lead Data Engineer (Databricks, PySpark & GCP)
Lead Data Engineer (Databricks, PySpark & GCP)

Egen • Hyderabad

On-site
INR 5,500,000 - 7,500,000
Healthcare benefits
Performance bonus
Technical Lead - Data Engineer (Data&AI)
Technical Lead - Data Engineer (Data&AI)

Srijan: Now Material • Gurugram District

On-site
INR 3,000,000 - 6,000,000
Technical Lead - Data Engineer (Data&AI)
Technical Lead - Data Engineer (Data&AI)

Srijan Technologies PVT LTD • Gurugram District

On-site
INR 4,000,000 - 7,500,000
Group Data Engineer I
Group Data Engineer I

The Peninsular and Oriental Steam Navigation Company • Bengaluru

On-site
INR 2,500,000 - 4,000,000
Data Engineer Lead
Data Engineer Lead

Nexus Corporation • Hyderabad

On-site
INR 3,500,000 - 7,000,000
Senior Data Engineer (GCP)
Senior Data Engineer (GCP)

Publicis Production • Bengaluru

On-site
INR 1,800,000 - 2,400,000
Senior Data Engineer
Senior Data Engineer

Proclink • Gandhamguda

On-site
INR 800,000 - 1,500,000
Senior AI Engineer
Senior AI Engineer

Latent View Analytics Limited • Chennai District

On-site
INR 3,000,000 - 4,200,000
DevOps Lead (GCP, AWS) – Data Engineering
DevOps Lead (GCP, AWS) – Data Engineering

Harmony Data Integration Technologies • India

Hybrid
INR 2,500,000 - 4,000,000