Lead Data Engineer

Hilabs

Pune District

On-site

INR 2,500,000 - 4,000,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Transport facility
Complimentary meals

Job summary

HiLabs, an AI-driven healthcare technology company, is hiring a Senior/Lead Data Engineer to design and scale modern data platforms powering AI and analytics solutions.

You will create high-performance data pipelines using Python and Spark, and work with AWS/GCP/Azure to deliver cloud-native data solutions across data ingestion, warehousing, and governance. Strong teamwork with data scientists and software engineers is required.

Qualifications

  • 5 to 10 years of hands-on experience in Data Engineering.
  • Strong experience building production-grade ETL and ELT pipelines.
  • Excellent programming skills in Python. Experience with Scala or Java is a plus.
  • Strong SQL skills and experience with relational and NoSQL databases.
  • Hands-on experience with Apache Spark (PySpark or Scala Spark).
  • Experience with workflow orchestration tools such as Apache Airflow.
  • Experience with cloud platforms such as AWS, GCP, or Azure.
  • Good understanding of cloud data services such as S3, Redshift, BigQuery, EMR, Glue, or equivalent.
  • Experience with data warehousing solutions like Snowflake, Redshift, or BigQuery.
  • Familiarity with Git, CI/CD pipelines, and Agile development practices.

Responsibilities

  • Design, develop, and maintain scalable ETL and ELT pipelines for large-scale data processing.
  • Build reliable and efficient data ingestion, transformation, and orchestration workflows.
  • Develop high-performance data pipelines using Python and Apache Spark.
  • Optimize SQL and NoSQL databases for performance, scalability, and reliability.
  • Work with cloud platforms such as AWS, GCP, or Azure to build cloud-native data solutions.
  • Build and optimize data warehouse solutions using Snowflake, Redshift, BigQuery, or similar technologies.
  • Implement data quality checks, validation frameworks, monitoring, and alerting.
  • Collaborate with cross-functional teams to support AI, analytics, and business reporting requirements.
  • Improve pipeline performance, scalability, and operational efficiency.
  • Follow best practices for data governance, security, version control, and CI/CD.

Skills

Python
Scala/Java
SQL
Apache Spark
Apache Airflow
AWS/GCP/Azure
Data warehousing (Snowflake/Redshift/\

Tools

Apache Spark
Apache Airflow
Git
CI/CD
Snowflake
Redshift
BigQuery
Hadoop
Kafka

Job description

HiLabs is an AI-driven healthcare technology company solving one of the biggest challenges in the US healthcare industry, dirty data. Our AI-powered platform processes billions of healthcare records and helps payers, providers, and life sciences organizations unlock accurate insights at scale.

At HiLabs, we combine Artificial Intelligence, Data Engineering, and Healthcare expertise to build innovative products that transform healthcare data quality and enable better business and patient outcomes.

Role Overview

We are looking for a passionate and experienced Senior/Lead Data Engineer to build and scale modern data platforms that power AI and analytics solutions. You will work on designing high-performance data pipelines, processing large-scale datasets, and building reliable cloud-native data solutions.

You will collaborate closely with Data Scientists, Product Managers, and Software Engineers to deliver scalable and production-ready data engineering solutions.

Key Responsibilities
  • Design, develop, and maintain scalable ETL and ELT pipelines for large-scale data processing.
  • Build reliable and efficient data ingestion, transformation, and orchestration workflows.
  • Develop high-performance data pipelines using Python and Apache Spark.
  • Optimize SQL and NoSQL databases for performance, scalability, and reliability.
  • Work with cloud platforms such as AWS, GCP, or Azure to build cloud-native data solutions.
  • Build and optimize data warehouse solutions using Snowflake, Redshift, BigQuery, or similar technologies.
  • Implement data quality checks, validation frameworks, monitoring, and alerting.
  • Collaborate with cross-functional teams to support AI, analytics, and business reporting requirements.
  • Improve pipeline performance, scalability, and operational efficiency.
  • Follow best practices for data governance, security, version control, and CI/CD.
Required Skills
  • 5 to 10 years of hands-on experience in Data Engineering.
  • Strong experience building production-grade ETL and ELT pipelines.
  • Excellent programming skills in Python. Experience with Scala or Java is a plus.
  • Strong SQL skills and experience with relational and NoSQL databases.
  • Hands-on experience with Apache Spark (PySpark or Scala Spark).
  • Experience with workflow orchestration tools such as Apache Airflow.
  • Experience with cloud platforms such as AWS, GCP, or Azure.
  • Good understanding of cloud data services such as S3, Redshift, BigQuery, EMR, Glue, or equivalent.
  • Experience with data warehousing solutions like Snowflake, Redshift, or BigQuery.
  • Familiarity with Git, CI/CD pipelines, and Agile development practices.
  • Strong analytical, debugging, and problem-solving skills.
Preferred Skills
  • Experience with Hadoop, Kafka, or other distributed data technologies.
  • Exposure to MLOps or AI data pipelines.
  • Experience working in a product-based or SaaS company.
  • Healthcare or healthcare analytics experience is an added advantage.
Why Join HiLabs?
  • Build AI-powered healthcare products that process billions of healthcare records.
  • Work on large-scale data engineering challenges using modern cloud technologies.
  • Collaborate with highly talented engineers, data scientists, and AI researchers.
  • Fast-paced product engineering environment with significant ownership and learning opportunities.
  • Competitive compensation, career growth, transport facility, and complimentary meals.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Data Engineer
Lead Data Engineer

V2 Solutions • Pune District

On-site
INR 2,500,000 - 4,200,000
Competitive compensation
Transport facility
Complimentary meals
Data Engineer II /Sr Data Engineer
Data Engineer II /Sr Data Engineer

HiLabs Inc. • Bengaluru

On-site
INR 1,800,000 - 2,800,000
ESOPs
Medical coverage
Professional development
+2
Lead Data Scientist
Lead Data Scientist

HiLabs Inc. • Pune District

On-site
INR 3,800,000 - 6,200,000
H1B sponsorship
ESOPs
Medical coverage
+3
Senior Data Scientist
Senior Data Scientist

HiLabs Inc. • Pune District

On-site
INR 1,800,000 - 2,400,000
Competitive salary
ESOPs
H1B sponsorship
+3
QA Lead - Data
QA Lead - Data

Hilabs • Karnataka

On-site
INR 1,200,000 - 1,800,000
Senior Data Scientist
Senior Data Scientist

HiLabs Inc. • Bengaluru

On-site
INR 1,800,000 - 2,800,000
Competitive Salary
ESOPs
H1B sponsorship
+3
Principal Data Scientist
Principal Data Scientist

HiLabs Inc. • Pune District

On-site
INR 4,000,000 - 8,000,000
H1B sponsorship
ESOPs
Medical coverage
+1
Engineering Manager
Engineering Manager

HiLabs Inc. • Bengaluru

On-site
INR 4,000,000 - 7,000,000
H1B sponsorship
ESOPs
Medical coverage for you and your LOs
Engineering Manager
Engineering Manager

HiLabs Inc. • Pune District

On-site
INR 3,500,000 - 6,500,000
ESOPs
H1B sponsorship
Healthcare coverage
+2
Director Technical Architect
Director Technical Architect

HiLabs Inc. • Bengaluru

On-site
INR 3,000,000 - 6,000,000