Lead Data Engineer – AWS

Jobtailor

Bengaluru

On-site

INR 1,200,000 - 1,800,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Jobtailor Bengaluru seeks an experienced Data Engineer to design scalable data pipelines using Spark/PySpark on Databricks and to optimize ETL/ELT workflows for large-scale processing.

You will work with AWS EMR/S3 and Hadoop, build data lakes/warehouses, and implement API data ingestion and orchestration with Airflow/Autosys. Strong Python and SQL skills are essential.

Qualifications

  • Minimum 5+ years of Data Engineering experience.
  • Strong hands-on experience in Databricks + Big Data stack.
  • Proven experience working on large-scale distributed systems.
  • Strong expertise in Apache Spark / PySpark.
  • Databricks mandatory.
  • SQL knowledge (PostgreSQL or similar).
  • Hands-on with AWS EMR, S3; Hadoop/Hive ecosystem.
  • Programming: Python mandatory; Scala a plus.
  • UNIX / Shell scripting knowledge.
  • ETL pipelines and data engineering fundamentals.
  • Experience with Elasticsearch for storage/retrieval.
  • Workflow orchestration: Airflow, Autosys.
  • API development & integration exposure.
  • Version control: Git / SVN.
  • Basic HTML know-how for web/data tasks.

Responsibilities

  • Design and build scalable data pipelines using Spark / PySpark on Databricks.
  • Develop and optimize ETL/ELT workflows for large-scale data processing.
  • Work with AWS EMR, S3, and Hadoop ecosystem for distributed data processing.
  • Build and maintain data lake and data warehouse solutions.
  • Develop and integrate APIs for data ingestion and consumption.
  • Implement data processing workflows using Airflow / Autosys.
  • Optimize data performance using partitioning, caching, and query tuning.
  • Handle structured and semi-structured data from multiple sources.
  • Ensure data quality, governance, and reliability of pipelines.
  • Collaborate with cross-functional teams (Analytics, Product, Engineering).

Skills

Spark/PySpark
Databricks
Big Data
SQL
Python
UNIX/Shell
Airflow
Autosys
APIs
Git/SVN
Elasticsearch
Hadoop/Hive

Tools

AWS EMR
S3

Job description

Responsibilities
  • Design and build scalable data pipelines using Spark / PySpark on Databricks
  • Develop and optimize ETL/ELT workflows for large-scale data processing
  • Work with AWS EMR, S3, and Hadoop ecosystem for distributed data processing
  • Build and maintain data lake and data warehouse solutions
  • Develop and integrate APIs for data ingestion and consumption
  • Implement data processing workflows using Airflow / Autosys
  • Optimize data performance using partitioning, caching, and query tuning
  • Handle structured and semi-structured data from multiple sources
  • Ensure data quality, governance, and reliability of pipelines
  • Collaborate with cross-functional teams (Analytics, Product, Engineering)
Requirements
  • Minimum 5+ years of Data Engineering experience
  • Strong hands-on experience in Databricks + Big Data stack
  • Proven experience working on large-scale distributed systems
  • Strong expertise in Apache Spark / PySpark
  • Databricks (mandatory)
  • SQL (PostgreSQL or similar)
  • Hands-on experience with AWS EMR, S3
  • Hadoop, Hive ecosystem
  • Programming skills: Python (must-have), Scala (good exposure)
  • Strong knowledge of UNIX / Shell scripting
  • ETL pipelines and data engineering fundamentals
  • Experience with Elasticsearch (data storage & retrieval)
  • Workflow orchestration tools (Airflow, Autosys)
  • Exposure to API development & integration
  • Version control tools (Git / SVN)
  • Basic knowledge of HTML (for web/data tasks)
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Specialist
Senior Data Specialist

Jobtailor • Bengaluru

On-site
INR 1,500,000 - 2,600,000
Lead Data Engineer, Databricks
Lead Data Engineer, Databricks

Jobtailor • Bengaluru

On-site
INR 900,000 - 1,500,000
Data Engineer
Data Engineer

HGS • Hyderabad

On-site
INR 1,200,000 - 2,100,000
Data Engineer ( AWS & Databricks)
Data Engineer ( AWS & Databricks)

Tekskills • Kolkata District

Hybrid
INR 1,500,000 - 2,100,000
Data Engineer ( AWS & Databricks)
Data Engineer ( AWS & Databricks)

Tekskills • Pune District

Hybrid
INR 1,200,000 - 2,000,000
Lead Data Engineer (Databricks & PySpark)
Lead Data Engineer (Databricks & PySpark)

Experis • Pune District

Hybrid
INR 4,000,000 - 7,000,000
Data Engineer ( AWS & Databricks)
Data Engineer ( AWS & Databricks)

Tekskills • Hyderabad

Hybrid
INR 2,500,000 - 4,200,000
Data Engineer ( AWS & Databricks)
Data Engineer ( AWS & Databricks)

Tekskills • Bulandshahr

Hybrid
INR 1,200,000 - 2,100,000
Lead Software Engineer
Lead Software Engineer

Impetus • Bengaluru

On-site
INR 1,000,000 - 2,000,000
Senior Data Engineer
Senior Data Engineer

Kumaran Systems • Chennai District

On-site
INR 1,200,000 - 1,800,000