Data Engineer - PySpark

Tekskills

Bengaluru

On-site

INR 1,200,000 - 1,800,000

Full time

9 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Tekskills seeks a skilled data engineer with deep Apache Spark expertise (batch and streaming) and Hive mastery. Proficiency in Python, Scala, or Java is essential.

The role requires familiarity with Airflow or Control-M, DBT, and experience with Kafka, Solace, and S3/MinIO for scalable data pipelines. Candidate should be comfortable deploying with Docker/Kubernetes and working with Lakehouse formats like Iceberg, Delta Lake, or Hudi.

Qualifications

  • Proficient in Python, Scala or Java.
  • Strong knowledge of Spark (batch and streaming) and Hive.
  • Experience with orchestration tools (Airflow, Control-M).
  • Experience with SQL transformation frameworks (DBT preferred).
  • Experience with Kafka, Solace and object stores (S3/MinIO).
  • Exposure to Docker and Kubernetes for deployment.
  • Hands-on experience with data Lakehouse formats (Iceberg, Delta Lake, Hudi).

Responsibilities

  • Strong expertise in Apache Spark (batch + streaming) and Hive.
  • Proficiency in Python, Scala, or Java.
  • Knowledge of orchestration tools (Airflow / Control-M) and SQL transformation frameworks (DBT preferred).
  • Experience working with Kafka, Solace, and object stores (S3, MinIO).
  • Exposure to Docker/Kubernetes for deployment.
  • Hands on experience of data Lakehouse formats (Iceberg, Delta Lake, Hudi).

Skills

Python
Scala
Java

Tools

Apache Spark
Hive
Airflow
Control-M
DBT
Kafka
Solace
Docker
Kubernetes
Iceberg
Delta Lake
Hudi
S3
MinIO

Job description

Job Summary

Role & responsibilities Strong expertise in Apache Spark (batch + streaming) and Hive. Proficiency in Python, Scala, or Java. Knowledge of orchestration tools (Airflow / Control-M) and SQL transformation frameworks (DBT preferred). Experience working with Kafka, Solace, and object stores (S3, MinIO). Exposure to Docker/Kubernetes for deployment. Hands on experience of data Lakehouse formats (Iceberg, Delta Lake, Hudi).

Responsibilities
  • Strong expertise in Apache Spark (batch + streaming) and Hive.
  • Proficiency in Python, Scala, or Java.
  • Knowledge of orchestration tools (Airflow / Control-M) and SQL transformation frameworks (DBT preferred).
  • Experience working with Kafka, Solace, and object stores (S3, MinIO).
  • Exposure to Docker/Kubernetes for deployment.
  • Hands on experience of data Lakehouse formats (Iceberg, Delta Lake, Hudi).
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer - PySpark
Data Engineer - PySpark

Tekskills • Chennai District

On-site
INR 1,500,000 - 2,100,000
Data Engineer (Hadoop, Python, PySpark)
Data Engineer (Hadoop, Python, PySpark)

Tekskills • Chennai District

On-site
INR 900,000 - 1,300,000
Data Engineer (Hadoop, Python, PySpark)
Data Engineer (Hadoop, Python, PySpark)

Tekskills • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Data engineer (Apache Hadoop, Python and PySpark)
Data engineer (Apache Hadoop, Python and PySpark)

Tekskills • Bengaluru

On-site
INR 1,200,000 - 2,400,000
Data Engineer With Java + Spark
Data Engineer With Java + Spark

Wissen Technology • Pune District

Hybrid
INR 1,200,000 - 1,800,000
Data Engineer - SQL/PySpark
Data Engineer - SQL/PySpark

Forward Eye Technologies • Pune District

On-site
INR 1,500,000 - 2,100,000
Data Engineer (PySpark + Cloudera)
Data Engineer (PySpark + Cloudera)

Zorba AI • Maharashtra

On-site
INR 1,000,000 - 1,500,000
Data Engineer (PySpark):
Data Engineer (PySpark):

The Hiring Club • Bengaluru

On-site
INR 800,000 - 1,200,000
Data Engineer-Pyspark
Data Engineer-Pyspark

Deloitte US-India Offices • Pune District

On-site
INR 2,000,000 - 4,000,000
DATA Engineer (JAVA SPARK)
DATA Engineer (JAVA SPARK)

Wissen Technology • Pune District

Hybrid
INR 1,400,000 - 2,200,000