Senior Data Engineer – PySpark, Cloud & Kafka - 5+ YoE - Immediate Joiner - Any UST Location

UST

Bengaluru

On-site

INR 1,000,000 - 1,500,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

A tech solutions company in India is seeking a professional to design, develop, and maintain scalable data pipelines using PySpark. Key responsibilities include building batch and real-time processing systems and optimizing Spark jobs. The ideal candidate should have hands-on experience with PySpark, strong SQL skills, and familiarity with cloud platforms like AWS. Experience with Kafka and CI/CD pipelines for data solutions is also preferred. This role offers an opportunity to collaborate with data teams and ensure data quality and security.

Qualifications

  • Strong hands-on experience with PySpark / Apache Spark.
  • Solid understanding of distributed data processing concepts.
  • Experience with Apache Kafka (producers, consumers, topics, partitions).
  • Hands-on experience with any one cloud platform (AWS preferred).
  • Proficiency in Python.
  • Strong experience with SQL and data modeling.
  • Experience working with large-scale datasets.
  • Familiarity with Linux/Unix environments.
  • Understanding of ETL/ELT frameworks.
  • Experience with CI/CD pipelines for data applications.

Responsibilities

  • Design, develop, and maintain scalable data pipelines using PySpark.
  • Build and manage batch and real-time data processing systems.
  • Develop and integrate Kafka-based streaming solutions.
  • Optimize Spark jobs for performance, cost, and scalability.
  • Work with cloud-native services to deploy and manage data solutions.
  • Ensure data quality, reliability, and security across platforms.
  • Collaborate with data scientists, analysts, and application teams.
  • Participate in code reviews, design discussions, and production support.

Job description

Candidates ready to join immediately can share their details via email for quick processing.

nitin.patil@ust.com

Key Responsibilities
  • Design, develop, and maintain scalable data pipelines using PySpark
  • Build and manage batch and real-time data processing systems
  • Develop and integrate Kafka-based streaming solutions
  • Optimize Spark jobs for performance, cost, and scalability
  • Work with cloud-native services to deploy and manage data solutions
  • Ensure data quality, reliability, and security across platforms
  • Collaborate with data scientists, analysts, and application teams
  • Participate in code reviews, design discussions, and production support
Must-Have Skills
  • Strong hands-on experience with PySpark / Apache Spark
  • Solid understanding of distributed data processing concepts
  • Experience with Apache Kafka (producers, consumers, topics, partitions)
  • Hands-on experience with any one cloud platform:
  • AWS (S3, EMR, Glue, EC2, IAM) or
  • Proficiency in Python
  • Strong experience with SQL and data modeling
  • Experience working with large-scale datasets
  • Familiarity with Linux/Unix environments
  • Understanding of ETL/ELT frameworks
  • Experience with CI/CD pipelines for data applications
Good-to-Have Skills
  • Experience with Spark Structured Streaming
  • Knowledge of Kafka Connect and Kafka Streams
  • Exposure to Databricks
  • Experience with NoSQL databases (Cassandra, MongoDB, HBase)
  • Familiarity with workflow orchestration tools (Airflow, Oozie)
  • Knowledge of containerization (Docker, Kubernetes)
  • Experience with data lake architectures
  • Understanding of security, governance, and compliance in cloud environments
  • Exposure to Scala or Java is a plus
  • Prior experience in Agile/Scrum environments
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer – PySpark/Python | Big Data & Cloud | 6+ Years Experience | Any UST Location | Im[...]
Data Engineer – PySpark/Python | Big Data & Cloud | 6+ Years Experience | Any UST Location | Im[...]

UST • Bengaluru

On-site
INR 1,200,000 - 1,800,000
6 + YoE - Data Engineer – Big Data / PySpark - any UST Location - Immediate Joiner
6 + YoE - Data Engineer – Big Data / PySpark - any UST Location - Immediate Joiner

UST • Bengaluru

On-site
INR 1,500,000 - 2,200,000
Senior Data Platform Engineer (Python/Spark)
Senior Data Platform Engineer (Python/Spark)

Genpact • Bengaluru

On-site
INR 4,000,000 - 7,500,000
Sr. Python Data Engineer – Offline Evaluation & PySpark | 7+ Years | Bangalore | Immediate Joiner
Sr. Python Data Engineer – Offline Evaluation & PySpark | 7+ Years | Bangalore | Immediate Joiner

UST • Bengaluru

On-site
INR 1,000,000 - 1,800,000
Walk-in | Pyspark Developer
Walk-in | Pyspark Developer

Tata Consultancy Services • Chennai District

On-site
INR 1,400,000 - 2,000,000
Data Engineer - SQL/PySpark
Data Engineer - SQL/PySpark

Forward Eye Technologies • Dadri

Hybrid
INR 1,400,000 - 2,000,000
Python Data Engineer (Blr/Chn/Hyd/Kochi/Kol/Pune)
Python Data Engineer (Blr/Chn/Hyd/Kochi/Kol/Pune)

Tata Consultancy Services • Hyderabad, Pune District, Bengaluru

On-site
INR 1,500,000 - 2,600,000
Data platform Engineer
Data platform Engineer

UST • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Contractor - PySpark Engineer
Contractor - PySpark Engineer

Vivantify • Hyderabad

On-site
INR 1,800,000 - 2,600,000
Data Engineer - SQL/PySpark
Data Engineer - SQL/PySpark

Forward Eye Technologies • Pune District

On-site
INR 1,500,000 - 2,100,000