Walk-in | Pyspark Data engineer

Tata Consultancy Services

Pune District

On-site

INR 1,000,000 - 1,400,000

Full time

5 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Tata Consultancy Services is seeking a data engineer in Pune to design, build, and optimize scalable data pipelines using PySpark and Python. You will work with large, diverse datasets, implement batch and real-time processing, and collaborate with analysts and scientists to deliver robust data solutions.

Responsibilities include developing Spark jobs for cleansing and transformation, optimizing performance, and ensuring governance, security, and compliance across data workloads.

Responsibilities

  • Design, develop, and maintain scalable data pipelines using PySpark and Python.
  • Build ETL/ELT frameworks for ingesting, transforming, and loading large datasets.
  • Develop Spark jobs for data cleansing, transformation, aggregation, and validation.
  • Optimize Spark applications for performance using partitioning, caching, broadcast joins, and tuning techniques.
  • Work with structured and unstructured data from multiple sources.
  • Implement batch and real-time data processing solutions.
  • Develop and optimize SQL queries, stored procedures, and data models.
  • Collaborate with business analysts, architects, and data scientists to deliver data solutions.
  • Ensure data quality, governance, security, and compliance standards.
  • Implement CI/CD and DevOps practices for data engineering workloads.
  • Troubleshoot production issues and provide operational support.

Job description

Role & responsibilities


  • Design, develop, and maintain scalable data pipelines using PySpark and Python.
  • Build ETL/ELT frameworks for ingesting, transforming, and loading large datasets.
  • Develop Spark jobs for data cleansing, transformation, aggregation, and validation.
  • Optimize Spark applications for performance using partitioning, caching, broadcast joins, and tuning techniques.
  • Work with structured and unstructured data from multiple sources.
  • Implement batch and real-time data processing solutions.
  • Develop and optimize SQL queries, stored procedures, and data models.
  • Collaborate with business analysts, architects, and data scientists to deliver data solutions.
  • Ensure data quality, governance, security, and compliance standards.
  • Implement CI/CD and DevOps practices for data engineering workloads.
  • Troubleshoot production issues and provide operational support.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Python /Pyspark Developer
Python /Pyspark Developer

CIEL HR • Bengaluru

On-site
INR 1,200,000 - 2,500,000
Pyspark Developer
Pyspark Developer

Sightspectrum • Chennai District

On-site
INR 900,000 - 1,300,000
Walk-in | Pyspark Developer
Walk-in | Pyspark Developer

Tata Consultancy Services • Chennai District

On-site
INR 1,400,000 - 2,000,000
Developer - PySpark
Developer - PySpark

Compunnel, Inc. • Pune District

On-site
INR 800,000 - 1,500,000
Data Engineer-Pyspark
Data Engineer-Pyspark

Deloitte US-India Offices • Chennai District

Hybrid
INR 900,000 - 1,300,000
PySpark Data Engineer
PySpark Data Engineer

Infosys • Hyderabad, Pune District, Bengaluru

On-site
INR 900,000 - 1,500,000
Python Pyspark Developer
Python Pyspark Developer

V2 Solutions • Ernakulam

On-site
INR 1,800,000 - 3,000,000
Data Engineer
Data Engineer

Alike Thoughts • Bengaluru

On-site
INR 1,200,000 - 1,900,000
Python, Pyspark Developer
Python, Pyspark Developer

Infosys • Hyderabad

On-site
INR 1,100,000 - 2,300,000
Pyspark Data Engineer
Pyspark Data Engineer

Clover Infotech • Chennai District

On-site
INR 1,200,000 - 1,800,000