Tcs Hiring For Pyspark Data Engineer

Tata Consultancy Services

Chennai District

On-site

INR 1,200,000 - 2,400,000

Full time

3 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Tata Consultancy Services in Chennai is seeking a PySpark Data Engineer to design, develop, and optimize large-scale data processing pipelines using PySpark and Spark ecosystem.

You will work with Data Scientists and Analysts to ingest, transform, and analyze data, build scalable ETL/ELT solutions, and ensure data quality across sources.

Qualifications

  • Strong hands-on experience in PySpark development.
  • Expertise in Spark SQL, DataFrames, and Spark optimization techniques.
  • Experience with Hadoop ecosystem (Hive, HDFS, Kafka).
  • Good knowledge of SQL and database concepts.
  • Experience in ETL pipeline development.
  • Knowledge of Azure Databricks, Azure Data Factory, or AWS EMR is preferred.
  • Experience with Git, CI/CD, and Agile methodologies.

Responsibilities

  • Develop and optimize data pipelines using PySpark and Apache Spark.
  • Perform data ingestion, transformation, and processing of large datasets.
  • Design scalable ETL/ELT solutions for data warehousing and analytics.
  • Work with distributed computing frameworks and big data technologies.
  • Integrate data from multiple sources and ensure data quality.
  • Collaborate with Data Scientists, Analysts, and Business teams.
  • Troubleshoot performance issues and optimize Spark jobs.
  • Follow best practices for data governance, security, and code quality.

Skills

PySpark development
Spark SQL DataFrames
SQL and database concepts
ETL pipeline development
CI/CD Agile methodologies

Tools

Hadoop Hive HDFS Kafka
Azure Databricks
Azure Data Factory
AWS EMR
Git
CI/CD

Job description

Job Description PySpark Data Engineer

Job Title: PySpark Data Engineer Experience: 5-10 Years Location: Chennai

Job Summary

We are looking for a skilled PySpark Data Engineer to design, develop, and optimize large-scale data processing pipelines. The ideal candidate should have strong expertise in PySpark, Spark ecosystem, SQL, ETL development, and cloud/big data technologies.

Key Responsibilities
  • Develop and optimize data pipelines using PySpark and Apache Spark.
  • Perform data ingestion, transformation, and processing of large datasets.
  • Design scalable ETL/ELT solutions for data warehousing and analytics.
  • Work with distributed computing frameworks and big data technologies.
  • Integrate data from multiple sources and ensure data quality.
  • Collaborate with Data Scientists, Analysts, and Business teams.
  • Troubleshoot performance issues and optimize Spark jobs.
  • Follow best practices for data governance, security, and code quality.
Required Skills
  • Strong hands-on experience in PySpark development.
  • Expertise in Spark SQL, DataFrames, and Spark optimization techniques.
  • Experience with Hadoop ecosystem (Hive, HDFS, Kafka).
  • Good knowledge of SQL and database concepts.
  • Experience in ETL pipeline development.
  • Knowledge of Azure Databricks, Azure Data Factory, or AWS EMR is preferred.
  • Experience with Git, CI/CD, and Agile methodologies.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Pyspark Developer
Pyspark Developer

Tata Consultancy Services • Kolkata District, Hyderabad, Chennai District

On-site
INR 2,600,000 - 3,800,000
Pyspark Data Engineer
Pyspark Data Engineer

Tata Consultancy Services • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Data Engineer (Spark and Scala)
Data Engineer (Spark and Scala)

Tata Consultancy Services • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Pyspark Developer
Pyspark Developer

Tata Consultancy Services • Pune City

On-site
INR 800,000 - 1,500,000
Hadoop + Pyspark
Hadoop + Pyspark

Alike Thoughts • Hyderabad, Chennai District

On-site
INR 1,400,000 - 2,200,000
Pyspark Developer
Pyspark Developer

Leading Global Technology Services Company • Chennai District, Bengaluru

Hybrid
INR 1,200,000 - 1,800,000
Data Engineer - ETL
Data Engineer - ETL

Forward Eye Technologies • Pune District

On-site
INR 2,000,000 - 3,200,000
PySpark Data Engineer
PySpark Data Engineer

Infosys • Hyderabad, Pune District, Bengaluru

On-site
INR 900,000 - 1,500,000
Data Engineer – PySpark & Snowflake
Data Engineer – PySpark & Snowflake

DTC Infotech. Pvt. Ltd. • Bengaluru

On-site
INR 600,000 - 1,200,000
Data Engineer
Data Engineer

Lorven Technologies Inc. • Tamil Nadu

On-site
INR 1,500,000 - 2,100,000