Python,Pyspark,GCP

Tata Consultancy Services

Hyderabad

On-site

INR 2,500,000 - 3,600,000

Full time

12 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Tata Consultancy Services is seeking a seasoned Data Engineer to design and maintain scalable data pipelines using Python and PySpark. The role emphasizes hands-on development, ETL processes, and performance optimization across distributed data environments.

Responsibilities include collaborating with cross-functional teams, ensuring data quality, and implementing best practices for reliability and scalability in cloud-enabled data architectures.

Qualifications

  • 6–8+ years of experience with Python and PySpark.
  • Strong proficiency in Python programming.
  • Hands-on experience with PySpark and Apache Spark.
  • Knowledge of Big Data technologies (Hadoop, Hive, Kafka, etc.).
  • Experience with SQL and relational/non-relational databases.
  • Familiarity with distributed computing and parallel processing.
  • Understanding of data engineering best practices.

Responsibilities

  • Develop and maintain scalable data pipelines using Python and PySpark.
  • Design and implement ETL (Extract, Transform, Load) processes.
  • Optimize and troubleshoot PySpark applications for performance.
  • Collaborate with cross-functional teams to understand data requirements.
  • Write clean, efficient, and well-documented code.
  • Conduct code reviews and participate in design discussions.

Skills

Python
PySpark
Apache Spark
Big Data
SQL
REST APIs
JSON/XML
Distributed computing
ETL
Data pipelines
GCP
AWS
Azure
Hadoop
Hive
Kafka
Code reviews

Job description

Role & responsibilities
  • 6 to 8 + Years of experience using Python and Pyspark.
  • Strong proficiency in Python programming.
  • Hands-on experience with PySpark and Apache Spark.
  • Knowledge of Big Data technologies (Hadoop, Hive, Kafka, etc.).
  • Experience with SQL and relational/non-relational databases.
  • Familiarity with distributed computing and parallel processing.
  • Understanding of data engineering best practices.
  • Experience with REST APIs, JSON/XML, and data serialization.
  • Exposure to GCP services and cloud computing environments.
  • Develop and maintain scalable data pipelines using Python and PySpark.
  • Design and implement ETL (Extract, Transform, Load) processes.
  • Optimize and troubleshoot existing PySpark applications for performance.
  • Collaborate with cross-functional teams to understand data requirements.
  • Write clean, efficient, and well-documented code.
  • Conduct code reviews and participate in design discussions.
  • Ensure data integrity and quality across the data lifecycle.
  • Integrate with cloud platforms like GCP, AWS or Azure.
  • Implement data storage solutions and manage large-scale datasets.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Python_PySpark Developer on GCP
Python_PySpark Developer on GCP

Tata Consultancy Services • Hyderabad

On-site
INR 1,200,000 - 2,200,000
Python_PySpark Developer on GCP
Python_PySpark Developer on GCP

TymblHub • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Digital : Python(GCP, Pyspark)
Digital : Python(GCP, Pyspark)

Tata Consultancy Services • Hyderabad

On-site
INR 1,100,000 - 1,700,000
Data Engineer- Python Pyspark GCP
Data Engineer- Python Pyspark GCP

Tata Consultancy Services • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Data Engineer- Python Pyspark GCP
Data Engineer- Python Pyspark GCP

Mployee.me • Hyderabad

On-site
INR 1,800,000 - 2,800,000
Digital : Python ,Pyspark, GCP
Digital : Python ,Pyspark, GCP

Tata Consultancy Services • Hyderabad

On-site
INR 1,800,000 - 2,400,000
Python /Pyspark Developer
Python /Pyspark Developer

CIEL HR • Bengaluru

On-site
INR 1,200,000 - 2,500,000
PySpark / Spark Developer
PySpark / Spark Developer

Infosys • Hyderabad, Pune District, Bengaluru

On-site
INR 1,500,000 - 2,300,000
Python, Pyspark, SQL
Python, Pyspark, SQL

Cognizant • Hyderabad

Hybrid
INR 1,000,000 - 1,600,000
Lead Data Engineer (Databricks, PySpark & GCP)
Lead Data Engineer (Databricks, PySpark & GCP)

Egen • Hyderabad

On-site
INR 5,500,000 - 7,500,000
Healthcare benefits
Performance bonus