Python PySpark Developer

Tata Consultancy Services

Chennai District, Bengaluru, Pune District

On-site

INR 600,000 - 1,200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Tata Consultancy Services is seeking a skilled data engineer based in Chennai, Tamil Nadu. The role involves implementing data ingestion pipelines, working with both structured and unstructured data, and developing scalable cloud-based solutions using PySpark.

The ideal candidate should have experience in ETL processes and data warehouse transformation. A strong understanding of data quality and performance optimization is essential. This position offers the chance to apply best practices in automation and testing to improve data processing frameworks.

Qualifications

  • Experience in implementing data ingestion pipelines from various sources.
  • Proficiency in building ETL/Data Warehouse transformation processes.
  • Experience with both structured and unstructured data.

Skills

Data ingestion pipeline implementation
ETL/Data Warehouse experience
Working with structured and unstructured data
Big Data solutions in PySpark
Scalable frameworks for data processing
Data quality and consistency
Performance analysis and optimization
Best practices in Automation and Testing

Job description

Must have Skills:
  • Implementing data ingestion pipelines from different types of data sources i.e Databases, S3, Files etc..
  • Experience in building ETL/ Data Warehouse transformation process.
  • Experience working with structured and unstructured data.
  • Developing Big Data and non-Big Data cloud-based enterprise solutions in PySpark and SparkSQL and related frameworks/libraries.
  • Developing scalable and re-usable, self-service frameworks for data ingestion and processing.
  • Integrating end to end data pipelines to take data from data source to target data repositories ensuring the quality and consistency of data.
  • Processing performance analysis and optimization.
  • Bringing best practices in following areas: Design & Analysis, Automation (Pipelining, IaC), Testing, Monitoring, Documentation.
Good to have (Knowledge):
  • Experience in cloud-based solutions,
  • Knowledge of data management principles
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Developer
Developer

GSB Solutions • Hyderabad

On-site
INR 1,000,000 - 1,500,000
Developer
Developer

GSB Solutions • Hyderabad

On-site
INR 1,200,000 - 1,500,000
Python, Pyspark Developer
Python, Pyspark Developer

Infosys • Hyderabad

On-site
INR 1,100,000 - 2,300,000
Python-Pyspark Developer
Python-Pyspark Developer

Infosys • Hyderabad

On-site
INR 2,400,000 - 3,600,000
Developer - PySpark
Developer - PySpark

Compunnel, Inc. • Pune District

On-site
INR 800,000 - 1,500,000
Python, PySpark, ETL Developer
Python, PySpark, ETL Developer

Infosys • Hyderabad

On-site
INR 800,000 - 1,600,000
Python with Data Engineer
Python with Data Engineer

Tata Consultancy Services • Chennai District, Bengaluru, Hyderabad

On-site
INR 1,800,000 - 3,000,000
Data Engineer (PySpark):
Data Engineer (PySpark):

The Hiring Club • Bengaluru

On-site
INR 800,000 - 1,200,000
PySpark Engineer
PySpark Engineer

Pagaar India • Hyderabad

On-site
INR 1,200,000 - 2,100,000
Senior Data Engineer - Databricks, PySpark & Lakehouse
Senior Data Engineer - Databricks, PySpark & Lakehouse

Tata Consultancy Services • Bengaluru, Pune District, Chennai District

On-site
INR 1,500,000 - 3,000,000