Python PySpark Developer

Hexaware Technologies

Hyderabad, Pune District, Bengaluru

On-site

INR 1,200,000 - 2,100,000

Full time

6 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Hexaware Technologies seeks a skilled Python, PySpark, and SQL Developer to design, develop, and optimize scalable data processing solutions for large datasets.

The ideal candidate will have strong experience in data engineering, ETL development, and managing complex SQL queries, with exposure to Spark-based pipelines and cloud platforms. Chennai/Pune/Mumbai/Bangalore/Hyderabad locations are supported, with opportunities to work across India.

Qualifications

  • 4-9 years of Python/PySpark/SQL development experience.
  • Experience designing and optimizing ETL/ELT pipelines.
  • Strong knowledge of data warehousing concepts and dimensional modeling.
  • Experience with Linux/Unix environments and version control.

Responsibilities

  • Design, develop, and maintain scalable data pipelines using Python and PySpark.
  • Develop and optimize SQL queries, stored procedures, and database objects.
  • Build, automate, and support ETL/ELT processes for data integration and transformation.
  • Process and analyze large volumes of structured and unstructured data with Apache Spark.
  • Perform data cleansing, validation, and transformation to ensure data quality.
  • Monitor, troubleshoot, and optimize data processing jobs for performance and reliability.
  • Collaborate with data analysts, data engineers, and business stakeholders to gather requirements.
  • Implement data security, governance, and best practices for data management.
  • Work with cloud platforms and distributed computing environments.
  • Create technical documentation and support deployment activities.

Skills

Python programming
PySpark
Apache Spark
SQL
ETL pipelines
Data warehousing concepts
Performance tuning
Git
Linux/Unix
Problem solving
Analytical skills
Azure
AWS
GCP
Databricks
Azure Data Factory
Azure Synapse
Apache Airflow
CI/CD
DevOps
Data lake
Data warehouse architectures

Education

Bachelor's degree in Computer Science/Information Technology/Engineering or related field

Tools

Databricks
Azure Data Factory (ADF)
Azure Synapse Analytics
Git

Job description

Job Description: Python + PySpark + SQL Developer

Job Title: Python / PySpark / SQL Developer

Experience: 4-9 Years (customize as needed)

Location: Chennai/Pune/Mumbai/Bangalore/Hyderabad


Job Summary

We are seeking a skilled Python, PySpark, and SQL Developer to design, develop, and optimize scalable data processing solutions. The ideal candidate should have strong experience in data engineering, ETL development, big data technologies, and database management while working with large-scale datasets.

Key Responsibilities
  • Design, develop, and maintain scalable data pipelines using Python and PySpark.
  • Develop and optimize complex SQL queries, stored procedures, and database objects.
  • Build, automate, and support ETL/ELT processes for data integration and transformation.
  • Process and analyze large volumes of structured and unstructured data using Apache Spark.
  • Perform data cleansing, validation, and transformation to ensure data quality.
  • Monitor, troubleshoot, and optimize data processing jobs for performance and reliability.
  • Collaborate with data analysts, data engineers, and business stakeholders to understand requirements.
  • Implement data security, governance, and best practices for data management.
  • Work with cloud platforms and distributed computing environments.
  • Create technical documentation and support deployment activities.
Required Skills
  • Strong proficiency in Python programming.
  • Hands-on experience with PySpark and Apache Spark.
  • Strong knowledge of SQL and database concepts.
  • Experience in developing and maintaining ETL pipelines.
  • Understanding of data warehousing concepts and dimensional modeling.
  • Knowledge of performance tuning and query optimization.
  • Familiarity with version control tools such as Git.
  • Experience working with Linux/Unix environments.
  • Strong problem-solving and analytical skills.
Preferred Skills
  • Experience with cloud platforms such as Azure, AWS, or Google Cloud Platform (GCP).
  • Knowledge of Databricks, Azure Data Factory (ADF), Azure Synapse, or similar data platforms.
  • Experience with workflow orchestration tools such as Apache Airflow.
  • Exposure to CI/CD pipelines and DevOps practices.
  • Knowledge of data lake and data warehouse architectures.
Qualifications
  • Bachelor's degree in Computer Science, Information Technology, Engineering, or related field.
  • Relevant experience in Python, PySpark, and SQL development.
  • Certifications in Azure, Databricks, Spark, or cloud technologies are an added advantage.
Nice-to-Have Technologies
  • Databricks
  • Azure Data Factory (ADF)
  • Azure Synapse Analytics
  • Power BI
  • Delta Lake
  • Apache Kafka
  • Airflow
  • Snowflake
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Pyspark Developer
Pyspark Developer

Leading Global Technology Services Company • Chennai District, Bengaluru

Hybrid
INR 1,200,000 - 1,800,000
PySpark / Spark Developer
PySpark / Spark Developer

Infosys • Hyderabad, Pune District, Bengaluru

On-site
INR 1,500,000 - 2,300,000
PySpark Developer (2 To 3 Years)
PySpark Developer (2 To 3 Years)

Infosys • Dadri, Chennai District, Bengaluru

Hybrid
INR 600,000 - 900,000
Contractor - PySpark Engineer
Contractor - PySpark Engineer

Vivantify • Hyderabad

On-site
INR 1,800,000 - 2,600,000
Developer - PySpark
Developer - PySpark

Compunnel, Inc. • Pune District

On-site
INR 800,000 - 1,500,000
Data Engineer (Python & PySpark)
Data Engineer (Python & PySpark)

Techknomatic Services • Pune District

On-site
INR 700,000 - 1,200,000
Python Data Engineer (Blr/Chn/Hyd/Kochi/Kol/Pune)
Python Data Engineer (Blr/Chn/Hyd/Kochi/Kol/Pune)

Tata Consultancy Services • Hyderabad, Pune District, Bengaluru

On-site
INR 1,500,000 - 2,600,000
Data Engineer - SQL/PySpark
Data Engineer - SQL/PySpark

Forward Eye Technologies • Dadri

Hybrid
INR 1,400,000 - 2,000,000
Data Engineer - ETL/PySpark
Data Engineer - ETL/PySpark

Forward Eye Technologies • Pune District

On-site
INR 1,200,000 - 2,400,000
Walk-in | Pyspark Developer
Walk-in | Pyspark Developer

Tata Consultancy Services • Chennai District

On-site
INR 1,400,000 - 2,000,000