Python & PySpark Developer ( 5-7 Years ) Hyderabad

HypTechie

Hyderabad

On-site

INR 1,500,000 - 2,100,000

Full time

5 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

HypTechie in Hyderabad seeks an experienced Python & PySpark Developer to build robust data pipelines using PySpark, Spark SQL, Delta Lake and Python-based data processing. The candidate should also have hands-on experience with GenAI/AI coding assistants to improve development, testing, debugging and documentation.

The role covers developing PySpark ETL jobs, debugging and optimizing Spark processing, and implementing data ingestion for CSV/XML files while ensuring test coverage with PyTest and

Qualifications

  • 5-7 years of Python/PySpark development experience.
  • Strong hands-on with PySpark, Spark SQL and DataFrames.
  • Experience with ETL pipelines and file ingestion.
  • Knowledge of Delta Lake and Spark processing.
  • Experience with PyTest and production troubleshooting.
  • Good Python debugging and code-quality practices.

Responsibilities

  • Develop, maintain and enhance PySpark ETL jobs.
  • Debug and optimize Spark DataFrame, Spark SQL and Delta Lake processing.
  • Develop data pipelines for CSV, XML and other file-based ingestion.
  • Implement reporting-date and business-date processing logic.
  • Write and execute PyTest-based unit and integration tests.
  • Troubleshoot production issues using Python logging and debugging.
  • Work with Python packaging, dependencies and code quality tools such as Pylint.
  • Optimize PySpark transformations and Delta Lake read/write operations.
  • Integrate SQL logic with PySpark-based data processing.
  • Analyze existing code and support impact assessment and documentation.

Skills

Python
PySpark
Spark SQL
Hadoop
Delta Lake
ETL / Data Pipelines
PySpark DataFrame
SQL to PySpark Integration
CSV / XML File Processing
PyTest
Python Debugging & Logging

Job description

Python & PySpark Developer

Job ID: 1484887

Location: Hyderabad

Experience: 5-7 Years

Openings: 1

Job Summary

We are looking for an experienced Python & PySpark Developer with strong expertise in PySpark, Spark SQL, ETL development, Delta Lake and Python-based data processing. The candidate should also have hands-on experience using GenAI/AI coding assistants to improve development, testing, debugging and documentation.

Primary Skills
  • Python
  • PySpark
  • Spark SQL
  • Hadoop
  • Delta Lake
  • ETL / Data Pipelines
  • PySpark DataFrame
  • SQL to PySpark Integration
  • CSV / XML File Processing
  • PyTest
  • Python Debugging & Logging
Key Responsibilities
  • Develop, maintain and enhance PySpark ETL jobs.
  • Debug and optimize Spark DataFrame, Spark SQL and Delta Lake processing.
  • Develop data pipelines for CSV, XML and other file-based ingestion.
  • Implement reporting-date and business-date processing logic.
  • Write and execute PyTest-based unit and integration tests.
  • Troubleshoot production issues using Python logging, exception handling and debugging techniques.
  • Work with Python packaging, runtime dependencies and code quality tools such as Pylint.
  • Optimize PySpark transformations and Delta Lake read/write operations.
  • Integrate SQL logic with PySpark-based data processing.
  • Analyze existing code and support impact assessment, root-cause analysis and technical documentation.
GenAI / AI Skills
  • Hands-on experience with GitHub Copilot, ChatGPT or similar AI coding assistants.
  • Ability to use GenAI for Python, PySpark and SQL development.
  • Experience using AI tools for ETL code analysis, debugging, impact assessment and documentation.
  • Ability to use AI-assisted code review to identify bugs, performance issues, security risks and code-quality gaps.
  • Experience generating or improving PyTest test cases using AI assistance.
  • Understanding of prompt engineering for SQL analysis, Spark debugging, log analysis and documentation.
  • Awareness of responsible AI usage, data privacy and validation of AI-generated outputs.
  • Ability to use AI tools to understand legacy codebases and create technical documentation.
  • Exposure to ML/AI, embeddings, vector search, RAG or LLM-based knowledge assistants is an advantage.
Required Experience
  • 5-7 years of experience in Python/PySpark development.
  • Strong hands-on experience with PySpark, Spark SQL and DataFrames.
  • Experience with ETL pipelines and file ingestion.
  • Good knowledge of Delta Lake and Spark processing.
  • Experience with PyTest and production troubleshooting.
  • Strong Python debugging and code-quality practices.
  • Good communication and documentation skills.
Mandatory Skills

Python | PySpark | Spark SQL | Hadoop | Delta Lake | ETL | DataFrames | PyTest | SQL | GenAI/AI Coding Assistants

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Contractor - Python & PySpark Developer
Contractor - Python & PySpark Developer

Vivantify • Hyderabad

On-site
INR 1,800,000 - 3,200,000
Competitive salary
Growth opportunities
Collaborative environment
Pyspark Developer (5 locations)
Pyspark Developer (5 locations)

Tata Consultancy Services • Bengaluru

On-site
INR 1,800,000 - 3,200,000
Contractor - PySpark Engineer
Contractor - PySpark Engineer

Vivantify • Hyderabad

On-site
INR 1,800,000 - 2,600,000
PySpark Developer / Senior Data Engineer
PySpark Developer / Senior Data Engineer

Alignity Solutions • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Developer – Big Data PySpark
Developer – Big Data PySpark

HypTechie • India

On-site
INR 1,000,000 - 1,100,000
PySpark Developer / Senior Data Engineer-MNC Client For 6 To 15Yrs imm
PySpark Developer / Senior Data Engineer-MNC Client For 6 To 15Yrs imm

Shell Infotech • Chennai District, Bengaluru, Hyderabad

Hybrid
INR 1,200,000 - 2,000,000
Senior PySpark Developer
Senior PySpark Developer

Basebiz • Chennai District, Bengaluru

On-site
INR 2,400,000 - 3,600,000
Python + Hadoop + Pyspark
Python + Hadoop + Pyspark

NITYO • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Data Engineer_ Python/pyspark + AWS
Data Engineer_ Python/pyspark + AWS

Tata Consultancy Services • Hyderabad

On-site
INR 1,400,000 - 2,300,000
Big Data Engineer -Pyspark, Spark, Hadoop, SQL
Big Data Engineer -Pyspark, Spark, Hadoop, SQL

Pi Square Technologies • Hyderabad, Pune District, Bengaluru

Hybrid
INR 1,800,000 - 3,000,000