Pyspark Developer

Infosys

Pune District

On-site

INR 1,200,000 - 2,100,000

Full time

2 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Continuous learning programs
Certifications and career growth

Job summary

Infosys in Pune, India seeks a PySpark Developer with 3-5 years in building scalable data solutions. You will work with PySpark, Hadoop, Hive, and related big data technologies to design and implement robust data pipelines.

The role involves collaborating with stakeholders, participating in scrum rituals, and contributing to data warehousing concepts and ETL processes. The ideal candidate communicates well and stays current with evolving big data practices.

Qualifications

  • 3–5 years of experience in PySpark and big data tech.
  • Strong expertise in PySpark with Spark, Hadoop, SQL exposure to Hive, Sqoop.
  • Experience with distributed data processing pipelines.
  • Understanding of data warehousing concepts and ETL processes.

Responsibilities

  • Design and development of large-scale PySpark data pipelines using Hadoop stack.
  • Collaborate with business stakeholders and SMEs to translate requirements.
  • Participate in scoping and estimating project plans.
  • Work in an Agile environment with daily standups and sprints.
  • Ensure data quality and performance optimization.

Skills

PySpark
Hadoop
Spark Core
Spark SQL
Batch processing
Spark Streaming
Unix shell scripting
Analytical skills
Communication skills

Education

Master of Computer Applications
Master of Technology
Bachelor of Computer Applications
Bachelor of Science
Bachelor of Engineering
Bachelor of Technology

Tools

HDFS
Hive
Sqoop
Unix shell
Cloud platforms

Job description

Job Description

We are seeking a skilled Pyspark Developer with 3-5 years of experience in building scalable solutions on Pyspark. The ideal candidate should have strong expertise in Pyspark with Spark, Hadoop, SQL exposure to Hive, Sqoop, UNIX shell scripting.

Roles Responsibilities
  • At least 3+ years of experience in designing and developing large scale, distributed data processing pipelines using PySpark, Hadoop and related technologies.
  • Having expertise in Pyspark, Hadoop, Spark Core, Spark SQL, Batch processing and Spark Streaming
  • Experience with Hadoop, HDFS, Hive and other BigData technologies.
  • Familiarity with Data warehousing and ETL concepts and techniques
  • UNIX shell scripting will be an added advantage in scheduling/running application jobs.
  • At least 3 years of experience in Project development life cycle activities and development/maintenance projects
  • Work with business stakeholders and other SMEs to understand high level business requirements.
  • Work with the Solution Designers and contribute to the development of project plans by participating in the scoping and estimating of proposed project.
  • Apply technical background understanding, business knowledge, system knowledge in the elicitation of Systems Requirements for projects.
  • Possess good knowledge on Spark architecture and transformations using Spark and PySpark.
  • Work in an Agile environment and participation in scrum daily standups, sprint planning reviews and retrospectives.
  • Understand project requirements and translate them into technical solutions which meets the project quality standards
  • Ability to work in team in diverse/multiple stakeholder environment and collaborate with upstream/downstream functional teams to identify, troubleshoot and resolve data issues.
    Technical RequirementPrimary skills:
    Pyspark, Hadoop, Spark Core, Spark SQL, Batch processing and Spark Streaming
    Hadoop, HDFS, Hive and other BigData technologies.
    Strong problem solving and Good Analytical skills.
  • Excellent verbal and written communication skills.
  • Experience and desire to work in a Global delivery environment.
  • Stay up to date with new technologies and industry trends in Development.
    Additional ResponsibilityInfosys is a global leader in next-generation digital services and consulting. Within the Data Analytics unit, you will work on cutting-edge data engineering and analytics initiatives for global clients across industries. The role offers opportunities to work with cloud-native data platforms, modern analytics ecosystems, AI-driven solutions, and large-scale enterprise data modernization programs. Employees benefit from continuous learning programs, certifications, and career growth opportunities within Infosys' data and analytics practice
    Educational RequirementMaster Of Comp. Applications,Master Of Technology,Bachelor Of Comp. Applications,Bachelor Of Science,Bachelor of Engineering,Bachelor Of Technology
    Preferred SkillsTechnology->Big Data - Data Processing->PySpark
    Service LineData Analytics Unit
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Pyspark Developer
Pyspark Developer

Infosys • Hyderabad

On-site
INR 1,500,000 - 2,400,000
Pyspark Developer
Pyspark Developer

Infosys • Chennai District

On-site
INR 1,200,000 - 1,800,000
Pyspark Developer
Pyspark Developer

Infosys Limited • Pune District

On-site
INR 900,000 - 1,200,000
PySpark / Spark Developer
PySpark / Spark Developer

Infosys • Hyderabad, Pune District, Bengaluru

On-site
INR 1,500,000 - 2,300,000
Hadoop / PySpark
Hadoop / PySpark

Infosys • Bengaluru

On-site
INR 1,800,000 - 3,200,000
Python PySpark Developer
Python PySpark Developer

Hexaware Technologies • Hyderabad, Pune District, Bengaluru

On-site
INR 1,200,000 - 2,100,000
PySpark Developer (2 To 3 Years)
PySpark Developer (2 To 3 Years)

Infosys • Dadri, Chennai District, Bengaluru

Hybrid
INR 600,000 - 900,000
Azure / Databricks/ PySpark Developer - Hyderabad
Azure / Databricks/ PySpark Developer - Hyderabad

TymblHub • Hyderabad

On-site
INR 800,000 - 1,100,000
Pyspark developer
Pyspark developer

Aligned Automation • Pune District

On-site
INR 2,200,000 - 3,400,000
Pyspark
Pyspark

Tata Consultancy Services • Chennai District

On-site
INR 1,800,000 - 2,600,000