Pyspark Developer

Infosys Limited

Pune District

On-site

INR 900,000 - 1,200,000

Full time

36 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Infosys Limited is seeking a Pyspark Developer in Pune with 3-5 years of experience to design and build scalable data pipelines using PySpark, Hadoop, Spark Core, and Spark SQL. You will contribute to batch processing and Spark Streaming, work with HDFS/Hive, and engage in Agile development with cross-functional teams across data and analytics initiatives.

The role emphasizes data warehousing knowledge, ETL concepts, Unix shell scripting, and collaboration with business stakeholders and solution

Qualifications

  • 3+ years of experience in designing large-scale, distributed data pipelines using PySpark, Hadoop and related technologies.
  • Experience with Hadoop, HDFS, Hive and other BigData technologies.
  • Familiarity with data warehousing and ETL concepts and techniques.

Responsibilities

  • Design and develop large-scale data pipelines using PySpark, Hadoop and related technologies.
  • Work with Spark Core, Spark SQL, Batch processing and Spark Streaming.
  • Collaborate with stakeholders to understand requirements and translate them into technical solutions.

Education

Master of Computer Applications
Master of Technology
Bachelor of Computer Applications
Bachelor of Science
Bachelor of Engineering
Bachelor of Technology

Tools

PySpark
Hadoop
Spark Core
Spark SQL
Batch processing
Spark Streaming
HDFS
Hive
UNIX shell scripting

Job description

Job ID/Reference Code INFSYS-INDEED1-254353

Work Experience 3 - 5 Years

Job Title Pyspark Developer

Educational Requirements Master Of Comp. Applications,Master Of Technology,Bachelor Of Comp. Applications,Bachelor Of Science,Bachelor of Engineering,Bachelor Of Technology

Service Line Data & Analytics Unit

Responsibilities
  • At least 3+ years of experience in designing and developing large scale, distributed data processing pipelines using PySpark, Hadoop and related technologies.
  • Having expertise in Pyspark, Hadoop, Spark Core, Spark SQL, Batch processing and Spark Streaming
  • Experience with Hadoop, HDFS, Hive and other BigData technologies.
  • Familiarity with Data warehousing and ETL concepts and techniques
  • UNIX shell scripting will be an added advantage in scheduling/running application jobs.
  • At least 3 years of experience in Project development life cycle activities and development/maintenance projects
  • Work with business stakeholders and other SMEs to understand high level business requirements.
  • Work with the Solution Designers and contribute to the development of project plans by participating in the scoping and estimating of proposed project.
  • Apply technical background understanding, business knowledge, system knowledge in the elicitation of Systems Requirements for projects.
  • Possess good knowledge on Spark architecture and transformations using Spark and PySpark.
  • Work in an Agile environment and participation in scrum daily standups, sprint planning reviews and retrospectives.
  • Understand project requirements and translate them into technical solutions which meets the project quality standards
  • Ability to work in team in diverse/multiple stakeholder environment and collaborate with upstream/downstream functional teams to identify, troubleshoot and resolve data issues.

Additional Responsibilities: Infosys is a global leader in next-generation digital services and consulting. Within the Data & Analytics unit, you will work on cutting-edge data engineering and analytics initiatives for global clients across industries. The role offers opportunities to work with cloud-native data platforms, modern analytics ecosystems, AI-driven solutions, and large-scale enterprise data modernization programs. Employees benefit from continuous learning programs, certifications, and career growth opportunities within Infosys' data and analytics practice

Technical and Professional Requirements: Primary skills:Pyspark, Hadoop, Spark Core, Spark SQL, Batch processing and Spark StreamingHadoop, HDFS, Hive and other BigData technologies.Strong problem solving and Good Analytical skills. Excellent verbal and written communication skills. Experience and desire to work in a Global delivery environment. Stay up to date with new technologies and industry trends in Development.

Preferred Skills:

Technology->Big Data - Data Processing->PySpark

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Pyspark Developer
Pyspark Developer

Infosys • Chennai District

On-site
INR 1,200,000 - 1,800,000
Azure | Databricks | PySpark - Pan India
Azure | Databricks | PySpark - Pan India

Infosys • Hyderabad, Pune District, Bengaluru

Hybrid
INR 600,000 - 1,000,000
PySpark Developer (2 To 3 Years)
PySpark Developer (2 To 3 Years)

Infosys • Dadri, Chennai District, Bengaluru

Hybrid
INR 600,000 - 900,000
PySpark Developer
PySpark Developer

Infosys • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Python PySpark Developer
Python PySpark Developer

Hexaware Technologies • Hyderabad, Pune District, Bengaluru

On-site
INR 1,200,000 - 2,100,000
Pyspark Developer (Open For Pune /Kolkata also)
Pyspark Developer (Open For Pune /Kolkata also)

Tata Consultancy Services • Hyderabad, Chennai District, Bengaluru

On-site
INR 1,200,000 - 1,800,000
Azure / Databricks/ PySpark Developer - Hyderabad
Azure / Databricks/ PySpark Developer - Hyderabad

TymblHub • Hyderabad

On-site
INR 800,000 - 1,100,000
Pyspark Data Engineer
Pyspark Data Engineer

Synechron • Bengaluru

On-site
INR 900,000 - 1,500,000
PySpark / Spark Developer
PySpark / Spark Developer

Infosys • Hyderabad, Pune District, Bengaluru

On-site
INR 1,500,000 - 2,300,000
Hiring Big Data Developer - Python/Pyspark - Chennai/Bangalore
Hiring Big Data Developer - Python/Pyspark - Chennai/Bangalore

Tech Mahindra • Chennai District, Bengaluru

Hybrid
INR 900,000 - 1,400,000