Lead Software Engineer - Pyspark, AWS, Python

JPMorgan Chase & Co.

Pune District

On-site

INR 1,500,000 - 2,500,000

Full time

3 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

JPMorganChase in Pune, India, invites a Lead Software Engineer to shape scalable data engineering in a cloud-first environment. You will own end-to-end pipelines, from design to deployment, partnering with analytics teams and product partners.

The role emphasizes Spark-based processing, Python development, and robust data governance within a secure, compliant platform. On-site in Pune with opportunities to influence enterprise data strategy.

Qualifications

  • Formal training or certification in software engineering concepts.
  • 5+ years of applied experience in software/data engineering.
  • Strong experience with distributed data processing and cloud data lakes.

Responsibilities

  • Lead design and maintenance of cloud-based data processing pipelines.
  • Architect data models for large-scale datasets and analytics.
  • Collaborate with cross-functional teams to translate business requirements into data solutions.
  • Drive data quality, governance, and secure data handling.

Skills

Spark
Python
SQL
AWS
Airflow
Docker

Education

Software engineering certification

Tools

Databricks
Kubernetes
Snowflake

Job description

Lead Software Engineer - Pyspark, AWS, Python

Pune, Maharashtra, India

  • Job Identification 210786029
  • Job Category Software Engineering
  • Business Unit Consumer & Community Banking
  • Posting Date 10/05/2026, 09:46 AM
  • Locations International Tech Park Pune Kharadi, 13th Floor, Block 2, Gat Nos. 1344/3 & 1344/4, Wagholi, Survey no 63/1/6, Kharadi,, Pune, IN-MH, 411014, IN
  • Apply Before 10/09/2026, 04:00 AM
  • Job Schedule Full time
Job Description

We have an opportunity to impact your career and provide an adventure where you can push the limits of what's possible.

As a Lead Software Engineer at JPMorganChase within the Consumer and Community Banking, youare an integral part of an agile team that works to enhance, build, and deliver trusted market-leading technology products in a secure, stable, and scalable way. As a core technical contributor, you are responsible for conducting critical technology solutions across multiple technical areas within various business functions in support of the firm’s business objectives.

Job responsibilities

  • Lead the design, development, and maintenance of robust, scalable cloud-based data processing pipelines and infrastructure, ensuring adherence to engineering standards, governance frameworks, and industry best practices.
  • Architect and refine data models for large-scale datasets, optimizing for efficient storage, high-performance retrieval, and advanced analytics while upholding data integrity and quality.
  • Partner with cross-functional teams to translate complex business requirements into effective, scalable data engineering solutions that drive organizational value.
  • Champion a culture of innovation and continuous improvement, proactively identifying and implementing enhancements to data infrastructure, processing workflows, and analytics capabilities.
  • Define and execute data strategy, including the development of enterprise data models and the management of end-to-end data infrastructure—from design and construction to installation and ongoing maintenance of large-scale processing systems.
  • Drive data quality initiatives, ensure seamless data accessibility for analysts and data scientists, and maintain strict compliance with data governance and regulatory requirements.
  • Align data engineering practices with business objectives, ensuring solutions are both technically sound and strategically relevant.
  • Author, review, and approve technical requirements and architectural designs, and lead process re-engineering efforts to deliver cost-effective, high-impact business solution
  • Drives team adoption of enterprise-authorized AI-assisted engineering practices within the work environment to improve code quality, delivery speed, and operational outcomes (e.g., AI-assisted code review/refactoring, test strategy acceleration, incident/root-cause analysis support), while establishing consistent validation standards (secure coding, peer review, automated testing) and promoting reuse of effective patterns across the team.
  • Applies knowledge of tools within the Software Development Life Cycle toolchain, including enterprise-authorized AI-assisted development and automation capabilities, to improve the value realized by automation.

Required qualifications, capabilities, and skills

  • Formal training or certification on software engineering concepts and 5+ years applied experience
  • Expert in at least one distributed data processing framework (Spark). Expert in at least one cloud data Lakehouse platforms (AWS Data lake services or Databricks, if not Hadoop),
  • Expert in at least one scheduling/orchestration tools ( Airflow, alternatively AWS Step Functions or similar) & Expert with relational and NoSQL databases. Expert in data structures, data serialization formats (JSON, AVRO, Protobuf, or similar), and big-data storage formats (Parquet, Iceberg, or similar)
  • Hands-on professional experience in one or more programming language(s), including Java or Python, proficiency in Python, SQL, and at least one additional language (e.g. Java or Scala) for data engineering tasks
  • Hands-on experience utilizing Apache Spark for large-scale data processing, including developing and optimizing data pipelines, performing real-time and batch analytics, and leveraging Spark’s libraries for machine learning and data transformation to drive actionable business insights.
  • Proficiency in microservices architecture, serverless computing and distributed cluster computing tools such as Docker, Kubernetes etc. Experience in one or more data modelling techniques (Dimensional, Data Vault, Kimball, Inmon, etc.)
  • Experience with test-driven development (TDD) or behavior-driven development (BDD) practices, as well as working with continuous integration and continuous deployment (CI/CD) tools.
  • Experience organizing and leading design workshops, coding sessions, and hackathons to promote a culture of excellence and innovation in data engineering. Expertise in architecting reusable, future-ready design patterns that address diverse use cases across the organization.
  • Expertise in working with streaming platforms like Kafka, MQ etc.
  • Demonstrated experience leading effective use of approved AI-assisted software development tools (e.g., for coding, code review, test acceleration, troubleshooting) with the ability to set team expectations for validating AI outputs for correctness, performance, and security.
  • Strong understanding of responsible AI use in engineering workflows, including data sensitivity considerations, secure handling of inputs/outputs, and adherence to resiliency and security expectations; experience coaching engineers on safe, compliant adoption within delivery practices

Preferred qualifications, capabilities, and skills

  • Hands-on experience with Infrastructure as Code (IaC) tools, preferably Terraform; experience with AWS CloudFormation is also valued.
  • Proficiency in cloud-based data pipeline technologies such as Spinnaker or similar platforms.
  • Strong working knowledge of the Snowflake data platform.
  • Experience in budgeting and resource allocation for data engineering projects.
  • Proven ability to manage vendor relationships effectively.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Lead Software Engineer - Data Engineer
Lead Software Engineer - Data Engineer

JPMorganChase • Hyderabad

On-site
INR 3,500,000 - 6,000,000
Lead Software Engineer - Data Engineer
Lead Software Engineer - Data Engineer

JPMorgan Chase & Co. • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Lead Software Engineer Java Fullstack React AWS
Lead Software Engineer Java Fullstack React AWS

JPMorgan Chase & Co. • Bengaluru

On-site
INR 2,500,000 - 4,000,000
Sr Lead Software Engineer - Java/Python, Data Engineering
Sr Lead Software Engineer - Java/Python, Data Engineering

JPMorgan Chase & Co. • Hyderabad

On-site
INR 4,500,000 - 7,500,000
Lead Software Engineer
Lead Software Engineer

JPMorgan Chase & Co. • Bengaluru

On-site
INR 2,600,000 - 3,800,000
Lead Software Engineer Java, Python and Databricks
Lead Software Engineer Java, Python and Databricks

JPMorgan Chase Bank • Mumbai

On-site
INR 2,500,000 - 4,500,000
Lead Software Engineer - Python, AWS, BigData
Lead Software Engineer - Python, AWS, BigData

Fairygodboss • Bengaluru

On-site
INR 3,000,000 - 6,000,000
Lead Software Engineer - Java Fullstack, AWS
Lead Software Engineer - Java Fullstack, AWS

JPMorgan Chase & Co. • Mumbai

On-site
INR 1,800,000 - 3,000,000
Lead Software Engineer - Java
Lead Software Engineer - Java

JPMorgan Chase & Co. • Bengaluru

On-site
INR 1,600,000 - 2,600,000
Lead Software Engineer - Java, Cloud Platforms (AWS), Kafka,Generative AI
Lead Software Engineer - Java, Cloud Platforms (AWS), Kafka,Generative AI

JPMorgan Chase & Co. • Hyderabad

On-site
INR 3,500,000 - 7,000,000