Senior Developer - Python and Spark

Citibank (Switzerland) AG

Pune District

On-site

Confidential

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Hybrid work model (3 days in office, 2
Global engineering network
Wellbeing and work-life balance

Job summary

Citibank (Switzerland) AG is seeking a Senior Data Engineer to design and build scalable data pipelines using Python and PySpark. You will integrate AI/ML models, ensure data quality, and automate CI/CD in a hybrid environment in Pune, India.

The role emphasizes production-grade code, peer mentorship, and collaboration with cross-functional teams to accelerate data-driven initiatives and platform intelligence.

Qualifications

  • Production-grade Python and Spark experience in data systems.
  • Advanced SQL for complex transformations and performance.
  • Designing and maintaining large ETL/ELT pipelines.
  • Proficient with Pandas, NumPy, SciPy for data prep.
  • DevOps practices with GitHub and CI/CD automation.
  • Linux environment familiarity.

Responsibilities

  • Design and build scalable ETL/ELT pipelines in Python and PySpark.
  • Develop production-grade code and mentor peers.
  • Integrate AI/ML models into data workflows.
  • Enforce data quality across pipelines and lifecycle.
  • Automate CI/CD for builds/deployments and improve DevOps maturity.
  • Collaborate with engineers and product owners to align tech with delivery goals.

Skills

Python
Apache Spark
SQL
ETL/ELT
Pandas
NumPy
SciPy
CI/CD
GitHub
Linux

Tools

GitHub
CI/CD pipelines
Linux usage

Job description

## Senior Developer - Python and SparkApplyremote type: Hybridlocations: Pune Maharashtra Indiatime type: Full timeposted on: Posted Todayjob requisition id: 26972902Citi is looking for a Senior Data Engineer to design and build the next-generation data processing and analytics platform that sits at the core of our AI-driven business innovation. In this role, you will architect scalable ETL/ELT pipelines using Python and PySpark, while also integrating Generative AI and Machine Learning models to push the boundaries of what our data systems can do. Working within a high-performing engineering team, you will shape how data flows, performs, and delivers value across the organisation at scale.## Responsibilities* Design and build highly scalable ETL/ELT pipelines in Python and PySpark to power reliable data ingestion, transformation, and integration at scale.* Develop production-grade code and elevate team capability through structured code reviews and hands-on peer mentorship.* Integrate Generative AI and Machine Learning models into data workflows, using AI development tools to deliver innovative solutions and resolve complex technical challenges.* Establish and enforce data quality standards across all pipelines, ensuring accuracy, consistency, and reliability of data throughout its lifecycle.* Implement an AI-first engineering approach by exploring and embedding Agentic AI workflows to increase development velocity and platform intelligence.* Automate CI/CD pipelines for builds and deployments, advancing DevOps maturity across the data platform.* Apply performance tuning and optimisation techniques to process large volumes of structured and semi-structured data efficiently and at speed.* Collaborate directly with engineers and product owners to align technical direction with delivery goals and accelerate outcomes.## Required Qualifications & Skills* Deep hands-on expertise in Python and Apache Spark, including Spark Core, Spark SQL, and DataFrames/Datasets, applied to production data systems.* Advanced SQL capability across complex query writing, data transformation logic, and query performance optimisation.* Demonstrated experience designing, building, and maintaining ETL/ELT pipelines for large-scale data ingestion and transformation.* Proficiency with Python data processing libraries including Pandas, NumPy, and SciPy for data wrangling and preprocessing tasks.* Practical experience with DevOps tools and CI/CD pipeline automation, with version control managed through GitHub.* Solid working knowledge of Linux environments, with familiarity across Windows systems.* Ability to collaborate across multidisciplinary teams to diagnose complex technical problems and deliver effective, lasting solutions.## Beneficial Skills & Qualifications* Working knowledge of Machine Learning and Generative AI, with direct exposure to AI development tooling in an engineering context.* Familiarity with Agentic AI workflows and their application to enhancing platform capabilities or development productivity.* Experience using Jupyter Notebook for rapid prototyping and iterative development of data solutions.## What We OfferAt Citi, you will work on technically challenging problems that matter, contributing to a platform that shapes how a global financial institution processes and derives intelligence from data. This is a role where technical depth, curiosity, and engineering rigour are genuinely valued.* Hybrid working model with 3 days in the office and 2 days working remotely, giving you flexibility and in-person collaboration.* Access to complex, large-scale engineering challenges at the intersection of data and AI, keeping your technical skills at the forefront of the industry.* Opportunities to grow into AI and platform architecture disciplines, supported by a team culture that invests in continuous technical development.* Exposure to a globally connected engineering network, collaborating with specialists across technology, data, and product functions.* A performance-driven environment where your contributions to platform innovation directly influence business outcomes at scale.* Wellbeing and work-life balance support, alongside a competitive benefits package designed to recognise the value you bring.Build data systems that power real decisions — apply now to join Citi's next-generation AI and data platform engineering team.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Developer - Python and Spark
Senior Developer - Python and Spark

Citi • Pune District

On-site
INR 1,800,000 - 3,000,000
Hybrid work model
Growth opportunities
Global engineering network
+2
Senior Developer - Python and Spark
Senior Developer - Python and Spark

Citigroup Inc. • Pune District

Hybrid
INR 2,800,000 - 4,200,000
Hybrid work arrangement
Exposure to AI/data platform projects
Growth in AI and platform architecture
Senior Developer - Python and Spark
Senior Developer - Python and Spark

Citi • Maharashtra

Hybrid
INR 1,200,000 - 1,800,000
Hybrid work schedule
Senior Big Data Engineer Pyspark Databricks - Vice President
Senior Big Data Engineer Pyspark Databricks - Vice President

Citibank (Switzerland) AG • Pune District

Hybrid
Confidential
PySpark Big Data Developer
PySpark Big Data Developer

Citibank (Switzerland) AG • Pune District

Hybrid
Confidential
Python Engineering AI Lead-Assistant Vice president
Python Engineering AI Lead-Assistant Vice president

Citibank (Switzerland) AG • Chennai District

Hybrid
Confidential
Data Platform Engineer (AI-Enabled)
Data Platform Engineer (AI-Enabled)

Citibank (Switzerland) AG • Pune District

Hybrid
Confidential
Senior Python Application Developer
Senior Python Application Developer

Citibank (Switzerland) AG • Pune District

On-site
Confidential
Applications Development Manager Java - C12 - PUNE
Applications Development Manager Java - C12 - PUNE

Citibank (Switzerland) AG • Pune District

On-site
Confidential
PySpark/Python developer - AVP
PySpark/Python developer - AVP

Citibank (Switzerland) AG • Pune District

On-site
Confidential