Big data/Python/Databricks Engineer Engineer

Citi

Chennai District

Hybrid

INR 1,500,000 - 2,100,000

Full time

Just now
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Hybrid working model
Learning opportunities
Career progression
Competitive compensation

Job summary

Citi is looking for an Applications Development Intermediate Programmer Analyst to design, build, and maintain large-scale data engineering solutions that support critical reporting and analytics functions across a global financial institution.

In this role, you will develop and optimize data pipelines, work across distributed computing platforms, and contribute to the full lifecycle of data application delivery.

Qualifications

  • 5 to 8 years of experience in data engineering, big data development, or software application development with a focus on large-scale data platforms.
  • Hands-on development experience using Python and PySpark for data ingestion, transformation, and pipeline orchestration at scale.
  • Expert in ETL and Oracle DB, SQL
  • Knowledge in Databricks
  • Practical knowledge of Hadoop ecosystem components including HDFS, Hive, and Hadoop cluster operations, with familiarity with Ozone storage.
  • Working experience in Linux environments, including scripting, job scheduling, and process management.
  • Experience building or supporting Tableau dashboards and reports to deliver data insights to business stakeholders.
  • Familiarity with RStudio for statistical analysis or data exploration in a data engineering context.
  • Deep expertise in Large Language Models (LLMs) including OpenAI, Gemini, Claude, Llama, and local/open‑source models
  • Bachelor's degree or equivalent experience in a relevant technical discipline.

Responsibilities

  • Build and maintain scalable data pipelines using Python and PySpark to process and transform large volumes of structured and unstructured data across distributed platforms.
  • Develop and optimize data workflows on Hadoop-based ecosystems, including HDFS, Hive ensuring reliable data availability for downstream reporting and analytics.
  • Conduct feasibility studies, time and cost estimates, and technical planning activities to support data engineering delivery across business areas.
  • Monitor and manage all phases of the data application development lifecycle, from analysis and design through to testing, implementation, and production support.
  • Collaborate with analytics and reporting teams to build and maintain Tableau dashboards and translating raw data into actionable business insights.
  • Administer and troubleshoot data processes within Linux environments, ensuring stability, performance, and operational continuity.
  • Assess risk across data engineering decisions, ensuring solutions align with security, data governance, and compliance standards.
  • Experience in managing and implementing successful projects
  • Working knowledge of consulting/project management techniques/methods
  • Ability to work under pressure and manage deadlines or unexpected changes in expectations or requirements

Skills

Python
PySpark
SQL
Oracle DB
Hadoop
Hive
Databricks
Linux
Tableau
RStudio
LLMs

Education

Bachelor's degree or equivalent

Tools

Hadoop

Job description

Citi is looking for an Applications Development Intermediate Programmer Analyst to design, build, and maintain large-scale data engineering solutions that support critical reporting and analytics functions across a global financial institution. In this role, you will develop and optimize data pipelines, work across distributed computing platforms, and contribute to the full lifecycle of data application delivery. This is an opportunity to work within a high-impact technology team that operates at the core of Citi's data infrastructure.

Responsibilities
  • Build and maintain scalable data pipelines using Python and PySpark to process and transform large volumes of structured and unstructured data across distributed platforms.
  • Develop and optimize data workflows on Hadoop-based ecosystems, including HDFS, Hive ensuring reliable data availability for downstream reporting and analytics.
  • Conduct feasibility studies, time and cost estimates, and technical planning activities to support data engineering delivery across business areas.
  • Monitor and manage all phases of the data application development lifecycle, from analysis and design through to testing, implementation, and production support.
  • Collaborate with analytics and reporting teams to build and maintain Tableau dashboards and translating raw data into actionable business insights.
  • Administer and troubleshoot data processes within Linux environments, ensuring stability, performance, and operational continuity.
  • Assess risk across data engineering decisions, ensuring solutions align with security, data governance, and compliance standards.
  • Experience in managing and implementing successful projects
  • Working knowledge of consulting/project management techniques/methods
  • Ability to work under pressure and manage deadlines or unexpected changes in expectations or requirements
Required Qualifications & Skills
  • 5 to 8 years of experience in data engineering, big data development, or software application development with a focus on large-scale data platforms.
  • Hands‑on development experience using Python and PySpark for data ingestion, transformation, and pipeline orchestration at scale.
  • Expert in ETL and Oracle DB, SQL
  • Knowledge in Data bricks
  • Practical knowledge of Hadoop ecosystem components including HDFS, Hive, and Hadoop cluster operations, with familiarity with Ozone storage.
  • Working experience in Linux environments, including scripting, job scheduling, and process management.
  • Experience building or supporting Tableau dashboards and reports to deliver data insights to business stakeholders.
  • Familiarity with RStudio for statistical analysis or data exploration in a data engineering context.
  • Deep expertise in Large Language Models (LLMs) including OpenAI, Gemini, Claude, Llama, and local/open‑source models
  • Bachelor's degree or equivalent experience in a relevant technical discipline.
Beneficial Skills & Qualifications
  • Experience working with cloud-native data platforms or migrating workloads from on‑premise Hadoop environments to modern data platforms.
  • Knowledge of data governance practices, metadata management, or data quality frameworks within large-scale environments.
  • Familiarity with consulting or project management techniques applied within a technology delivery context.
What We Offer

At Citi, you will work within a collaborative and performance‑driven global technology team where your contributions directly support the data infrastructure underpinning one of the world's leading financial institutions. This role offers technical depth, meaningful delivery ownership, and the opportunity to grow your expertise across big data and analytics platforms.

  • Hybrid working model with 2 days in the office and 3 days working remotely, supporting a sustainable work‑life balance.
  • Opportunity to work on large-scale, real‑world data engineering challenges using modern big data tools and platforms at enterprise scale.
  • Access to continuous learning and development resources to deepen technical skills across data engineering, analytics, and cloud technologies.
  • Autonomy to make technical decisions and operate with limited direct supervision, with the opportunity to act as a subject matter expert to senior stakeholders.
  • A structured career pathway with clear progression opportunities within Citi's global technology organization.
  • Competitive financial well‑being and benefits package aligned to your location and level.
Job Family Group

Technology

Job Family

Applications Development

Time Type

Full time

Citi is an equal opportunity employer, and qualified candidates will receive consideration without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, disability, status as a protected veteran, or any other characteristic protected by law.

If you are a person with a disability and need a reasonable accommodation to use our search tools and/or apply for a career opportunity review Accessibility at Citi.

View Citi’s EEO Policy Statement and the Know Your Rights poster.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Big data/Python/Databricks Engineer Engineer
Big data/Python/Databricks Engineer Engineer

Citigroup Inc. • Chennai District

Hybrid
INR 600,000 - 1,000,000
Hybrid working model
Learning resources
Career progression opportunities
Big data/Python/Databricks Engineer Engineer
Big data/Python/Databricks Engineer Engineer

Citi • Chennai District

Hybrid
INR 1,500,000 - 2,500,000
Hybrid work model
Continuous learning resources
Structured career pathway
Data Engineer - Big Data, Python
Data Engineer - Big Data, Python

Citigroup Inc. • Chennai District

On-site
INR 900,000 - 1,800,000
Data Engineer - Big Data Python
Data Engineer - Big Data Python

Citi • Chennai District

On-site
INR 1,200,000 - 2,100,000
Big Data / PySpark Engineering Lead - Vice President
Big Data / PySpark Engineering Lead - Vice President

Citigroup Inc. • Pune District

On-site
INR 1,800,000 - 2,500,000
Data Engineer Python Developer
Data Engineer Python Developer

Citi • Pune District

On-site
INR 1,200,000 - 1,800,000
Data Engineer - Assistant Vice President
Data Engineer - Assistant Vice President

Citi • Maharashtra

On-site
INR 1,200,000 - 1,800,000
Senior Developer - Python and Spark
Senior Developer - Python and Spark

Citigroup Inc. • Pune District

Hybrid
INR 2,800,000 - 4,200,000
Hybrid work arrangement
Exposure to AI/data platform projects
Growth in AI and platform architecture
Senior Developer - Python and Spark
Senior Developer - Python and Spark

Citi • Maharashtra

Hybrid
INR 1,200,000 - 1,800,000
Hybrid work schedule
Big Data Engineer - Python and Spark
Big Data Engineer - Python and Spark

Citigroup • Pune District

On-site
INR 500,000 - 900,000