Data Engineer with Scala and Spark/ Pyspark

Capco

Bengaluru

Hybrid

INR 1,800,000 - 3,200,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Hybrid work model

Job summary

Capco is seeking a Data Engineer to design and optimize large-scale data processing systems. You will build high-performance pipelines handling tens of millions of records daily, leveraging Scala, Apache Spark, and Java.

You will integrate Native APIs and microservices, collaborate with cross-functional teams, and drive data quality, governance, and performance improvements in a hybrid Capco office setup.

Qualifications

  • Strong hands-on experience in Scala, Apache Spark, and Java.
  • Proven experience building large-scale distributed data processing systems handling tens of millions of records daily.
  • Expertise in developing and consuming Native APIs and microservices.
  • Strong understanding of Spark architecture, performance tuning, and optimization techniques.
  • Experience with SQL and relational/non-relational databases.

Responsibilities

  • Design, develop, and optimize large-scale data processing applications using Scala, Apache Spark, and Java.
  • Build and maintain high-performance data pipelines capable of processing 80-90 million records daily.
  • Develop and integrate Native APIs and data services to support business-critical applications.
  • Collaborate with cross-functional teams to gather requirements and deliver robust data engineering solutions.
  • Optimize Spark jobs, data workflows, and distributed computing processes for high throughput and low latency.
  • Implement best practices for data quality, monitoring, governance, and operational excellence.
  • Troubleshoot and resolve performance bottlenecks across data processing and ingestion pipelines.
  • Participate in code reviews and contribute to engineering standards and architecture decisions.

Skills

Scala
Apache Spark
Java
Native APIs
Microservices
Data modeling
ETL/ELT

Tools

Kafka
Airflow
Hadoop ecosystem

Job description

About Us

"Capco, a Wipro company, is a global technology and management consulting firm. Awarded with Consultancy of the year in the British Bank Award and has been ranked Top 100 Best Companies for Women in India 2022 by Avtar & Seramount. With our presence across 32 cities across globe, we support 100+ clients acrossbanking, financial and Energy sectors. We are recognized for our deep transformation execution and delivery."\>

WHY JOIN CAPCO?

You will work on engaging projects with the largest international and local banks, insurance companies, payment service providers and other key players in the industry. The projects that will transform the financial services industry.

MAKE AN IMPACT

Innovative thinking, delivery excellence and thought leadership to help our clients transform their business. Together with our clients and industry partners, we deliver disruptive work that is changing energy and financial services.

#BEYOURSELFATWORK

Capco has a tolerant, open culture that values diversity, inclusivity, and creativity.

CAREER ADVANCEMENT

With no forced hierarchy at Capco, everyone has the opportunity to grow as we grow, taking their career into their own hands.

DIVERSITY & INCLUSION

We believe that diversity of people and perspective gives us a competitive advantage.

MAKE AN IMPACT

Job Title: Data Engineer

Position: Data Engineer
Experience: 4+ Years
Work Mode: Hybrid (Capco Office)

Key Responsibilities

  • Design, develop, and optimize large-scale data processing applications using Scala, Apache Spark, and Java.
  • Build and maintain high-performance data pipelines capable of processing 80-90 million records daily with a focus on scalability, reliability, and efficiency.
  • Develop and integrate Native APIs and data services to support business-critical applications and analytics platforms.
  • Collaborate with cross-functional teams to gather requirements and deliver robust data engineering solutions.
  • Optimize Spark jobs, data workflows, and distributed computing processes to ensure high throughput and low latency.
  • Implement best practices for data quality, monitoring, governance, and operational excellence.
  • Troubleshoot and resolve performance bottlenecks across data processing and ingestion pipelines.
  • Participate in code reviews and contribute to engineering standards, architecture decisions, and continuous improvement initiatives.

Required Skills & Experience

  • Strong hands-on experience in Scala, Apache Spark, and Java.
  • Proven experience building and supporting large-scale distributed data processing systems handling tens of millions of records daily.
  • Expertise in developing and consuming Native APIs and microservices.
  • Strong understanding of Spark architecture, performance tuning, partitioning, caching, and optimization techniques.
  • Experience with data modeling, ETL/ELT processes, and large-scale batch and streaming data pipelines.
  • Solid understanding of distributed systems, concurrency, and high-volume data processing.
  • Experience with SQL and relational/non-relational databases.
  • Strong debugging, analytical, and problem-solving skills.

Preferred Qualifications

  • Experience working in enterprise-scale data environments within Financial Services, Payments, or FinTech domains.
  • Exposure to cloud platforms (AWS, Azure, or GCP) and containerized deployments.
  • Familiarity with Kafka, Airflow, Hadoop ecosystem, or similar big data technologies.
  • Experience with CI/CD pipelines, DevOps practices, and Agile methodologies.

What We\'re Looking For

  • Engineers passionate about solving complex data challenges at scale.
  • Professionals who can design highly performant systems capable of processing and transforming massive datasets efficiently.
  • Team players who thrive in fast-paced, collaborative environments and take ownership of end-to-end delivery.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer with Scala and Spark/ Pyspark
Data Engineer with Scala and Spark/ Pyspark

Capco • India

Hybrid
INR 1,500,000 - 2,100,000
Senior Data Engineering - PSM-Libra
Senior Data Engineering - PSM-Libra

Capco • Bengaluru

On-site
INR 2,000,000 - 4,000,000
Senior Data Analyst
Senior Data Analyst

Capco • India

Hybrid
INR 1,200,000 - 1,800,000
Hybrid work model
Google Cloud Platform Data Engineer
Google Cloud Platform Data Engineer

Capco Technologies Pvt Ltd • Bengaluru

Hybrid
INR 1,400,000 - 2,100,000
AWS Data Engineer
AWS Data Engineer

Capco • Bengaluru

On-site
INR 3,000,000 - 5,500,000
Data Analyst
Data Analyst

Capco • Bengaluru

On-site
INR 900,000 - 1,500,000
Python with Spark Developer (5.1-7 years)-Chennai
Python with Spark Developer (5.1-7 years)-Chennai

Triwill Group • Chennai District

On-site
INR 1,200,000 - 2,400,000
Python with Spark Developer (5.1-7 years)-Chennai
Python with Spark Developer (5.1-7 years)-Chennai

Capco • Chennai District

On-site
INR 1,200,000 - 1,800,000
GCP Data Engineer
GCP Data Engineer

Capco • Bengaluru

Hybrid
INR 1,200,000 - 1,800,000
Data Scientist with R programming
Data Scientist with R programming

Capco • Mumbai

On-site
INR 1,200,000 - 1,800,000