Spark Java Data Engineer

Infosys

Bengaluru

On-site

INR 1,200,000 - 1,800,000

Full time

9 days ago
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Infosys in Bengaluru seeks an experienced Spark-Java engineer to design end-to-end data processing jobs, build ETL pipelines, and optimize Spark performance.

You will work with Databricks DBX to implement scalable data workflows, ensure data quality, and collaborate with stakeholders to deliver reliable data solutions with proper tests and monitoring.

Qualifications

  • 3–5 years of hands‑on experience in software/data engineering roles.
  • Bachelor’s or Master’s degree: BTECH, MTECH, MCA, MSC.
  • Strong programming experience in Java with solid OOP understanding.
  • Practical experience with Apache Spark for batch data processing.
  • Experience building and supporting ETL pipelines and data transformations.
  • Working knowledge of DBX for developing and running data workflows.

Responsibilities

  • Design, develop, and maintain Spark-based data processing jobs using Java aligned to business and technical requirements.
  • Build and enhance ETL workflows, ensuring accuracy, completeness, and consistency of processed datasets.
  • Implement integrations and workflows on DBX, supporting scalable execution and operational reliability.
  • Optimize Spark jobs for performance (partitioning, caching, shuffle tuning) and cost efficiency.
  • Write clean, maintainable code with appropriate logging, error handling, and unit/integration tests.
  • Troubleshoot production issues, perform root-cause analysis, and implement preventive fixes.
  • Collaborate with cross-functional teams to refine requirements, plan deliveries, and ensure smooth releases.
  • Document technical designs, data flows, and operational runbooks to support long-term maintainability.

Skills

Java programming
OOP concepts

Education

Bachelor’s or Master’s degree: BTECH, MTECH, MCA, MSC

Tools

Apache Spark
DBX
CI/CD
ETL workflows

Job description

Spark-Java, DBX Spark-Java, Databricks Preferred Qualifications:
  • Experience designing end-to-end data pipelines including ingestion, transformation, validation, and publishing layers.
  • Strong understanding of distributed processing concepts and Spark internals for performance tuning and stability.
  • Exposure to CI/CD practices for data/engineering workflows and disciplined release management.
  • Experience with production monitoring, alerting, and operational support for data pipelines.
  • Proven ability to collaborate with stakeholders, communicate trade-offs, and deliver within timelines.
Key Responsibilities:
  • Design, develop, and maintain Spark-based data processing jobs using Java aligned to business and technical requirements.
  • Build and enhance ETL workflows, ensuring accuracy, completeness, and consistency of processed datasets.
  • Implement integrations and workflows on DBX, supporting scalable execution and operational reliability.
  • Optimize Spark jobs for performance (partitioning, caching, shuffle tuning) and cost efficiency.
  • Write clean, maintainable code with appropriate logging, error handling, and unit/integration tests.
  • Troubleshoot production issues, perform root-cause analysis, and implement preventive fixes.
  • Collaborate with cross-functional teams to refine requirements, plan deliveries, and ensure smooth releases.
  • Document technical designs, data flows, and operational runbooks to support long-term maintainability.
Minimum Qualifications:
  • 3–5 years of hands‑on experience in software/data engineering roles.
  • Bachelor’s or Master’s degree: BTECH, MTECH, MCA, MSC.
  • Strong programming experience in Java with solid understanding of OOP and coding best practices.
  • Practical experience with Apache Spark for batch data processing.
  • Experience building and supporting ETL pipelines and data transformations.
  • Working knowledge of DBX for developing and running data workflows.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Spark-Java, DBX
Spark-Java, DBX

Infosys • Bengaluru

On-site
INR 1,500,000 - 2,300,000
Data Engineer - Spark and Scala
Data Engineer - Spark and Scala

Infosys • Bengaluru

On-site
INR 1,200,000 - 2,400,000
Spark-Scala, Databricks
Spark-Scala, Databricks

Infosys • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Spark
Spark

Infosys • Bengaluru

On-site
INR 900,000 - 1,400,000
Spark Data Engineer
Spark Data Engineer

Infosys • Bengaluru

On-site
INR 1,800,000 - 2,400,000
PySpark Developer - Data Engineering
PySpark Developer - Data Engineering

Infosys • Bengaluru

On-site
INR 1,500,000 - 2,100,000
Restaurant d'entreprise
Indemnités de stage/alternance
Senior Java Developer
Senior Java Developer

SG Analytics • Chennai District

On-site
INR 1,500,000 - 4,000,000
Big Data Engineer (Spark, Scala)
Big Data Engineer (Spark, Scala)

Dataconsol Private Limited • Hyderabad

On-site
INR 1,800,000 - 3,000,000
Sr. Data Engineer - Databricks
Sr. Data Engineer - Databricks

Anblicks Inc. • Hyderabad

On-site
INR 3,000,000 - 5,500,000
DATA Engineer (JAVA SPARK)
DATA Engineer (JAVA SPARK)

Wissen Technology • Pune District

Hybrid
INR 1,400,000 - 2,200,000