Databricks

Infosys

Bengaluru

On-site

INR 3,000,000 - 5,500,000

Full time

9 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Infosys is seeking a senior data engineer in Bengaluru to design and deliver scalable data solutions using Databricks and PySpark. You will work with consultants, data engineers, and stakeholders to build reliable pipelines and drive high-quality releases.

The role emphasizes ownership, guidance, and growth in modern lakehouse practices within a collaborative environment.

Qualifications

  • Bachelors or Masters degree in BTECH, MTECH, MCA, or MSC (or equivalent).
  • 23 years of hands-on experience with Databricks in data engineering or analytics projects.
  • Strong experience in PySpark for transformations and distributed data processing.
  • Solid understanding of data pipeline concepts, data modeling basics, and structured/semi-structured data handling.
  • Ability to debug and troubleshoot Spark jobs and collaborate within delivery teams.

Responsibilities

  • Develop and maintain data pipelines and transformations using Databricks and PySpark.
  • Implement scalable ETL/ELT workflows to ingest, cleanse, and curate data for downstream analytics and reporting.
  • Optimize Spark jobs for performance and cost by tuning partitions, caching, joins, and cluster configurations.
  • Build reusable notebooks and modular code to support consistent development and easier maintenance.
  • Perform data validation, reconciliation, and quality checks to ensure accuracy and reliability of datasets.
  • Collaborate with cross-functional teams to gather requirements, clarify data definitions, and deliver aligned solutions.
  • Support deployments and production operations by troubleshooting failures, analyzing logs, and resolving incidents.
  • Contribute to documentation, coding standards, and best practices for Databricks-based development.

Education

Bachelors/Masters in BTech/MTech/MCA/MSC

Tools

Databricks
PySpark
Delta Lake
Spark SQL
Data Modeling
Workflow Orchestration

Job description

Job DescriptionJoin a team where data engineering meets real-world impact. In this role, youll help design and deliver scalable data solutions using Databricks and PySpark, enabling teams to turn raw data into trusted, analytics-ready assets. Youll collaborate closely with consultants, data engineers, and stakeholders to understand business needs, build reliable pipelines, and support high-quality releases. This is a great opportunity for someone with 23 years of experience who enjoys solving data challenges, improving performance, and learning modern lakehouse practices. If youre motivated by clean engineering, continuous improvement, and working in a collaborative environment where your contributions are visible and valued, this role offers the right mix of ownership, guidance, and growth.

Roles ResponsibilitiesKey Responsibilities:
  • Develop and maintain data pipelines and transformations using Databricks and PySpark.
  • Implement scalable ETL/ELT workflows to ingest, cleanse, and curate data for downstream analytics and reporting.
  • Optimize Spark jobs for performance and cost by tuning partitions, caching, joins, and cluster configurations.
  • Build reusable notebooks and modular code to support consistent development and easier maintenance.
  • Perform data validation, reconciliation, and quality checks to ensure accuracy and reliability of datasets.
  • Collaborate with cross-functional teams to gather requirements, clarify data definitions, and deliver aligned solutions.
  • Support deployments and production operations by troubleshooting failures, analyzing logs, and resolving incidents.
  • Contribute to documentation, coding standards, and best practices for Databricks-based development.
Minimum Qualifications:
  • Bachelors or Masters degree in BTECH, MTECH, MCA, or MSC (or equivalent).
  • 23 years of hands-on experience working with Databricks in data engineering or analytics engineering projects.
  • Strong experience in PySpark for building transformations and distributed data processing.
  • Solid understanding of data pipeline concepts, data modeling basics, and structured/semi-structured data handling.
  • Ability to debug and troubleshoot Spark jobs and collaborate effectively within delivery teams.
Technical RequirementETL, PYSPARK, DATABRICKS, Delta Lake, Spark SQL, Data Modeling, Workflow Orchestration, Performance Tuning
  • Experience with Spark optimization techniques and practical performance tuning in Databricks environments.
  • Familiarity with Delta Lake concepts such as ACID tables, schema evolution, and incremental processing patterns.
  • Exposure to orchestrating workflows and managing dependencies for end-to-end pipeline execution.
  • Experience working in agile delivery models with strong ownership of tasks, timelines, and quality outcomes.

Strong communication skills to translate requirements into implementable data solutions and clearly document outcomes.

Educational Requirement

MCA,MSc,MTech,Bachelor of Engineering,BTech

Preferred Skills

Technology->Big Data - Data Processing->PySpark,Technology->Data Engineering->Databricks

Service Line

Data Analytics Unit

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Spark-Scala, Databricks
Spark-Scala, Databricks

Infosys • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Databricks Developer
Databricks Developer

Kumaran Systems • Hyderabad

On-site
INR 1,500,000 - 2,800,000
Databricks Data Engineer
Databricks Data Engineer

Adesso SE • Pune District

On-site
INR 1,400,000 - 2,800,000
Senior Data Engineer – Databricks
Senior Data Engineer – Databricks

Aspire, Jordan • India

On-site
INR 1,500,000 - 2,100,000
Databricks Engineer
Databricks Engineer

Impronics Technologies • Gurugram District

On-site
INR 4,000,000 - 7,000,000
Databricks Data Specialist - R01569707
Databricks Data Specialist - R01569707

Brillio • Bengaluru

On-site
INR 1,200,000 - 2,400,000
Databricks Data Engineer
Databricks Data Engineer

NextGen Digital Solutions - NDS • Mumbai

On-site
INR 900,000 - 1,500,000
Data Engineer- Databricks
Data Engineer- Databricks

r3 Consultant • Bengaluru

On-site
INR 1,000,000 - 2,000,000
Databricks Data Engineer
Databricks Data Engineer

Infosys • Ahmedabad District, Chennai District, Bengaluru

On-site
INR 800,000 - 1,400,000
Databricks - Data Engineer
Databricks - Data Engineer

Tredence Inc. • Bengaluru

On-site
INR 1,500,000 - 2,500,000