Big Data - Senior Engineer

Iris Software, Inc.

Hinoba-an

On-site

PHP 1,990,000 - 3,316,000

Full time

14 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Iris Software, Inc. Noida, Gurugram and Pune is looking for a seasoned Big Data Engineer to design scalable data platforms using Spark, Hadoop, and Azure Databricks to support enterprise analytics.

You will architect Lakehouse solutions with Apache Hudi or Iceberg, establish robust ingestion and transformation pipelines, and mentor teammates on best practices in data engineering. This role offers exposure to cutting-edge tech and collaborative projects in a growth-focused environment.

Qualifications

  • Experience with data ingestion tools such as Sqoop and Hadoop ecosystems.
  • Proficient in Azure Databricks and Hadoop ecosystem fundamentals (HBase, Impala).
  • Strong SQL and Hive-based data processing skills for scalable pipelines.
  • Experience with Hudi/Iceberg lakehouse architectures and Spark-based workloads.
  • PySpark experience is a plus.

Responsibilities

  • Design scalable Big Data solutions using Spark, Hadoop, Azure Databricks, and modern data platform technologies.
  • Lead development of distributed data processing pipelines using Spark (Scala or PySpark) and Hadoop ecosystem technologies.
  • Design and optimize SQL and Hive-based data processing solutions to improve performance and scalability.
  • Architect and optimize Azure Databricks solutions supporting large-scale data engineering and analytics workloads.
  • Design and implement data Lakehouse solutions leveraging Apache Hudi or Apache Iceberg.
  • Establish data ingestion, transformation, validation, and reconciliation frameworks to improve data reliability.
  • Drive performance tuning initiatives across Spark jobs, Databricks workloads, Hive queries, and Hadoop processing environments.
  • Review data engineering solutions to ensure adherence to architecture, performance, and engineering standards.
  • Troubleshoot complex data processing, performance, and platform issues through detailed root cause analysis.
  • Mentor team members on Spark, Hadoop, Databricks, Hudi/Iceberg, SQL optimization, and Big Data engineering best practices.
  • Collaborate with various teams and stakeholders to support end-to-end data platform delivery.
  • Drive continuous improvement initiatives focused on scalability, performance, reliability, and operational efficiency.
  • Demonstrates strong ownership while driving Big Data Engineering excellence.
  • Collaborate effectively with various teams and business stakeholders to ensure smooth delivery.
  • Promotes quality-focused engineering through proactive validation, optimization, and continuous improvement.
  • Applies strong analytical thinking to evaluate complex data engineering and platform challenges.
  • Demonstrate adaptability while managing evolving technologies, data ecosystems, and business requirements.
  • Communicates effectively regarding delivery status, risks, dependencies, and improvement opportunities.
  • Maintains high attention to detail across data architecture, processing design, testing, and implementation activities.
  • Encourages continuous improvement in data engineering practices and platform operations.
  • Supports knowledge sharing and mentoring to strengthen team capabilities.
  • Balances scalability, performance, reliability, and business priorities while driving delivery excellence.

Skills

Data Ingestion Tools
Hadoop
Azure Databricks
Hadoop Ecosystem Fundamentals
Hive
Scala
Apache Hudi
PySpark

Job description

Select how often (in days) to receive an alert: Create Alert

Why Join Iris?
Are you ready to do the best work of your career at one ofIndia’s Top 25 Best Workplaces in IT industry? Do you want to grow in an award-winning culture thattruly values your talent and ambitions?
Join Iris Software — one offastest-growing IT services companies— whereyou own and shape your success story.

About Us
At Iris Software, our vision is to be our client’s most trusted technology partner, and the first choice for the industry’s top professionals to realize their full potential.

With over 4,300 associates across India, U.S.A, and Canada, we help our enterprise clients thrive with technology-enabled transformation across financial services, healthcare, transportation & logistics, and professional services.

Our work covers complex, mission-critical applications with the latest technologies, such as high-value complex Application & Product Engineering, Data & Analytics, Cloud, DevOps, Data & MLOps, Quality Engineering, and Business Automation.

Working with Us
At Iris, every role is more than a job — it’s a launchpad for growth.

Our Employee Value Proposition, “Build Your Future. Own Your Journey.”reflects our belief that people thrive when they have ownership of their career and the right opportunities to shape it.

We foster a culture where your potential is valued, your voice matters, and your work creates real impact. With cutting-edge projects, personalized career development, continuous learning and mentorship, we support you to grow and become your best — both personally and professionally.

Curious what it’s like to work at Iris? Head to this video for an inside look at the people, the passion, and the possibilities. Watch it here .

Job Description

Experience: 6-7 Years

Location: Noida, Gurugram, Pune

Mandatory Skills:

Data Ingestion Tools (Sqoop), Hadoop (HDFS + YARN), Azure Databricks, Hadoop Ecosystem Fundamentals (HBase + Impala), Hive, Scala, Apache Hudi

Additional Skills:

PySpark

Key Responsibilities

  • Design scalable Big Data solutions using Apache Spark, Hadoop, Azure Databricks, and modern data platform technologies.
  • Define data processing architecture, transformation strategies, and engineering standards aligned with business objectives.
  • Lead development of distributed data processing pipelines using Spark (Scala or PySpark) and Hadoop ecosystem technologies.
  • Design and optimize SQL and Hive-based data processing solutions to improve performance and scalability.
  • Architect and optimize Azure Databricks solutions supporting large-scale data engineering and analytics workloads.
  • Design and implement data Lakehouse solutions leveraging Apache Hudi or Apache Iceberg.
  • Establish data ingestion, transformation, validation, and reconciliation frameworks to improve data reliability.
  • Drive performance tuning initiatives across Spark jobs, Databricks workloads, Hive queries, and Hadoop processing environments.
  • Review data engineering solutions to ensure adherence to architecture, performance, and engineering standards.
  • Troubleshoot complex data processing, performance, and platform issues through detailed root cause analysis.
  • Mentor team members on Spark, Hadoop, Databricks, Hudi/Iceberg, SQL optimization, and Big Data engineering best practices.
  • Collaborate with various teams and stakeholders to support end-to-end data platform delivery.
  • Drive continuous improvement initiatives focused on scalability, performance, reliability, and operational efficiency.
  • Demonstrates strong ownership while driving Big Data Engineering excellence.
  • Collaborate effectively with various teams and business stakeholders to ensure smooth delivery.
  • Promotes quality-focused engineering through proactive validation, optimization, and continuous improvement.
  • Applies strong analytical thinking to evaluate complex data engineering and platform challenges.
  • Demonstrate adaptability while managing evolving technologies, data ecosystems, and business requirements.
  • Communicates effectively regarding delivery status, risks, dependencies, and improvement opportunities.
  • Maintains high attention to detail across data architecture, processing design, testing, and implementation activities.
  • Encourages continuous improvement in data engineering practices and platform operations.
  • Supports knowledge sharing and mentoring to strengthen team capabilities.
  • Balances scalability, performance, reliability, and business priorities while driving delivery excellence.

Programming Language - Scala - Scala

Perks and Benefits for IrisiansIris provides world-class benefits for a personalized employee experience. These benefits are designed to support financial, health and well-being needs of Irisians for a holistic professional and personal growth. Click here to view the benefits.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Databricks - Manager
Databricks - Manager

Iris Software, Inc. • Hinoba-an

On-site
PHP 600,000 - 1,200,000
Java - Lead
Java - Lead

Iris Software, Inc. • Hinoba-an

Hybrid
PHP 1,653,000 - 2,645,000
Java FullStack React- Lead
Java FullStack React- Lead

Iris Software, Inc. • Hinoba-an

On-site
PHP 1,592,000 - 2,785,000
Java - Senior Engineer
Java - Senior Engineer

Iris Software, Inc. • Hinoba-an

On-site
PHP 793,000 - 1,190,000
Senior Big Data Engineer: Spark, Hadoop & Lakehouse
Senior Big Data Engineer: Spark, Hadoop & Lakehouse

Iris Software, Inc. • Hinoba-an

On-site
PHP 1,990,000 - 3,316,000
Senior Business Analyst - Domain
Senior Business Analyst - Domain

Iris Software, Inc. • Hinoba-an

On-site
PHP 595,000 - 793,000
QA Manual - Senior Engineer
QA Manual - Senior Engineer

Iris Software, Inc. • Hinoba-an

On-site
PHP 380,000 - 800,000
QA Automation - Lead
QA Automation - Lead

Iris Software, Inc. • Hinoba-an

On-site
PHP 1,326,000 - 2,321,000
UI React - Senior Engineer
UI React - Senior Engineer

Iris Software, Inc. • Hinoba-an

On-site
PHP 793,000 - 1,587,000
World-class benefits
QA Automation - Intern
QA Automation - Intern

Iris Software, Inc. • Hinoba-an

On-site
PHP 400,000 - 900,000
Benefits for Irisians