Senior Data Engineer (APAC Region)

Anrgi Tech Private Limited

Pune District

On-site

INR 2,000,000 - 2,800,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

ANRGI TECH Pvt. Ltd. in Pune is seeking a Senior Data Engineer to design, build, and operate scalable data pipelines on the Databricks platform.

You will develop end-to-end data workflows from ingestion to consumption, implement robust error handling, and ensure pipeline reliability and performance. The role involves migrating legacy ETL to modern ELT patterns, refactoring code to PySpark, and collaborating with data architects, analysts, and stakeholders.

Qualifications

  • Strong foundation in data engineering principles and ETL/ELT design patterns.
  • Hands-on PySpark development with DataFrames API and Spark SQL.
  • Experience with Databricks workspace, cluster management, notebooks, and job orchestration.
  • Familiarity with Delta Lake features, ACID transactions, and schema evolution.
  • Proficient Python programming for data processing and automation.

Responsibilities

  • Design, build, and operate scalable and reliable data pipelines on the Databricks platform.
  • Develop end-to-end data workflows from ingestion through transformation to consumption.
  • Implement robust error handling, monitoring, and alerting mechanisms.
  • Ensure data pipeline reliability, performance, and maintainability.
  • Collaborate with data architects, analysts, and business stakeholders.

Skills

Databricks
PySpark
Python
SQL
Data engineering
Data governance

Tools

Databricks
Delta Lake
Git
CI/CD

Job description

  • Design, build, and operate scalable and reliable data pipelines on the Databricks platform
  • Develop end-to-end data workflows from ingestion through transformation to consumption
  • Implement robust error handling, monitoring, and alerting mechanisms
  • Ensure data pipeline reliability, performance, and maintainability
  • Optimize pipeline performance through efficient Spark job design and cluster configuration
  • Manage and orchestrate complex data workflows using Databricks Jobs and workflows
Legacy Code Modernization
  • Refactor legacy code and data pipelines to PySpark for improved performance and scalability
  • Migrate traditional ETL processes to modern ELT patterns on Databricks
  • Assess existing codebases and identify opportunities for optimization and modernization
  • Ensure backward compatibility and data integrity during migration processes
  • Document refactoring approaches and create migration playbooks
  • Collaborate with stakeholders to minimize disruption during code transitions
  • Implement data quality checks and validation frameworks
  • Design and maintain Delta Lake tables with appropriate optimization strategies
  • Develop reusable code libraries and frameworks for common data engineering tasks
  • Follow software engineering best practices including version control, testing, and CI/CD
  • Participate in code reviews and provide constructive feedback to team members
  • Troubleshoot and resolve data pipeline issues in production environments
  • Work closely with data architects, analysts, and business stakeholders
  • Collaborate with Infrastructure (Infra), Applications (Apps), and Cyber teams
  • Share knowledge and best practices with Team *****
  • Mentor junior data engineers on PySpark and Databricks technologies
  • Document technical solutions and maintain comprehensive documentation
Essential Technical Skills
  • Data Engineering: Strong foundation in data engineering principles, ETL/ELT processes, and data pipeline design patterns
  • PySpark: Proven hands‑on experience developing data pipelines using PySpark, including DataFrames API, Spark SQL, and performance optimization
  • Databricks Platform: Practical experience with Databricks workspace, cluster management, notebooks, and job orchestration
  • Workspace AI Agent: Knowledge of Databricks Workspace AI Agent capabilities and integration
  • Delta Lake: Understanding of Delta Lake features including ACID transactions, schema evolution, and optimization techniques
  • Python: Strong Python programming skills for data processing and automation
Additional Technical Skills
  • SQL proficiency for data querying and transformation
  • Experience with cloud platforms (Azure, AWS, or GCP)
  • Understanding of data governance and security best practices
  • Knowledge of streaming data processing (Structured Streaming)
  • Familiarity with DevOps practices and CI/CD pipelines
  • Experience with version control systems (Git)
  • Understanding of data quality frameworks and testing methodologies
Professional Experience
  • Minimum 5 years in data engineering or related roles
  • At least 2-3 years of hands‑on experience with Databricks platform
  • Proven track record of refactoring legacy code to modern frameworks
  • Experience building and maintaining production data pipelines at scale
  • Background working across multiple data sources and formats
Required Certifications - mandatory to have at least one certification
Additional Certifications (Preferred)
  • Databricks Certified Associate Developer for Apache Spark
  • Cloud platform certifications (Azure Data Engineer Associate, AWS Certified Data Analytics, or Google Cloud Professional Data Engineer)
  • Relevant data engineering or big data certifications
Soft Skills
  • Strong problem-solving and analytical thinking abilities
  • Excellent communication skills to explain technical concepts clearly
  • Ability to work collaboratively in cross‑functional teams
  • Self‑motivated with strong attention to detail
  • Adaptable to changing priorities and technologies
  • Client‑focused mindset with commitment to quality delivery

ANRGI TECH Pvt. Ltd. is an Equal Opportunity Employer and does not discriminate on the basis of race or ethnicity, religion, sex, national origin, age, veteran disability or genetic information or any other reason prohibited by law in employment.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Data Engineer
Senior Data Engineer

Tekskills • Pune District, Bengaluru, Delhi

Hybrid
INR 1,500,000 - 2,100,000
Senior Data Engineer – Databricks
Senior Data Engineer – Databricks

Aspire, Jordan • India

On-site
INR 1,500,000 - 2,100,000
Databricks Developer
Databricks Developer

Kumaran Systems • Hyderabad

On-site
INR 1,500,000 - 2,800,000
Databricks Data Specialist - R01569707
Databricks Data Specialist - R01569707

Brillio • Bengaluru

On-site
INR 1,200,000 - 2,400,000
Databricks Engineer
Databricks Engineer

Impronics Technologies • Gurugram District

On-site
INR 4,000,000 - 7,000,000
Data Architect
Data Architect

Infosys • Bengaluru Urban

On-site
INR 1,200,000 - 2,000,000
Cutting-edge cloud and data technologies
Collaborative culture focusing on innovation
Cross-functional project visibility
Senior Data Engineer
Senior Data Engineer

Wissen • Bengaluru

On-site
INR 2,800,000 - 4,800,000
Databricks Developer
Databricks Developer

Kumaran • Hyderabad

On-site
INR 1,000,000 - 1,400,000
Senior Data Engineer
Senior Data Engineer

CoffeeBeans • Bengaluru

On-site
INR 1,800,000 - 2,400,000
Senior Data Engineer
Senior Data Engineer

Quess IT Solutions • Bengaluru

On-site
INR 3,000,000 - 6,000,000