Data Engineer -Databricks

Lever, Inc.

Coimbatore District

On-site

INR 1,500,000 - 2,500,000

Full time

8 days ago
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Insurance package
Learning and development resources

Job summary

ShyftLabs is seeking a Data Engineer to build scalable data pipelines on the Databricks Lakehouse Platform. You will leverage Python, PySpark, and SQL to ingest, transform, and export data via REST APIs, integrating multiple sources including S3 and databases.

Responsibilities include designing ETL/ELT pipelines, managing Delta Lake tables with Medallion Architecture, and optimizing Spark workloads while ensuring data quality and reliability across production systems.

Qualifications

  • Strong expertise in Python, PySpark, and Advanced SQL.
  • Hands-on experience with the Databricks Lakehouse Platform.
  • Good understanding of Unity Catalog, Delta Lake, Databricks Workflows/Jobs, Clusters, Notebooks, Repos, and Medallion Architecture.
  • Experience integrating with REST APIs for data ingestion and data export.
  • Strong knowledge of ETL/ELT development, batch processing, incremental loading, and data transformation.
  • Experience with data modeling (Star Schema, Snowflake Schema, Fact & Dimension tables, SCD concepts).
  • Understanding of data warehousing concepts and best practices.
  • Experience working with structured and semi-structured data (CSV, JSON, Parquet, Delta).
  • Knowledge of partitioning, file optimization, Spark performance tuning, and query optimization.
  • Experience with Git and CI/CD best practices

Responsibilities

  • Design, develop, and maintain scalable ETL/ELT pipelines using Databricks, PySpark, and SQL.
  • Integrate data from multiple sources, including databases, Amazon S3, files, and REST APIs.
  • Build data pipelines with Databricks Unity Catalog.
  • Implement business logic, data transformations, and dimensional data models.
  • Create, schedule, monitor, and optimize Databricks Jobs and Workflows.
  • Design and manage Delta Lake tables using Medallion Architecture (Bronze, Silver, Gold).
  • Ensure data quality through validations, error handling, logging, and monitoring.
  • Optimize Spark workloads for performance, scalability, and reliability.
  • Collaborate with cross-functional teams to deliver production-ready data solutions.

Skills

Python
PySpark
Advanced SQL
Databricks Lakehouse Platform
REST API integrations
ETL/ELT
Dimensional data modeling
Data warehousing concepts
Structured & semi-structured data
Git & CI/CD
Spark performance tuning

Tools

Databricks
Unity Catalog
Delta Lake
Databricks Workflows/Jobs

Job description

Position Overview

We are looking for a Data Engineer with hands-on experience in building scalable datapipelines and data engineering solutions on the Databricks Lakehouse Platform. The idealcandidate should have strong expertise in Python, PySpark, SQL, Databricks, AWS, andREST API integrations for data ingestion, managing large volumes of data, and data export

ShyftLabs is a growing data product company that was founded in early 2020 and works primarily with Fortune 500 companies. We deliver digital solutions built to help accelerate the growth of businesses in various industries, by focusing on creating value through innovation.

Job Responsibilities
  • Design, develop, and maintain scalable ETL/ELT pipelines using Databricks, PySpark, and SQL.
  • Integrate data from multiple sources, including databases, Amazon S3, files, andREST APIs.
  • Build data pipelines with Databricks Unity Catalog.
  • Implement business logic, data transformations, and dimensional data models.
  • Create, schedule, monitor, and optimize Databricks Jobs and Workflows.
  • Design and manage Delta Lake tables using Medallion Architecture (Bronze, Silver,Gold).
  • Ensure data quality through validations, error handling, logging, and monitoring.
  • Optimize Spark workloads for performance, scalability, and reliability.
  • Collaborate with cross-functional teams to deliver production-ready data solutions.
Basic Qualification
  • Strong expertise in Python, PySpark, and Advanced SQL.
  • Hands-on experience with the Databricks Lakehouse Platform.
  • Good understanding of Unity Catalog, Delta Lake, Databricks Workflows/Jobs, Clusters, Notebooks, Repos, and Medallion Architecture.
  • Experience integrating with REST APIs for data ingestion and data export.
  • Strong knowledge of ETL/ELT development, batch processing, incremental loading, and data transformation.
  • Experience with data modeling (Star Schema, Snowflake Schema, Fact & Dimension tables, SCD concepts).
  • Understanding of data warehousing concepts and best practices.
  • Experience working with structured and semi-structured data (CSV, JSON, Parquet, Delta).
  • Knowledge of partitioning, file optimization, Spark performance tuning, and query optimization.
  • Experience with Git and CI/CD best practices
Preferred Qualifications
  • 4+ years of experience in Data Engineering with 2+ years of hands-on Databricks experience.
  • Experience with Auto Loader, Spark Declarative pipelines, Kafka, Airflow, or dbt is aplus.
  • Databricks certification is an added advantage.

We are proud to offer a competitive salary alongside a strong insurance package. We pride ourselves on the growth of our employees, offering extensive learning and development resources.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data Engineer -Databricks
Data Engineer -Databricks

Shyftlabs • Coimbatore District

On-site
INR 900,000 - 1,500,000
Insurance package
Learning and development resources
Forward Deployed Engineer
Forward Deployed Engineer

ShyftLabs • Coimbatore District

On-site
INR 4,000,000 - 7,000,000
Insurance package
Learning & development resources
Technical Data Engineer (Databricks)
Technical Data Engineer (Databricks)

ShyftLabs • Coimbatore District

On-site
INR 1,500,000 - 2,500,000
Competitive salary
Insurance package
Forward- Lead Data Engineer Databricks
Forward- Lead Data Engineer Databricks

ShyftLabs • Coimbatore District

On-site
INR 1,500,000 - 2,800,000
Insurance package
Learning & development program
Lead Data Engineer - Databricks
Lead Data Engineer - Databricks

Lever, Inc. • Coimbatore District

On-site
INR 4,000,000 - 7,000,000
Insurance package
Learning and development resources
Lead Data Engineer - Databricks
Lead Data Engineer - Databricks

ShyftLabs • Coimbatore District

On-site
INR 1,800,000 - 2,400,000
Lead Data Engineer (Databricks)
Lead Data Engineer (Databricks)

Shyftlabs • Coimbatore District

On-site
INR 3,000,000 - 6,000,000
Competitive salary
Insurance package
Learning & development resources
Databricks - Data Engineer
Databricks - Data Engineer

Tredence Inc. • Bengaluru

On-site
INR 1,500,000 - 2,500,000
Azure Data Engineer
Azure Data Engineer

Texplorers • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Senior Data Engineer – Databricks
Senior Data Engineer – Databricks

Aspire, Jordan • India

On-site
INR 1,500,000 - 2,100,000