Senior Data Engineer

bakertilly

Bengaluru

On-site

INR 1,400,000 - 2,200,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

BTVK Advisory is seeking a senior Data Engineer to design, build, and maintain robust data pipelines on Azure Databricks across medallion layers, enabling analytics and AI-ready datasets.

You will work with Unity Catalog, ADLS Gen2, Delta Lake, and cross-cloud data movement, focusing on reliability, governance, and performance in a Bengaluru-based, fast-paced environment.

Qualifications

  • Bachelor's degree required, preferably in Information Technology, Computer Science, Data Science, Analytics, or Statistics.
  • 5-8 years of hands-on experience in data engineering, data integration, or analytics engineering, including substantial time delivering on Databricks.
  • Deep, hands-on experience with Azure Databricks and its critical components, including workspaces, clusters/SQL warehouses, Unity Catalog, and ADLS Gen2 integration.
  • Strong proficiency in Spark/PySpark and SQL, including the ability to develop complex queries and perform advanced data transformations.
  • Hands-on experience ingesting structured, semi-structured, and unstructured data via custom streams (Structured Streaming / Auto Loader), co

Responsibilities

  • Design, build, and maintain robust, scalable ETL/ELT pipelines on Azure Databricks to ingest, transform, and curate data across the medallion architecture - from raw/bronze landing through silver and gold to platinum/consumption-ready layers - ensuring reliability, performance, and timeliness.
  • Build and manage Unity Catalog assets (catalogs, schemas, tables, and Volumes) and configure or mount external locations/Volumes for governed access to cloud storage such as ADLS Gen2.
  • Design and implement audit, control, and operational-metadata tables that capture pipeline run history, data lineage, record counts, data-quality outcomes, and exception handling, enabling end-to-end observability, reconciliation, and auditability.
  • Define and apply data governance strategies across the lakehouse - including access control, data classification, PII handling, lineage, retention, and data-quality frameworks - leveraging Unity Catalog and aligned to enterprise standards.
  • Develop reusable, performant Delta Lake data models optimized for downstream AI/ML, analytics, and reporting consumption, applying techniques such as partitioning, Z-ordering / liquid clustering, and OPTIMIZE.
  • Prepare and structure analytics- and AI-ready datasets with upstream AI/ML activities in mind - supporting feature engineering, feature stores, embeddings/vector data, and ML-ready data products for the organization's AI initiatives.
  • Enable cross-cloud data movement and sharing - migrating or sharing data from other cloud platforms (e.g., AWS, GCP) and cloud data warehouses into the Azure Databricks lakehouse, including via Delta Sharing.
  • Troubleshoot and resolve data pipeline and data-quality issues to ensure consistent, dependable data delivery, and partner with BI, analytics, and business teams to deliver curated, analytics-ready datasets for Power BI and Tableau.
  • Orchestrate and automate workflows using Databricks Workflows/Jobs and Lakeflow Declarative Pipelines (DLT), with CI/CD via Git and Databricks Asset Bundles; contribute to data engineering standards, documentation, code reviews, testing, and continuous improvement.
  • Deliver projects using Agile or hybrid methodologies, manage tasks independently, provide accurate status updates, and proactively identify risks or dependencies.

Skills

Azure Databricks
Spark/PySpark
SQL
ETL/ELT pipelines
Data Integration

Education

Bachelor's degree

Tools

Unity Catalog
ADLS Gen2
Delta Lake
Delta Sharing

Job description

BTVK Advisory is a leading advisory firm whose specialized professionals guide clients through an ever-changing business world, helping them win now and anticipate tomorrow. BTVK Advisory, and its affiliated entities, have operations in North America, South America, Europe, Asia, and Australia. BTVK Advisory's ultimate parent entity, Baker Tilly US, LLP, is an independent member of Baker Tilly International, a worldwide network of independent accounting and business advisory firms in 141 territories, with 43,000 professionals and a combined worldwide revenue of $5.2 billion.Baker Tilly is an equal opportunity/affirmative action employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, disability or protected veteran status, gender identity, sexual orientation, or any other legally protected basis, in accordance with applicable federal, state or local law.To be added to all ET through Experienced requisitions Any unsolicited resumes submitted through our website or to Baker Tilly Advisory Group, LP, employee e-mail accounts are considered property of Baker Tilly Advisory Group, LP, and are not subject to payment of agency fees. In order to be an authorized recruitment agency ("search firm") for Baker Tilly Advisory Group, LP, there must be a formal written agreement in place and the agency must be invited, by Baker Tilly's Talent Attraction team, to submit candidates for review via our applicant tracking system.Job Description:

Responsibilities
  • Design, build, and maintain robust, scalable ETL/ELT pipelines on Azure Databricks to ingest, transform, and curate data across the medallion architecture - from raw/bronze landing through silver and gold to platinum/consumption-ready layers - ensuring reliability, performance, and timeliness.
  • Build and manage Unity Catalog assets (catalogs, schemas, tables, and Volumes) and configure or mount external locations/Volumes for governed access to cloud storage such as ADLS Gen2.
  • Design and implement audit, control, and operational-metadata tables that capture pipeline run history, data lineage, record counts, data-quality outcomes, and exception handling, enabling end-to-end observability, reconciliation, and auditability.
  • Define and apply data governance strategies across the lakehouse - including access control, data classification, PII handling, lineage, retention, and data-quality frameworks - leveraging Unity Catalog and aligned to enterprise standards.
  • Develop reusable, performant Delta Lake data models optimized for downstream AI/ML, analytics, and reporting consumption, applying techniques such as partitioning, Z-ordering / liquid clustering, and OPTIMIZE.
  • Prepare and structure analytics- and AI-ready datasets with upstream AI/ML activities in mind - supporting feature engineering, feature stores, embeddings/vector data, and ML-ready data products for the organization's AI initiatives.
  • Enable cross-cloud data movement and sharing - migrating or sharing data from other cloud platforms (e.g., AWS, GCP) and cloud data warehouses into the Azure Databricks lakehouse, including via Delta Sharing.
  • Troubleshoot and resolve data pipeline and data-quality issues to ensure consistent, dependable data delivery, and partner with BI, analytics, and business teams to deliver curated, analytics-ready datasets for Power BI and Tableau.
  • Orchestrate and automate workflows using Databricks Workflows/Jobs and Lakeflow Declarative Pipelines (DLT), with CI/CD via Git and Databricks Asset Bundles; contribute to data engineering standards, documentation, code reviews, testing, and continuous improvement.
  • Deliver projects using Agile or hybrid methodologies, manage tasks independently, provide accurate status updates, and proactively identify risks or dependencies.
Requirements
  • Bachelor's degree required, preferably in Information Technology, Computer Science, Data Science, Analytics, or Statistics.
  • 5-8 years of hands-on experience in data engineering, data integration, or analytics engineering, including substantial time delivering on Databricks.
  • Deep, hands-on experience with Azure Databricks and its critical components, including workspaces, clusters/SQL warehouses, Unity Catalog, and ADLS Gen2 integration.
  • Strong proficiency in Spark/PySpark and SQL, including the ability to develop complex queries and perform advanced data transformations.
  • Hands-on experience ingesting structured, semi-structured, and unstructured data via custom streams (Structured Streaming / Auto Loader), co
Your Match

How well this role fits your profile.

No H-1B filings for this employer appear in the public Department of Labor record. That is not the same as a refusal to sponsor — smaller employers and first-time filers often have no history, and company names do not always match the filing name. Search the sponsor database .

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer
Senior Data Engineer

Baker Tilly • Bengaluru

On-site
INR 1,200,000 - 3,000,000
Senior Data Engineer
Senior Data Engineer

Baker Tilly US • Bengaluru

On-site
INR 700,000 - 1,000,000
Senior Databricks Engineer
Senior Databricks Engineer

Keka Technologies Private Limited • Hyderabad

On-site
INR 1,500,000 - 2,500,000
Data Engineering Lead
Data Engineering Lead

Kumaran Systems • Hyderabad

On-site
INR 2,800,000 - 4,000,000
Senior Data Engineer
Senior Data Engineer

Toppan Merril • Chennai District

On-site
INR 800,000 - 1,200,000
Senior Data Engineer
Senior Data Engineer

Kumaran Systems • Hyderabad

On-site
INR 4,000,000 - 6,000,000
Data Engineer
Data Engineer

Kumaran Systems • Hyderabad

On-site
INR 2,400,000 - 4,200,000
Data Engineer
Data Engineer

Kumaran Systems • Chennai District

On-site
INR 4,000,000 - 7,000,000
Senior Databricks Engineer
Senior Databricks Engineer

DataBeat • Hyderabad

On-site
INR 2,500,000 - 4,200,000
Databrick Data Engineer
Databrick Data Engineer

BDO India • Mumbai

On-site
INR 2,500,000 - 4,500,000