Senior Data Engineer (India)

Openkrill

United States

Remote

USD 150,000 - 190,000

Full time

9 days ago
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Openkrill seeks a Lead Data Engineer to design and operationalize scalable data solutions for analytics and AI/ML initiatives using Databricks, Azure Fabric, PySpark, and SQL.

You will architect pipelines across data warehouses, data lakes, and real-time integration, mentor engineers, and partner with stakeholders to align data strategy with business goals.

Experience with Azure Data Factory, ADLS Gen2, Azure SQL, Power BI, governance (GxP, HIPAA/GDPR), and CI/CD is highly valued.

Qualifications

  • Bachelor’s degree in computer science, information systems, engineering, or related field (Master’s preferred).
  • 5–8 years of experience designing and developing enterprise-scale data solutions.

Responsibilities

  • Architect end-to-end data solutions across data warehouses, data lakes, and real-time pipelines.
  • Design, build, and optimize data models and pipelines using Databricks, PySpark, and SQL.
  • Lead automation, CI/CD, and production readiness of data workloads.
  • Mentor junior engineers and align data initiatives with business goals.
  • Collaborate with analytics and IT stakeholders to enable self-service analytics.

Skills

Leadership
Mentoring
Communication
Stakeholder management

Education

Bachelor’s degree in CS/IS/Engineering
Master’s degree preferred

Tools

Databricks
Azure Fabric
PySpark
SQL
Azure Data Factory
ADLS Gen2
Azure SQL Server
Azure DevOps
Power BI
Tableau/Looker
Git
Terraform
ARM/Bicep

Job description

The Lead Data Engineer will design, build, and operationalize scalable data solutions to support enterprise analytics and AI/ML initiatives. This role requires expert-level proficiency in Databricks, Azure Fabric, PySpark, SQL, and the Azure ecosystem, with deep experience across data warehouses, data lakes, and real-time integration. The Lead Data Engineer will architect end-to-end pipelines using industry-standard tools, drive automation, and move solutions effectively into production. The incumbent will ensure compliance with data governance requirements (including GxP and HIPAA/GDPR) while building reusable, integrated pipelines and analytical models that promote self-service analytics. This role provides technical leadership across the team, mentors junior engineers, and partners with business stakeholders to align data engineering with organizational objectives.

About the Role. Data Architecture & Engineering
  • Architect, design, and implement end-to-end data solutions using Azure Databricks, PySpark, Azure Data Factory, and Azure SQL.
  • Design, build, and maintain data pipelines from data sources through integration to consumption for specific use cases.
  • Implement robust data modeling standards across bronze, silver, and gold layers in the data lake.
  • Develop data models (conceptual, logical, and/or physical) as required.
  • Optimize Spark and SQL workloads for performance, scalability, and cost efficiency.
  • Manage metadata using data preparation, integration, and AI-enabled tools and techniques.
Data Integration & Automation
  • Drive automation in data integration; recommend and lead implementation of techniques to automate repeatable data preparation and integration tasks.
  • Build API-based integrations (REST/JSON) and real-time ingestion frameworks.
  • Automate data workflows using Azure DevOps pipelines and Git-based CI/CD practices.
  • Implement parameterized, reusable pipeline templates for ingestion and transformation.
  • Develop automated unit, regression, and integration testing frameworks for data jobs.
Analytics & Data Enablement
  • Prepare and curate high-quality datasets for BI, reporting, and advanced analytics.
  • Partner with analytics teams using Power BI, Tableau, or similar platforms to define semantic models and KPIs.
  • Implement performance-optimized data models for self-service analytics.
  • Will occasionally provide support to end users on the use of data visualization solutions.
Stakeholder Engagement & Leadership
  • Lead technical design reviews, mentor junior engineers, and promote best practices.
  • Assist cross-functional groups, business analysts, and stakeholders to gather, define, and refine data requirements.
  • Collaborate with business and IT stakeholders to align data engineering with organizational objectives.
  • Propose innovative data ingestion, preparation, and integration techniques to address stakeholder requirements.
  • Contribute to architectural roadmaps and technology evaluations for the data platform.
  • In collaboration with functional leaders, identify inefficiencies and recommend improvements to the executive team.
About YouJob Experience & Education Requirements:

Bachelor’s degree in Computer Science, Information Systems, Engineering, or related field (Master’s preferred)

And

5–8 years of experience designing and developing enterprise-scale data solutions (data warehouses, data lakes, operational databases)

Other:
  • Expert-level proficiency in Databricks, Azure Fabric, PySpark, SQL, and Azure DevOps.
  • Proven experience with Azure Data Factory, ADLS Gen2, and Azure SQL Server.
  • Strong experience with Microsoft Azure data management architectures including Data Warehouse, Data Lake, and Data Catalogue, and supporting processes such as Data Integration, Governance, and Metadata Management.
  • Experience with Power BI required; Tableau or Looker a plus.
  • Working knowledge of CI/CD automation, version control (Git), and infrastructure as code (ARM, Bicep, or Terraform).
  • Experience in life sciences or healthcare industries is a strong plus.
  • Good understanding of GxP, GDPR/HIPAA, and applicable CFR/CTR/CTD regulations.
  • Demonstrated success working with both IT and business stakeholders while integrating analytics and data science output into business processes and workflows.
  • Must have excellent written and verbal communication skills.
  • Proven ability to work independently and as part of a team and meet important deadlines.
  • Statistical analysis skills are an asset.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Data Engineer
Senior Data Engineer

United Network for Organ Sharing (UNOS) • Richmond (VA)

On-site
USD 120,000 - 180,000
Senior Data Engineer
Senior Data Engineer

HeartCentrix Solutions • United States

On-site
USD 110,000 - 160,000
Lead Data Architect
Lead Data Architect

FinThrive • United States

On-site
USD 90,000 - 150,000
Senior Data Engineer
Senior Data Engineer

Peyton Resource Group • Irving (TX)

On-site
USD 100,000 - 130,000
Sr. Data Engineer, Data Platform
Sr. Data Engineer, Data Platform

Mirion Technologies • United States

On-site
USD 120,000 - 160,000
Senior Data Engineer
Senior Data Engineer

RBA, Inc. • United States

On-site
USD 120,000 - 190,000
Lead Data Engineer
Lead Data Engineer

Accrescent Group • Cary (NC)

On-site
USD 140,000 - 180,000
Senior Azure Databricks Data Engineer
Senior Azure Databricks Data Engineer

EXL • New York (NY)

Hybrid
USD 140,000 - 190,000
Data Engineer
Data Engineer

Compunnel, Inc. • San Jose (CA)

On-site
USD 120,000 - 180,000
Sr. Azure Data Engineer with Databricks
Sr. Azure Data Engineer with Databricks

VeeRteq Solutions Inc. • Philadelphia

Hybrid
USD 120,000 - 160,000
Hybrid work model
Onsite presence three days per week