Databricks Data Engineer (Azure + Databricks AI)

OneData Software Solutions

Coimbatore District

On-site

INR 2,500,000 - 5,200,000

Full time

17 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

OneData Software Solutions seeks a Databricks Data Engineer to design, build, and optimize scalable data and AI pipelines on Microsoft Azure. You will own lakehouse architecture using Delta Lake and Unity Catalog, enabling GenAI and ML production use cases.

You will collaborate with data scientists, analysts, and stakeholders to transform raw data into AI-ready assets and ensure governance, performance, and cost optimization across pipelines.

Qualifications

  • 5+ years of data engineering experience, including 3+ years with Databricks on Azure.
  • Strong proficiency in Python (PySpark) and SQL.
  • Deep knowledge of Delta Lake, Spark internals, and performance tuning.

Responsibilities

  • Design and build batch and streaming pipelines in Azure Databricks using PySpark, Spark SQL, and Delta Lake.
  • Develop declarative pipelines and orchestrate workloads using Lakeflow and Databricks Workflows/Jobs.
  • Implement data governance, lineage, and access controls using Unity Catalog.
  • Ingest data from multiple sources using Azure Data Factory, Event Hubs, ADLS Gen2, Auto Loader, and REST APIs.

Skills

Python
PySpark
SQL
Delta Lake
Delta Live Tables
Databricks
Azure Data Factory
Azure AD / Entra ID
Unity Catalog
MLflow

Education

Bachelor's degree in Computer Science/IT or related field

Tools

Azure Databricks
ADLS Gen2
Event Hubs
Model Serving / Databricks Mosaic AI

Job description

Location: Coimbatore, Tamil Nadu, India

Employment Type: Full Time

Experience: 5–8+ years in data engineering, with 3+ years hands-on in Databricks

Notice Period: Immediate joiners to 30 days preferred

Shift: Partial US overlap, e.g., 12:30 PM – 9:30 PM IST

About the Role

We are looking for a Databricks Data Engineer to design, build, and optimize scalable data and AI pipelines on Microsoft Azure. You will own lakehouse architecture using Delta Lake and Unity Catalog, and help bring GenAI and machine learning use cases into production using Databricks Mosaic AI capabilities. You will work closely with data scientists, analysts, and business stakeholders to turn raw data into trusted, AI-ready assets.

Key Responsibilities
  • Design and build batch and streaming pipelines in Azure Databricks using PySpark, Spark SQL, and Delta Lake, following medallion (Bronze/Silver/Gold) architecture.
  • Develop declarative pipelines and orchestrate workloads using Lakeflow (Delta Live Tables) and Databricks Workflows/Jobs.
  • Implement data governance, lineage, and access controls using Unity Catalog.
  • Ingest data from multiple sources using Azure Data Factory, Event Hubs, ADLS Gen2, Auto Loader, and REST APIs.
  • Build and support AI/GenAI solutions on Databricks, including RAG pipelines with Vector Search, model deployment through Model Serving, and LLM integration via Databricks Foundation Model APIs or Azure OpenAI.
  • Manage the ML lifecycle with MLflow, covering experiment tracking, model registry, evaluation, and monitoring.
  • Prepare feature sets and curated datasets for ML models and AI agents, using Databricks Feature Store and Agent Framework where applicable.
  • Tune Spark jobs and clusters for performance and cost, including partitioning, liquid clustering/Z-ordering, Photon, and serverless compute.
  • Implement CI/CD for data and AI assets using Databricks Asset Bundles, Azure DevOps or GitHub Actions, and infrastructure as code (Terraform/Bicep).
  • Ensure data quality through pipeline expectations, validation frameworks, and Lakehouse Monitoring.
  • Collaborate with stakeholders to gather requirements, estimate effort, and document solutions.
Required Skills:
  • Bachelor's degree in Computer Science, IT, or a related field (B.E./B.Tech/MCA or equivalent).
  • 5+ years of data engineering experience, including 3+ years with Databricks on Azure.
  • Strong proficiency in Python (PySpark) and SQL.
  • Deep knowledge of Delta Lake, Spark internals, and performance tuning.
  • Hands‑on experience with Azure services: ADLS Gen2, Azure Data Factory, Event Hubs, Key Vault, Entra ID (Azure AD), and Azure Monitor.
  • Practical experience with Unity Catalog and lakehouse governance.
  • Experience with Databricks AI features such as Mosaic AI Model Serving, Vector Search, AI Functions, or MLflow for GenAI.
  • Understanding of LLM concepts including embeddings, chunking, RAG, prompt engineering, and evaluation.
  • Experience with Git‑based development and CI/CD pipelines.
  • Solid grasp of data modeling (dimensional, Data Vault, or lakehouse patterns).
Preferred Qualifications
  • Experience building AI agents or chatbots with Databricks Agent Framework, LangChain, or LlamaIndex.
  • Familiarity with Databricks SQL, AI/BI Dashboards, and Genie spaces.
  • Experience with real‑time streaming (Structured Streaming, Kafka).
  • Exposure to Microsoft Fabric, Power BI, or Azure Synapse.
  • Experience working with US/UK clients or in regulated domains such as BFSI, healthcare, or insurance.
Soft Skills
  • Strong problem‑solving and communication skills, with the ability to explain technical trade‑offs to non‑technical audiences.
  • Comfortable working in Agile/Scrum teams and owning deliverables end to end.
  • Ability to mentor junior engineers and contribute to best practices.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Databricks - Data Engineer
Databricks - Data Engineer

Tredence Inc. • Bengaluru

On-site
INR 1,500,000 - 2,500,000
Azure Data Engineer
Azure Data Engineer

Texplorers • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Azure Databricks Developer
Azure Databricks Developer

Birlasoft • India

On-site
INR 2,500,000 - 4,500,000
Senior Data Engineer (Jaipur, India)
Senior Data Engineer (Jaipur, India)

Thoughtswinsystems • Jaipur

On-site
INR 1,200,000 - 2,000,000
Senior Data Engineer
Senior Data Engineer

USEReady • Bengaluru

On-site
INR 1,500,000 - 2,100,000
Senior Databricks Data Engineer
Senior Databricks Data Engineer

Omnicom Global Solutions • Bengaluru

Hybrid
INR 1,800,000 - 2,400,000
Hybrid work model
Azure Databricks Specialist
Azure Databricks Specialist

Birlasoft • Pune District

On-site
INR 1,500,000 - 2,500,000
Databricks Data Architect
Databricks Data Architect

Unison Group • Chennai District

On-site
INR 1,800,000 - 3,000,000
Data Bricks Expert
Data Bricks Expert

eSolutionsFirst • Hyderabad

On-site
INR 4,500,000 - 7,000,000
Data Scientist (With Databricks)- Immediate Hiring
Data Scientist (With Databricks)- Immediate Hiring

Diggibyte Technologies • Ahmedabad District, Chennai District, Bengaluru

Hybrid
INR 1,200,000 - 1,800,000