Data Engineer- GCP/Databricks

AuxoAI Engineering Pvt. Ltd.

Bengaluru

Hybrid

INR 2,400,000 - 3,600,000

Full time

8 days ago
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Hybrid work model

Job summary

AuxoAI Engineering Pvt. Ltd. seeks a Senior Data Engineer to lead the design, development, and optimisation of modern data pipelines and cloud-native platforms.

This role focuses on scalable batch and streaming workflows across lakehouse environments, with strong hands-on engineering and mentorship across teams. You will work with AI engineers and data scientists to enable high‑quality data for AI and analytics use cases at scale, while driving governance, performance, and CI/CD practices across

Qualifications

  • 5+ years hands-on data engineering in production pipelines.
  • Experience with Databricks on GCP including BigQuery, GCS, Spark and Delta Lake.
  • Strong Python/Scala programming and SQL for transformation and modeling.
  • Experience with data warehousing concepts, ETL/ELT, and schema evolution.
  • Familiarity with CI/CD, Git, and data quality monitoring.

Responsibilities

  • Design and build scalable batch and streaming data pipelines across lakehouse layers.
  • Develop Databricks-based pipelines and orchestrate with Databricks Jobs.
  • Create analytical data layers in BigQuery or Databricks SQL with performance tuning.
  • Implement transformations for wide and semi-structured data and event-time joins.
  • Collaborate with AI/solution architects, mentor junior engineers, and contribute to docs and reviews.
  • Ensure data governance, lineage, and security best practices across pipelines.

Skills

Databricks
GCP
BigQuery
Spark
Delta Lake
Python/Scala
SQL
Data modeling
ETL/ELT
Data warehousing
CI/CD
Data quality monitoring
Git
Collaboration
Problem solving

Tools

Unity Catalog
Dataplex
Terraform
Docker
Kubernetes

Job description

AuxoAI is seeking a Senior Data Engineer to lead the design, development, and optimisation of modern data pipelines and cloud-native platforms. This role is ideal for someone with deep experience building scalable batch and streaming data workflows across cloud and lakehouse environments, strong hands‑on engineering skills, and a drive to mentor junior engineers.

You will work closely with AI engineers, solution architects, and cross‑functional teams to build production‑grade pipelines spanning ingestion, transformation, and curated data delivery — enabling high‑quality data for AI and analytics use cases at scale.

Location: Bangalore / Mumbai / Hyderabad / Gurgaon (Hybrid — 3 days in office)

Responsibilities
  • - Design and build scalable batch and streaming data pipelines across bronze, silver, and gold medallion layers.
  • - Build historical and incremental ingestion using Auto Loader/Spark Structured Streaming/Kafka feeds, with GCS/Azure/AWS storage and Databricks Jobs orchestration.
  • - Develop and maintain Databricks-based pipelines using Spark and Delta Lake for lakehouse architecture, including migration of legacy or on‑premises data sources.
  • - Design and maintain analytical data layers in BigQuery or Databricks SQL, applying best practices in partitioning, clustering, and performance tuning.
  • - Implement SQL/PySpark transformations for wide and semi‑structured data, including wide‑to‑long processing and typed or hybrid models suited to consumer requirements.
  • - Collaborate with AI engineers and data scientists to build pipelines that feed ML models, AI agents, and analytical systems.
  • - Implement data governance, quality controls, and security best practices including schema enforcement, lineage tracking, and access controls.
  • - Drive engineering best practices across CI/CD, testing, monitoring, and pipeline observability.
  • - Partner with solution architects to translate data requirements into technical designs.
  • - Mentor junior data engineers and contribute to documentation, code reviews, and agile ceremonies.
Requirements
  • - 5+ years of hands‑on experience in data engineering, building and operating production‑grade pipelines.
  • - Hands‑on experience with Databricks on GCP, including BigQuery, GCS, Databricks, Spark, Delta Lake, and structured streaming.
  • - Hands‑on experience with Databricks and Apache Spark, including Delta Lake and end‑to‑end lakehouse implementations.
  • - Strong programming skills in Python and/or Scala, with solid SQL for modelling and transformation.
  • - Experience with data modelling, ETL/ELT, pipeline orchestration, and data warehousing concepts, including experience working with large, evolving JSON/map/array payloads, wide‑to‑long transformations, event‑time context joins and schema‑change handling.
  • - Familiarity with Git, CI/CD pipelines, and data quality monitoring frameworks.
  • - Solid understanding of data architecture, schema design, and performance tuning.
  • - Experience with Unity Catalog, source reconciliation, schema evolution, correction handling and replay/recovery testing.
  • - Strong problem‑solving and collaboration skills.
Bonus Skills
  • - GCP Professional Data Engineer certification.
  • - Experience with Vertex AI, Cloud Functions, Dataproc, or real‑time streaming architectures.
  • - Experience with factory or industrial data sources — MES systems, IoT sensor streams, or operational telemetry.
  • - Familiarity with data governance and cataloguing tools such as Dataplex, Unity Catalog, Atlan, or Collibra.
  • - Exposure to Docker, Kubernetes, API integration, and infrastructure‑as‑code (Terraform).
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data Engineer- GCP/Databricks
Data Engineer- GCP/Databricks

Zohorecruit • Bengaluru

Hybrid
INR 2,400,000 - 4,800,000
Data Engineer- GCP/Databricks
Data Engineer- GCP/Databricks

Zohorecruit • Bangalore Rural

On-site
INR 3,200,000 - 5,200,000
Data Engineer- GCP/Databricks
Data Engineer- GCP/Databricks

Zohorecruit • Bengaluru Urban

On-site
INR 2,400,000 - 4,200,000
Auxo AI - Senior Data Engineer - Google Cloud Platform
Auxo AI - Senior Data Engineer - Google Cloud Platform

Auxo AI • Hyderabad

On-site
INR 800,000 - 1,100,000
Senior Data Engineer - GCP
Senior Data Engineer - GCP

AuxoAI Inc. • Bengaluru

On-site
INR 1,200,000 - 2,000,000
Data Engineer-GenAI
Data Engineer-GenAI

Auxo AI • Gurugram District

On-site
INR 1,500,000 - 2,500,000
Lead Data Engineer (Databricks, PySpark & GCP)
Lead Data Engineer (Databricks, PySpark & GCP)

Egen • Hyderabad

On-site
INR 5,500,000 - 7,500,000
Healthcare benefits
Performance bonus
Sr Data Engineer
Sr Data Engineer

AuxoAI Engineering • Bengaluru Urban

On-site
INR 1,800,000 - 3,200,000
Data Engineer
Data Engineer

AuxoAI Inc. • Bengaluru

On-site
INR 800,000 - 1,500,000
Databricks Data Engineer (Azure + Databricks AI)
Databricks Data Engineer (Azure + Databricks AI)

OneData Software Solutions • Coimbatore District

On-site
INR 2,500,000 - 5,200,000