Sr Data Engineer

Illumina

India

On-site

INR 3,000,000 - 5,500,000

Full time

6 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Illumina is hiring a Senior Data Engineer in India to design, build, and scale data products on a cloud lakehouse. You will work with Databricks and Snowflake, using Python and SQL to model data, implement governance, and drive analytics with AI/ML capabilities.

You will mentor engineers across a global team, contribute to architecture decisions, and own end-to-end data pipelines from ingestion through analytics-ready datasets in a fast-paced, innovative environment.

Qualifications

  • 8+ years of professional data engineering experience building and scaling data products on cloud platforms such as Databricks and/or Snowflake.
  • Strong proficiency in Python, including reusable framework development using functional and object-oriented programming.
  • Advanced SQL and strong data modeling skills (relational, dimensional, and lakehouse).
  • Solid understanding of distributed systems and system design for large-scale data processing.
  • Hands-on experience with open table formats (Delta Lake, and/or Apache Iceberg) and big-data file formats (Parquet).
  • Experience with Spark and modern ELT tooling (e.g., dbt).
  • Experience with data observability, governance, security, and compliance practices (RBAC, PII, SOX).
  • Demonstrated adoption of AI in data and analytics engineering workflows.
  • Solid software engineering foundation - Git, REST APIs, JSON, CI/CD on at least one cloud environment (AWS preferred).
  • Strong written and verbal communication skills, with the ability to work effectively across business, AI, and platform teams and lead technical discussion.
  • Bachelor's degree in Computer Science, Data Science, Information Systems, Engineering, Mathematics, or a related field, or equivalent demonstrable experience.

Responsibilities

  • Translate domain needs into well-modeled, governed and scalable data products.
  • Design, build, and scale end-to-end data products on Databricks (and Snowflake) from ingestion to analytics-ready datasets (Bronze/Silver/Gold).
  • Develop reusable Python frameworks for ingestion, transformation, validation, and publishing.
  • Design robust data models and build distributed data pipelines using Spark, Delta Lake, dbt, and SQL.
  • Embed data quality, reconciliation, validation, and governance into pipelines (Unity Catalog: lineage, RBAC, masking, PII handling).
  • Monitor, alert, troubleshoot, perform root-cause analysis and SLA adherence for critical datasets.
  • Adopt AI in data and analytics workflows to accelerate development and optimization.
  • Provide technical leadership — set standards, review code, mentor engineers, and communicate trade-offs.

Skills

Python
SQL
Data modeling
Distributed systems
CI/CD
REST APIs
Git

Education

Bachelor's degree in CS/DS/Engineering or related field

Tools

Databricks
Snowflake
dbt
Delta Lake
Spark
Unity Catalog

Job description

What if the work you did every day could impact the lives of people you know? Or all of humanity?

At Illumina, we are expanding access to genomic technology to realize health equity for billions of people around the world. Our efforts enable life-changing discoveries that are transforming human health through the early detection and diagnosis of diseases and new treatment options for patients.

Working at Illumina means being part of something bigger than yourself. Every person, in every role, has the opportunity to make a difference. Surrounded by extraordinary people, inspiring leaders, and world changing projects, you will do more and become more than you ever thought possible.

Role Overview

The Senior Data Engineer is a seasoned, hands-on engineer who designs, builds, and scales data products on our cloud lakehouse, powering analytics, reporting, and AI/ML across Illumina. We are looking for someone with strong proficiency in Python, SQL, and data modeling, a solid understanding of distributed systems and system design who has built and scaled data products on modern cloud platforms such as Databricks and Snowflake.

This is a hands-on, senior individual-contributor role with end-to-end ownership and leadership spanning multiple domains such as Supply Chain, Manufacturing and Quality, including mentoring engineers on our global (India-based) team.

Key Responsibilities
  • Partner across business, AI, and platform teams translating domain needs (e.g., SAP, Manufacturing, Quality) into well-modeled, governed and scalable data products.
  • Design, build, and scale end-to-end data products on Databricks (and interoperating with Snowflake) - from ingestion through curated, analytics-ready datasets following a medallion (Bronze/Silver/Gold) architecture.
  • Develop reusable frameworks, libraries, and standardized patterns in Python (functional and OOP as appropriate) for ingestion, transformation, validation, and publishing.
  • Design robust data models (relational, dimensional, and lakehouse) and apply strong system-design judgment to build performant, reliable distributed data pipelines using Spark, Delta Lake / open table formats, dbt, and SQL.
  • Embed data quality, reconciliation, validation, and governance into pipelines (Unity Catalog: lineage, RBAC, masking, PII handling.
  • Monitoring, alerting, troubleshooting, root-cause analysis, and SLA adherence for business-critical datasets.
  • Adopt AI in day-to-day data and analytics engineering to accelerate development, testing, and optimization.
  • Act as a technical leader - set standards, lead code reviews, contribute to architecture decisions, mentor engineers, and communicate trade-offs to peers and stakeholders.
Required Qualifications
  • 8+ years of professional data engineering experience building and scaling data products on cloud platforms such as Databricks and/or Snowflake.
  • Strong proficiency in Python, including reusable framework development using functional and object-oriented programming.
  • Advanced SQL and strong data modeling skills (relational, dimensional, and lakehouse).
  • Solid understanding of distributed systems and system design for large-scale data processing.
  • Hands-on experience with open table formats (Delta Lake, and/or Apache Iceberg ) and big-data file formats (Parquet).
  • Experience with Spark and modern ELT tooling (e.g., dbt).
  • Experience with data observability , governance, security, and compliance practices (RBAC, PII, SOX).
  • Demonstrated adoption of AI in data and analytics engineering workflows.
  • Solid software engineering foundation - Git, REST APIs, JSON, CI/CD on at least one cloud environment (AWS preferred).
  • Strong written and verbal communication skills, with the ability to work effectively across business, AI, and platform teams and lead technical discussion.
  • Bachelor's degree in Computer Science, Data Science, Information Systems, Engineering, Mathematics, or a related field, or equivalent demonstrable experience.
Preferred Qualifications
  • Strong plus with domain knowledge of SAP, Manufacturing, and/or Quality data and processes.
  • Experience delivering in GxP / 21 CFR Part 11 or comparable regulated environments (life sciences, pharma, medical devices).
  • Experience with Unity Catalog and lakehouse governance at scale.
  • Snowflake-to-Databricks migration experience.
  • Exposure to SAP data (ECC / S/4HANA, CDS views) and SAP data integration patterns (e.g., SAP Business Data Cloud). Bonus if candidate have additional domain knowledge such as Commercial and Finance.
  • Databricks and/or dbt certifications.
  • Familiarity with Power BI / Tableau and enabling BI and conversational-analytics.
Competencies We Value
  • Ownership: End-to-end accountability for data products, from design through production support.
  • Engineering Craft: Clean, reusable, well-designed code and pride in auditable, reliable data.
  • System Thinking: Sound design judgment across scale, performance, and cost.
  • Technical Leadership: Raising the bar through standards, reviews, mentorship.
  • Learning
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr Data Engineer
Sr Data Engineer

Illumina • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Data Engineer 2
Data Engineer 2

Illumina • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Sr Data Engineer, Data Eng & Governance Bangalore,India
Sr Data Engineer, Data Eng & Governance Bangalore,India

Via Licensing Corporation • Bengaluru

On-site
INR 2,500,000 - 4,200,000
Sr Data Engineer, Data Eng & Governance
Sr Data Engineer, Data Eng & Governance

Via Licensing Corporation • Bengaluru

On-site
INR 1,800,000 - 3,000,000
Lead Data Engineer
Lead Data Engineer

PocketFM • Bengaluru

On-site
INR 3,000,000 - 5,400,000
Health insurance
Paid time off
Remote learning budget
Senior Data Engineer
Senior Data Engineer

Amgen • Hyderabad

On-site
INR 2,500,000 - 4,500,000
Data Engineer & Business Intelligence Analyst
Data Engineer & Business Intelligence Analyst

Model Business Services • SECTOR 78

On-site
INR 1,200,000 - 2,000,000
Opportunity for Data Engineer
Opportunity for Data Engineer

Hinduja Tech Limited • Pune District

On-site
INR 1,800,000 - 3,000,000
Senior Data Engineer
Senior Data Engineer

Publicis Production • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Senior Databricks Engineer
Senior Databricks Engineer

DataBeat • Hyderabad

On-site
INR 2,500,000 - 4,200,000