Associate Specialist, Data Engineering

EyeBio

Hyderabad

Hybrid

INR 1,200,000 - 2,000,000

Full time

3 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

EyeBio in Hyderabad is seeking an Associate Specialist: Data Engineering with 2–4 years of hands-on experience to build and support data pipelines, ETL/ELT workflows, and analytics-ready datasets.

The role focuses on Python, PySpark, SQL, AWS, and Databricks, with exposure to data lakes, lakehouse patterns, and production support. You will collaborate with Data Analysts, Data Scientists, and product managers in a hybrid work model.

Qualifications

  • 2–4 years of hands-on experience building and supporting data pipelines, ETL/ELT workflows, and analytics-ready datasets.
  • Strong fundamentals in Python, PySpark, SQL, AWS, and Databricks with exposure to data lakes and lakehouse patterns.
  • Experience with data quality checks, production support, and CI/CD practices is a plus.
  • Experience collaborating with Data Analysts, Data Scientists, and product managers to deliver analytics-ready datasets.
  • Comfort with agile delivery ceremonies and writing technical documentation for pipelines.

Responsibilities

  • Build, enhance, and support batch and streaming data pipelines from defined designs.
  • Develop ETL/ELT transformations across data lake, lakehouse, and warehouse environments.
  • Collaborate with cross-functional teams to understand requirements and deliver curated datasets.
  • Implement data quality checks, validations, and basic anomaly detection to improve data trust.
  • Run, monitor, and troubleshoot pipelines using orchestration and observability tools.
  • Contribute to CI/CD pipelines and automated testing for data workloads.
  • Document pipeline specifications, data contracts, and runbooks.

Skills

Python
PySpark
SQL
AWS
Databricks
Delta Lake
Batch/Streaming

Tools

Databricks Workflows
Terraform
Docker
AWS S3
AWS Glue
AWS Lambda
AWS Step Functions
AWS EMR
CloudWatch

Job description

Job Description

Associate Specialist: Data Engineering

The Opportunity:

Join a global biopharma company with a 130-year legacy and mission to achieve new milestones in healthcare. Be part of a technology-driven, data-led organization supporting a diversified portfolio of medicines, vaccines, and animal health products. Work alongside passionate teams that use data, analytics, and insights to drive decisions and tackle some of the world's greatest health threats.


Our Technology Centers are globally distributed hubs that enable our digital transformation and business outcomes across IT. They bring together diverse teams to collaborate, share best practices, and deliver solutions that save and improve lives.


This role is based at our Hyderabad Tech Center and follows a hybrid working model (3 days onsite, 2 days remote). Candidates are expected to reside within commuting distance of the Hyderabad office.


Role Overview

We are looking for a Data Engineer with 2-4 years of hands-on experience in building and supporting data pipelines, ETL/ELT workflows, and analytics-ready datasets. The ideal candidate should have strong fundamentals in Python, PySpark, SQL, AWS, and Databricks, with practical exposure to data lakes, lakehouse patterns, data warehousing, data quality, and production support. This role is best suited for a hands-on engineer who can work from defined requirements, contribute to reliable data solutions, collaborate with cross-functional teams, and grow into larger ownership over time.


What will you do in this role


  • Build, enhance, and support batch and streaming data pipelines using defined technical designs and backlog requirements.

  • Develop and maintain ETL/ELT transformations using Python, PySpark, and SQL across data lake, lakehouse, and warehouse environments.

  • Work closely with Data Analysts, Data Scientists, senior engineers, tech leads, and product managers to understand requirements and deliver curated, analytics-ready datasets.

  • Implement data quality checks, validations, reconciliations, and basic anomaly checks to improve trust and usability of data outputs.

  • Run, monitor, and troubleshoot pipelines using orchestration and observability tools such as Databricks Workflows, AWS Step Functions, scheduling, logging, monitoring, and alerting.

  • Follow engineering practices including unit testing, integration testing, automated data tests, code reviews, and quality gates within CI/CD.

  • Support BI and analytics use cases by applying dimensional modeling concepts such as facts, dimensions, star/snowflake schemas, and slowly changing dimensions (SCD).

  • Write and tune SQL queries for data profiling, transformations, validations, debugging, and performance improvements.

  • Use AWS services such as S3, Glue, Lambda, Step Functions, EMR, and CloudWatch to support data engineering workloads while following security practices such as IAM, encryption, and least privilege.

  • Contribute to cloud resource provisioning and environment configuration using Terraform, with guidance from senior engineers.

  • Package, deploy, and support workloads using Docker and related runtime configurations, including ECS/Fargate where applicable.

  • Use GitHub for version control, branching, pull requests, code reviews, and contribution to CI/CD pipelines.

  • Develop scalable data processing logic on Databricks / Apache Spark using PySpark and lakehouse concepts such as Delta Lake, ACID transactions, and schema evolution.

  • Use Jupyter/Databricks notebooks for exploration, debugging, and PoCs; convert validated logic into reusable modules, tests, and deployment-ready pipelines.

  • Participate in Agile delivery ceremonies, provide task-level estimates, share progress updates, and raise risks or dependencies early.

  • Create and maintain technical documentation such as pipeline specifications, data contracts, runbooks, and support notes.


Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Associate Specialist , Data Engineering
Associate Specialist , Data Engineering

Merck • India

Hybrid
INR 1,200,000 - 1,800,000
Associalte Specialist , Data Engineering
Associalte Specialist , Data Engineering

Merck • India

Hybrid
INR 1,000,000 - 1,400,000
Spcialist , Data Engineering
Spcialist , Data Engineering

Merck • India

Hybrid
INR 3,000,000 - 5,500,000
Specialist , Data Engineering
Specialist , Data Engineering

MSD Malaysia • Hyderabad

Hybrid
INR 1,800,000 - 3,000,000
Specialist , Data Engineering
Specialist , Data Engineering

EyeBio • Hyderabad

Hybrid
INR 1,800,000 - 3,000,000
Associate - Data Engineering (Pharma)
Associate - Data Engineering (Pharma)

Chryselys • Hyderabad

Hybrid
INR 1,200,000 - 1,800,000
Senior Specialist, ESE ML Engineer
Senior Specialist, ESE ML Engineer

Merck • India

Hybrid
INR 1,500,000 - 3,500,000
Hybrid work model
Sr. Data Engineer
Sr. Data Engineer

Cohere Health • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Senior Data Engineer
Senior Data Engineer

Recro • Bengaluru

On-site
INR 2,500,000 - 4,500,000
Data Engineer
Data Engineer

Enterprise Minds, Inc • Pune District

On-site
INR 1,400,000 - 2,400,000