Process Data Engineer I - Pharmaceutical Product Development

Bristol-Myers Squibb

Hyderabad

On-site

INR 1,200,000 - 1,800,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Bristol Myers Squibb hires an early‑career Data Engineer to build data products for product development across molecular features, lab data, and manufacturing parameters. You will design scalable pipelines, transform raw data into trusted, model‑ready data assets, and collaborate with scientists and data scientists to enable AI and analytics across global teams.

Ideal candidates have 2+ years of hands-on data engineering experience with SQL, Python, Databricks, and cloud platforms, in a

Qualifications

  • 2+ years of industry experience in data engineering or related field.
  • Hands-on with modern data technologies including SQL, Python, ETL/ELT, Delta Lake and data warehousing.
  • Experience with Databricks, dbt, and cloud data platforms is preferred.

Responsibilities

  • Build scientific data products for product development across molecular features and lab data.
  • Develop scalable data pipelines supporting analytics, AI and modelling.
  • Transform raw data into model-ready datasets for cross-functional teams.
  • Improve data accessibility, reliability, and quality through automation.
  • Contribute to Databricks, Data Fabric and cloud-native architectures.
  • Automate data workflows and implement engineering best practices.
  • Develop reusable data assets for Product Development across US, Europe, and India.

Skills

SQL
Python
ETL/ELT
Data Modelling
Data Warehousing
Databricks
dbt
PySpark
Delta Lake
Cloud platforms

Education

Bachelor's or Master's in CS/Engineering/IS/Bioinformatics

Tools

Databricks
dbt
AWS
PySpark

Job description

At Bristol Myers Squibb, our employees often ask, “Who are you working for?”—a question that fuels collaboration, accountability, and urgency in our work. Our purpose-driven culture inspires us to discover, develop, and deliver innovative medicines to prevail over serious diseases. We offer uniquely interesting and meaningful work, opportunities for growth, and a supportive environment that values inclusion, wellbeing, flexibility, and comprehensive benefits. This is work that transforms the lives of patients, and the careers of those who do it.

Build the data foundation that helps us get to an AI native state Advanced AI is only as powerful as the data that enables it.

We are seeking an early‑career Data Engineer passionate about creating high‑quality scientific data products and contributing to the development of data fabric to support advanced analytics, AI, machine learning, and scientific decision‑making across Pharmaceutical Product Development.

The ideal candidate will be an engineer who sees data not as a collection of tables, but as a strategic asset that powers scientific discovery and AI innovation.

This role will provide an opportunity to work on complex datasets, modern cloud technologies, and cutting‑edge digital transformation initiatives, with applications across chemical process development, biologics development, drug product development, and analytical development.

What You Will Do
  • Build scientific data products for product development focused on datasets for molecular features, material properties, laboratory data, process and manufacturing parameters, stability and product performance data.
  • Contribute to initiatives for data structuring and data contextualization.
  • Develop and maintain scalable data pipelines supporting analytics, AI, and scientific modelling initiatives.
  • Transform raw scientific and operational data into trusted, model‑ready data products.
  • Work with cross‑functional teams to improve data accessibility, reliability, and quality.
  • Support enterprise initiatives involving Databricks, Data Fabric, and cloud‑native architectures.
  • Automate data workflows and reduce manual effort through engineering best practices.
  • Contribute to development of reusable data assets supporting Product Development innovation across US, Europe and India.
What Makes You Successful

You are naturally curious and:

  • Ask why before asking how.
  • Investigate root causes rather than treating symptoms.
  • Enjoy solving messy, ambiguous data problems.
  • Balance technical rigor with practical execution.
  • Work effectively with scientists, analysts, data scientists, and engineers.
Qualifications
  • Bachelor's or Master’s degree in Computer Science, Chemical Engineering, Information Systems, Bioinformatics, Biotechnology or related field with 2+ years of industry experience.
  • Technical hands‑on experience with modern data technologies. Data Engineering: SQL, Python, ETL/ELT Development, Delta Lake, Lakehouse Architecture, Medallion Architecture, Data Modelling, Data Warehousing, Distributed Computing, Data Validation Cloud & Modern Data Platforms: Databricks, dbt, AWS Modern Data Practices: Data Product Design, Vector Databases, Data Quality Engineering, Data Observability, Metadata Management, Master Data Management, Data Lineage, Data Governance principles, Data Cataloguing Emerging Technologies: AI-Ready Data Foundations, Vector Database fundamentals, Semantic Layer Design, Knowledge Graph Concepts, Data Foundations for GenAI & Agentic AI Applications Software Engineering & Delivery Practices: Git & Version Control, API Integration, Workflow Automation, CI/CD Fundamentals, Agile Delivery.
  • Strong analytical and problem‑solving skills. Excellent communication and collaboration abilities.
Preferred Experience
  • Working with pharmaceutical product development datasets in the scientific, manufacturing, or laboratory data domains.
  • Exposure to scientific, analytical, or laboratory‑based methods through academic coursework, research with an aptitude for understanding the data generated by these techniques.
  • Familiarity with Data products and AI/ML workflows.
  • Experience with PySpark, Python, SQL, dbt, Databricks.
How We Work

Where you work matters – because collaboration, innovation and patient impact happen in many settings. Our roles are structured across four work models: site-essential, site-by-design, field-based and remote-by-design. The model assigned to this role is based on its core responsibilities.

Supporting People with Disabilities

BMS is dedicated to ensuring that people with disabilities can excel through a transparent recruitment process, reasonable workplace accommodations/adjustments and ongoing support in their roles. Applicants can request a reasonable workplace accommodation/adjustment prior to accepting a job offer. If you require reasonable accommodations/adjustments in completing this application, or in any part of the recruitment process, direct your inquiries to adastaffingsupport@bms.com.

Visit careers.bms.com/eeo-accessibility to access our complete Equal Employment Opportunity statement.

Candidate Rights

BMS will consider qualified applicants with arrest and conviction records, pursuant to applicable laws in your area.

For roles based in Los Angeles County only: If you live in or expect to work from Los Angeles County if hired for this position, please visit this page for important additional information: https://careers.bms.com/california-residents/

Data Protection

We will never request payments, financial information, or social security numbers during our application or recruitment process.

If this posting is missing required information required by local law or incorrect, contact BMS at TAEnablement@bms.com with the Job Title and Requisition number. Do not send application‑related inquiries to this email.

We’re creating innovative medicines for patients fighting serious diseases. We’re also nurturing our own diverse team with inspiring work and challenging career options. No matter the role, each one of us makes a contribution. And that makes all the difference.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Manager, Pharmaceutical Product Development GenAI & Data Science Hyderabad - TS - IN R1605511 Posted 15 hours ago
Manager, Pharmaceutical Product Development GenAI & Data Science Hyderabad - TS - IN R1605511 Posted 15 hours ago

Bristol-Myers Squibb • Hyderabad

On-site
INR 4,000,000 - 6,000,000
Manager, Pharmaceutical Product Development GenAI & Data Science
Manager, Pharmaceutical Product Development GenAI & Data Science

Bristol-Myers Squibb • Hyderabad

On-site
INR 1,800,000 - 3,000,000
Software-Engineer--Analytical-Engineering
Software-Engineer--Analytical-Engineering

Bristol Myers Squibb • Hyderabad

On-site
INR 1,200,000 - 2,400,000
Manager 2, Rapid Digital Solutions Hyderabad - TS - IN R1602054 Flexible Location
Manager 2, Rapid Digital Solutions Hyderabad - TS - IN R1602054 Flexible Location

Bristol-Myers Squibb • Hyderabad

Hybrid
INR 4,000,000 - 8,000,000
Software-Engineer--Analytical-Engineering Hyderabad - TS - IN R1605373 Posted 15 hours ago
Software-Engineer--Analytical-Engineering Hyderabad - TS - IN R1605373 Posted 15 hours ago

Bristol-Myers Squibb • Hyderabad

On-site
INR 1,400,000 - 2,400,000
Manager, PV Analytics Center of Excellence Hyderabad - TS - IN R1605415 Posted 9 hours ago
Manager, PV Analytics Center of Excellence Hyderabad - TS - IN R1605415 Posted 9 hours ago

Bristol-Myers Squibb • Hyderabad

Hybrid
INR 1,500,000 - 2,000,000
Data Science Manager
Data Science Manager

Bristol Myers Squibb • Hyderabad

On-site
INR 2,800,000 - 3,800,000
Manager Pharmaceutical Product Development GenAI And Data Science
Manager Pharmaceutical Product Development GenAI And Data Science

Bristol Myers Squibb • Hyderabad

On-site
INR 400,000 - 700,000
Data Engineer II, Data & Analytics
Data Engineer II, Data & Analytics

Bristol Myers Squibb • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Senior Manager, Emerging Technology Risk
Senior Manager, Emerging Technology Risk

Bristol-Myers Squibb • Hyderabad

Hybrid
INR 3,500,000 - 6,000,000