Sr. Data Engineer – Clinical Data Hub

Amgen Inc. (IR)

Hyderabad

On-site

INR 1,200,000 - 2,100,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Amgen Inc. (IR) invites qualified candidates for a Senior Data Engineer role focused on designing, building, and maintaining scalable data pipelines that power clinical data products.

You will collaborate with data architects, scientists, and SMEs to deliver end-to-end data solutions, ensure data quality, and govern sensitive information. The position requires hands-on expertise with Databricks, Apache Spark, Python/R, SQL, and cloud platforms, with a shift that supports global operations.

Qualifications

  • Big data tech and platforms like Databricks and Apache Spark
  • Python/R for EDA, feature engineering, ML model training
  • Proficiency in SQL and data visualization tools
  • Strong data governance and privacy knowledge (GDPR/CCPA)
  • ETL/ELT and data integration familiarity

Responsibilities

  • Design, develop, and maintain data solutions for data generation, collection, and processing
  • Assist in design and development of the data pipeline
  • Create data pipelines and ensure data quality via ETL processes
  • Contribute to data pipelines, ETL/ELT, and data integration solutions
  • Own data pipeline projects from inception to deployment, manage scope/timelines
  • Collaborate with cross-functional teams to understand data requirements
  • Develop and maintain data models, dictionaries, and documentation
  • Implement data security and privacy measures
  • Leverage cloud platforms (AWS preferred) to build scalable data solutions
  • Collaborate with Data Architects, Business SMEs, and Data Scientists
  • Identify and resolve complex data-related challenges
  • Adhere to coding, testing, and designing reusable code/components
  • Explore new tools to improve ETL performance
  • Participate in sprint planning and estimations
  • Collaborate with product teams

Skills

Big data technologies
Databricks
Apache Spark (PySpark, SparkSQL)
ETL/ELT processes
Python/R for data analysis
SQL data analysis
Data governance
Data security & privacy
Data visualization tools

Education

Master’s or Bachelor's degree in Computer Science/IT

Tools

Databricks
Apache Spark (PySpark, SparkSQL)
Python
R
SQL
Data Visualization tools

Job description

Career Category Information Systems Job Description

ABOUT THE ROLE Clinical Data Hub Amgen’s Clinical Data Hub (CDH) is a Information Technology product team chartered to identify, design and implement technology that powers Amgen’s end‑to‑end drug development lifecycle. We are at an inflection point, accelerating through rapid, AI‑driven modernization to build clinical data products that enable both drug regulatory submissions and drug discovery and development. If you’re passionate about turning complex clinical data into resilient, scalable products that help speed life‑changing medicines to patients worldwide, this is a once‑in‑a‑decade opportunity to do your career‑best work on a global stage. Role Description: he role is responsible for designing, building, maintaining, analyzing, and interpreting data to provide actionable insights that drive business decisions. This role involves working with large datasets, developing reports, supporting and executing data governance initiatives and, visualizing data to ensure data is accessible, reliable, and efficiently managed. The ideal candidate has strong technical skills, experience with big data technologies, and a deep understanding of data architecture and ETL processes.

Roles & Responsibilities:
  • Design, develop, and maintain data solutions for data generation, collection, and processing
  • Be a key team member that assists in design and development of the data pipeline
  • Create data pipelines and ensure data quality by implementing ETL processes to migrate and deploy data across systems
  • Contribute to the design, development, and implementation of data pipelines, ETL/ELT processes, and data integration solutions
  • Take ownership of data pipeline projects from inception to deployment, manage scope, timelines, and risks
  • Collaborate with cross-functional teams to understand data requirements and design solutions that meet business needs
  • Develop and maintain data models, data dictionaries, and other documentation to ensure data accuracy and consistency
  • Implement data security and privacy measures to protect sensitive data
  • Leverage cloud platforms (AWS preferred) to build scalable and efficient data solutions
  • Collaborate with Data Architects, Business SMEs, and Data Scientists to design and develop end‑to‑end data pipelines to meet fast paced business needs across geographic regions
  • Identify and resolve complex data‑related challenges
  • Adhere to best practices for coding, testing, and designing reusable code/component
  • Explore new tools and technologies that will help to improve ETL platform performance
  • Participate in sprint planning meetings and provide estimations on technical implementation
  • Collaborate and communicate effectively with product teams
Basic Qualifications and Experience:
  • Master’s or Bachelor's degree with 8 - 12 years of experience in Computer Science, IT or related field
Functional Skills
  • Must-Have Skills: Hands on experience with big data technologies and platforms, such as Databricks, Apache Spark (PySpark, SparkSQL), workflow orchestration, performance tuning on big data processing Hands on experience with various Python/R packages for EDA, feature engineering and machine learning model training Proficiency in data analysis tools (eg. SQL) and experience with data visualization tools Excellent problem‑solving skills and the ability to work with large, complex datasets Strong understanding of data governance frameworks, tools, and best practices. Knowledge of data protection regulations and compliance requirements (e.g., GDPR, CCPA)
  • Good-to-Have Skills: Experience with ETL tools such as Apache Spark, and various Python packages related to data processing, machine learning model development Strong understanding of data modeling, data warehousing, and data integration concepts Knowledge of Python/R, Databricks, SageMaker, cloud data platforms Clinical development domain knowledge is a plus
  • Professional Certifications: Certified Data Engineer / Data Analyst (preferred on Databricks or cloud environments) Certified Data Scientist (preferred on Databricks or Cloud environments) Machine Learning Certification (preferred on Databricks or Cloud environments) SAFe for Teams certification (preferred)
  • Soft Skills: Excellent critical‑thinking and problem‑solving skills Strong communication and collaboration skills Demonstrated awareness of how to function in a team setting Demonstrated presentation skills
Shift Information:

This position requires you to work a later shift and may be assigned a second or third shift schedule. Candidates must be willing and able to work during evening or night shifts, as required based on business requirements.

Amgen is committed to unlocking the potential of biology for patients suffering from serious illnesses by discovering, developing, manufacturing and delivering innovative human therapeutics. This approach begins by using tools like advanced human genetics to unravel the complexities of disease and understand the fundamentals of human biology. Amgen focuses on areas of high unmet medical need and leverages its biologics manufacturing expertise to strive for solutions that improve health outcomes and dramatically improve people's lives. A biotechnology pioneer since 1980, Amgen has grown to be one of the world's leading independent biotechnology companies, has reached millions of patients around the world and is developing a pipeline of medicines with breakaway potential. For more information, visit www.amgen.com and follow us on www.twitter.com/amgen

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Specialist Data Analytics
Specialist Data Analytics

Amgen Inc. (IR) • Hyderabad

On-site
INR 1,500,000 - 2,800,000
Sr. Associate Systems Analyst
Sr. Associate Systems Analyst

Amgen Inc. (IR) • Hyderabad

On-site
INR 1,200,000 - 2,000,000
Sr. Data Engineer – Clinical Data Foundation
Sr. Data Engineer – Clinical Data Foundation

Amgen SA • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Senior Software Engineer – Clinical Data Foundation
Senior Software Engineer – Clinical Data Foundation

Amgen Inc. (IR) • Hyderabad

On-site
INR 1,400,000 - 2,200,000
Sr Mgr Data Engineer
Sr Mgr Data Engineer

Amgen Inc. (IR) • Hyderabad

On-site
INR 1,800,000 - 2,500,000
Senior Data Engineer
Senior Data Engineer

Amgen • Hyderabad

On-site
INR 1,200,000 - 2,000,000
Competitive benefits
Collaborative culture
Professional development opportunities
Sr Associate IS Engineer, Commercialization Technology
Sr Associate IS Engineer, Commercialization Technology

Amgen Inc. (IR) • Hyderabad

On-site
INR 1,500,000 - 2,500,000
Sr Data Engineer
Sr Data Engineer

Amgen SA • Hyderabad

On-site
INR 2,500,000 - 4,000,000
Data Management Sr Associate
Data Management Sr Associate

Amgen Inc. (IR) • Hyderabad

On-site
INR 8,646,000 - 11,527,000
Data Sciences - Information and Data Architecture Mgr
Data Sciences - Information and Data Architecture Mgr

Amgen Inc. (IR) • Hyderabad

On-site
INR 900,000 - 1,500,000