Sr. Data Engineer – Clinical Data Foundation

Amgen SA

Hyderabad

On-site

INR 4,000,000 - 7,000,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Amgen is seeking a Sr. Data Engineer in Hyderabad to design, build, and maintain scalable data pipelines and governance. You will work with Databricks, Spark, Python, and SQL to deliver reliable data insights for business decisions across global regions.

The role requires strong ETL/ELT expertise, data modeling, and collaboration with cross-functional teams to ensure data accessibility and security. Evening or night shifts may be required based on business needs.

Qualifications

  • Hands-on experience with Databricks and Spark for large-scale data processing.
  • Proficiency in Python or R for data analysis and model training.
  • Strong SQL and data visualization skills.
  • Understanding of data governance and regulatory compliance (GDPR/CCPA).
  • Experience with Delta Lake and DataFrames.

Responsibilities

  • Design, develop, and maintain data solutions for data generation, collection, and processing.
  • Build data pipelines and ensure data quality via ETL/ELT processes.
  • Collaborate with data architects, data scientists, and SMEs across regions.

Education

Master’s / Bachelor's degree in Computer Science / IT or related field

Tools

Databricks
Apache Spark (PySpark)
Python
SQL
Delta Lake
DataFrames

Job description

ABOUT AMGEN

Amgen harnesses the best of biology and technology to fight the world’s toughest diseases, and make people’s lives easier, fuller and longer. We discover, develop, manufacture and deliver innovative medicines to help millions of patients. Amgen helped establish the biotechnology industry more than 40 years ago and remains on the cutting-edge of innovation, using technology and human genetic data to push beyond what’s known today.

ABOUT THE ROLE

Role Description:

The Sr. Data Engineer is responsible for designing, building, maintaining, analyzing, and interpreting data to provide actionable insights that drive business decisions. This role involves working with large datasets, developing reports, supporting and executing data governance initiatives and, visualizing data to ensure data is accessible, reliable, and efficiently managed. The ideal candidate has strong technical skills, experience with big data technologies, and a deep understanding of data architecture and ETL processes

Roles & Responsibilities:

  • Design, develop, and maintain data solutions for data generation, collection, and processing
  • Be a key team member that assists in design and development of the data pipeline
  • Create data pipelines and ensure data quality by implementing ETL processes to migrate and deploy data across systems
  • Contribute to the design, development, and implementation of data pipelines, ETL/ELT processes, and data integration solutions
  • Take ownership of data pipeline projects from inception to deployment, manage scope, timelines, and risks
  • Collaborate with cross-functional teams to understand data requirements and design solutions that meet business needs
  • Develop and maintain data models, data dictionaries, and other documentation to ensure data accuracy and consistency
  • Implement data security and privacy measures to protect sensitive data
  • Leverage cloud platforms (AWS preferred) to build scalable and efficient data solutions
  • Collaborate with Data Architects, Business SMEs, and Data Scientists to design and develop end-to-end data pipelines to meet fast paced business needs across geographic regions
  • Identify and resolve complex data-related challenges
  • Adhere to best practices for coding, testing, and designing reusable code/component
  • Explore new tools and technologies that will help to improve ETL platform performance
  • Participate in sprint planning meetings and provide estimations on technical implementation
  • Collaborate and communicate effectively with product teams

Basic Qualifications and Experience:

  • Master’s /Bachelor’s degree with 9-12 years of experience in Computer Science, IT or related field

Functional Skills:

Must-Have Skills:

  • Hands on experience with big data technologies and platforms, such as Databricks, Apache Spark (PySpark, SparkSQL), workflow orchestration, performance tuning on big data processing
  • Hands on experience with various Python/R packages for data analysis, feature engineering and machine learning model training
  • Proficiency in data analysis tools (eg. "SQL") and experience with data visualization tools
  • Excellent problem-solving skills and the ability to work with large, complex datasets
  • Strong understanding of data governance frameworks, tools, and best practices.
  • Knowledge of data protection regulations and compliance requirements (e.g., GDPR, CCPA)
  • Delta Lake, DataFrames, broadcast join

Good-to-Have Skills:

  • Experience with ETL tools such as Apache Spark, and various Python packages related to data processing, machine learning model development
  • Strong understanding of data modeling, data warehousing, and data integration concepts
  • Knowledge of Python/R, Databricks, SageMaker, cloud data platforms
  • Experience with Clinical Development domain.

Professional Certifications:

  • Certified Data Engineer / Data Analyst (preferred on Databricks or cloud environments)
  • Machine Learning Certification (preferred on Databricks or Cloud environments)
  • SAFe for Teams certification (preferred)

Soft Skills:

  • Excellent critical-thinking and problem-solving skills
  • Strong communication and collaboration skills
  • Demonstrated awareness of how to function in a team setting
  • Demonstrated presentation skills

Shift Information:

This position requires you to work a later shift and may be assigned a second or third shift schedule. Candidates must be willing and able to work during evening or night shifts, as required based on business requirements.

EQUAL OPPORTUNITY STATEMENT

Amgen is an Equal Opportunity employer and will consider you without regard to your race, color, religion, sex, sexual orientation, gender identity, national origin, protected veteran status, or disability status.

We will ensure that individuals with disabilities are provided with reasonable accommodation to participate in the job application or interview process, to perform essential job functions, and to receive other benefits and privileges of employment. Please contact us to request an accommodation.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr. Data Engineer – Clinical Data Foundation
Sr. Data Engineer – Clinical Data Foundation

Biopharma Careers • Hyderabad

On-site
INR 4,000,000 - 6,500,000
Associcate Data Engineer
Associcate Data Engineer

Biopharma Careers • Hyderabad

On-site
INR 3,000,000 - 5,500,000
Associcate Data Engineer
Associcate Data Engineer

Amgen • Hyderabad

On-site
INR 1,400,000 - 2,100,000
Sr Data Engineer
Sr Data Engineer

Biopharma Careers • Hyderabad

On-site
INR 1,800,000 - 2,600,000
Specialist Data Analytics
Specialist Data Analytics

Biopharma Careers • Hyderabad

On-site
INR 1,800,000 - 2,800,000
Sr Data Engineer
Sr Data Engineer

Amgen SA • Hyderabad

On-site
INR 2,500,000 - 4,000,000
Sr. Data Engineer – Clinical Data Hub
Sr. Data Engineer – Clinical Data Hub

Amgen • Hyderabad

On-site
INR 1,800,000 - 2,400,000
Sr. Data Engineer – Clinical Data Hub
Sr. Data Engineer – Clinical Data Hub

Biopharma Careers • Hyderabad

On-site
INR 1,800,000 - 3,600,000
Sr. Associate Data Engineer
Sr. Associate Data Engineer

Amgen SA • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Data Engineer
Data Engineer

Amgen • Hyderabad

On-site
INR 800,000 - 1,200,000