Sr Data Engineer

Amgen

India

On-site

INR 4,000,000 - 7,000,000

Full time

2 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Amgen is seeking a highly motivated Senior Data Engineer to own the design and development of complex data pipelines, data integration frameworks, and metadata-driven architectures for R&D data. Responsibilities include building scalable ETL/ELT solutions, real-time and batch processing, and ensuring data security and governance across enterprise data fabric layers.

The role requires deep expertise in big data processing, distributed computing, data modeling, and governance to enable

Qualifications

  • Proficient with large-scale data pipelines and governance.
  • Strong knowledge of R&D data needs in Biotech/Pharma contexts.
  • Experience building data fabrics, data meshes, and self-service analytics.
  • Proven ability to design scalable ETL/ELT pipelines.
  • Experience with real-time and batch processing and data virtualization.

Responsibilities

  • Design, develop, and maintain scalable ETL/ELT pipelines for varied data.
  • Build real-time and batch data processing solutions across sources.
  • Optimize Spark, Hadoop and distributed frameworks for cost-efficiency.
  • Implement metadata management and data lineage tooling.
  • Ensure RBAC, security, and compliance across environments.
  • Improve query performance via indexing, partitioning, and caching.
  • Develop CI/CD pipelines for data workflows and monitoring.
  • Create data virtualization layers for cross-storage access.
  • Collaborate with data architects, analysts, and DevOps teams.
  • Stay current with new data technologies and governance practices.

Skills

Databricks
PySpark
SparkSQL
Apache Spark
AWS
Python
SQL
Scaled Agile methodologies
DevOps practices
Excellent communication
Team collaboration

Education

Master's degree in Computer Science, IT, or related field
Bachelor's degree in Computer Science, IT, or related field
AWS Certified Data Engineer
Databricks Certificate
Scaled Agile SAFe certification

Tools

Databricks
PySpark
SparkSQL
Apache Spark
AWS services
Python
SQL

Job description

ABOUT AMGEN

Amgen harnesses the best of biology and technology to fight the world's toughest diseases, making people's lives easier, fuller, and longer. We discover, develop, manufacture, and deliver innovative medicines to help millions of patients. Amgen helped establish the biotechnology industry more than 40 years ago and remains on the cutting edge of innovation, using technology and human genetic data to push beyond what's known today.

ABOUT THE ROLE

Let's do this. Let's change the world. We are looking for highly motivated expert Senior Data Engineer who can own the design & development of complex data pipelines, solutions and frameworks with detailed functional knowledge of R&D. The ideal candidate will be responsible to design , develop , and optimize data pipelines, data integration frameworks, and metadata-driven architectures that enable seamless data access and analytics. This role prefers deep expertise in big data processing, distributed computing, data modeling, and governance frameworks to support self-service analytics, AI-driven insights, and enterprise-wide data management.

Roles & Responsibilities
  • Design, develop, and maintain scalable ETL/ELT pipelines to support structured, semi-structured, and unstructured data processing across the Enterprise Data Engineering for Biotech or Pharma functional knowledge of R&D.
  • Implement real-time and batch data processing solutions, integrating data from multiple sources into a unified, governed data fabric architecture.
  • Optimize big data processing frameworks using Apache Spark, Hadoop, or similar distributed computing technologies to ensure high availability and cost efficiency.
  • Work with metadata management and data lineage tracking tools to enable enterprise-wide data discovery and governance.
  • Ensure data security, compliance, and role-based access control (RBAC) across data environments.
  • Optimize query performance, indexing strategies, partitioning, and caching for large-scale data sets.
  • Develop CI/CD pipelines for automated data pipeline deployments, version control, and monitoring.
  • Implement data virtualization techniques to provide seamless access to data across multiple storage systems.
  • Collaborate with cross-functional teams, including data architects, business analysts, and DevOps teams, to align data engineering strategies with enterprise goals.
  • Stay up to date with emerging data technologies and best practices, ensuring continuous improvement of Enterprise Data Fabric architectures.
Must-Have Skills
  • Hands-on experience in data engineering technologies such as Databricks, PySpark , SparkSQL Apache Spark, AWS, Python, SQL, and Scaled Agile methodologies.
  • Proficiency in workflow orchestration, performance tuning on big data processing.
  • Strong understanding of AWS services
  • Experience with Data Fabric, Data Mesh, or similar enterprise-wide data architectures.
  • Ability to quickly learn, adapt and apply new technologies
  • Strong problem-solving and analytical skills
  • Excellent communication and teamwork skills
  • Experience with Scaled Agile Framework ( SAFe ), Agile delivery practices, and DevOps practices.
Good-to-Have Skills
  • Good to have deep expertise in Biotech & Pharma industries
  • Experience in writing APIs to make the data available to the consumers
  • Experienced with SQL/NOSQL database, vector database for large language models
  • Experienced with data modeling and performance tuning for both OLAP and OLTP databases
  • Experienced with software engineering best-practices, including but not limited to version control (Git, Subversion, etc.), CI/CD (Jenkins, Maven etc.), automated unit testing, and Dev Ops
Education and Professional Certifications

Master's degree and 3 to 4 + years of relevant Computer Science, IT or related field experience
OR
Bachelor's degree and 5 to 8 + years of relevant Computer Science, IT or related field experience

  • AWS Certified Data Engineer preferred
  • Databricks Certificate preferred
  • Scaled Agile SAFe certification preferred
Soft Skills
  • Excellent analytical and troubleshooting skills.
  • Strong verbal and written communication skills
  • Ability to work effectively with global, virtual teams
  • High degree of initiative and self-motivation.
  • Ability to manage multiple priorities successfully.
  • Team-oriented, with a focus on achieving team goals.
  • Ability to learn quickly, be organized and detail oriented.
  • Strong presentation and public speaking skills .
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr Data Engineer
Sr Data Engineer

Amgen SA • Hyderabad

On-site
INR 2,500,000 - 4,000,000
Sr Machine Learning Engineer
Sr Machine Learning Engineer

Amgen • India

On-site
INR 1,500,000 - 2,100,000
Sr. Data Engineer - Clinical Data Foundation
Sr. Data Engineer - Clinical Data Foundation

Amgen • India

On-site
INR 3,000,000 - 6,000,000
Sr. Associate Data Engineer
Sr. Associate Data Engineer

Amgen SA • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Sr. Associate Data Engineer
Sr. Associate Data Engineer

Amgen • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Associcate Data Engineer
Associcate Data Engineer

Biopharma Careers • Hyderabad

On-site
INR 3,000,000 - 5,500,000
Sr. Data Engineer – Clinical Data Foundation
Sr. Data Engineer – Clinical Data Foundation

Amgen SA • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Associate Data Engineer
Associate Data Engineer

Amgen SA • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Sr Mgr Data Engineer
Sr Mgr Data Engineer

Amgen Inc • Hyderabad

On-site
INR 800,000 - 1,200,000
Sr Associate Software Engineer
Sr Associate Software Engineer

Amgen SA • Hyderabad

On-site
INR 1,000,000 - 1,500,000