Site Reliability Engineer – Data Platforms

Amgen Inc. (IR)

Hyderabad

On-site

INR 1,500,000 - 2,100,000

Full time

3 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Amgen Inc. (IR) in Hyderabad is seeking a Site Reliability Engineer for EDSE Data Platforms. You will balance governance with hands-on AWS SRE to ensure reliable, scalable infrastructure and seamless service handovers.

The role focuses on building IaC, automation, observability, incident response, and collaboration with security, product, and business teams to drive reliability and resilience.

Qualifications

  • Master's or Bachelor’s degree in Computer Science, Engineering, Information Systems, or related field
  • 5–8 years of relevant experience

Responsibilities

  • Service governance and operating model – establish and maintain a consistent handover process with readiness criteria covering ownership, architecture, dependencies, monitoring, security findings, runbooks, recovery, risks and knowledge transfer.
  • Hands-on AWS site reliability engineering – operate and improve secure, reliable, scalable AWS infrastructure, build IaC, delivery pipelines, automation, observability, incident response, and recovery.

Skills

AWS infrastructure
Networking & IAM
Security & Compliance
Observability & Monitoring
Infrastructure as Code
CI/CD pipelines
Incident response
Jira/Jira Align
Stakeholder management
Hands-on SRE

Education

Bachelor's degree in Computer Science or related
Master's degree in CS or related

Tools

Docker
Kubernetes
AWS
Terraform
Jenkins
Git

Job description

Career Category Engineering Job Description

Join Amgen’s Mission of Serving Patients. At Amgen, if you feel like you’re part of something bigger, it’s because you are. Our shared mission‑to serve patients living with serious illnesses‑drives all that we do. Since 1980, we’ve helped pioneer the world of biotech in our fight against the world’s toughest diseases. With our focus on four therapeutic areas − Oncology, Inflammation, General Medicine, and Rare Disease, we reach millions of patients each year. As a member of the Amgen team, you’ll help make a lasting impact on the lives of patients as we research, manufacture, and deliver innovative medicines to help people live longer, fuller happier lives. Our award‑winning culture is collaborative, innovative, and science based. If you have a passion for challenges and the opportunities that lay within them, you’ll thrive as part of the Amgen team. Join us and transform the lives of patients while transforming your career.

About Amgen

Amgen harnesses the best of biology and technology to fight the world’s toughest diseases, and make people’s lives easier, fuller and longer. We discover, develop, manufacture and deliver innovative medicines to help millions of patients. Amgen helped establish the biotechnology industry more than 40 years ago and remains on the cutting‑edge of innovation, using technology and human genetic data to push beyond what’s known today.

About the Role

EDSE Data Platforms is seeking a Site Reliability Engineer who is equally comfortable improving production systems in AWS and establishing how services should be operated. Our SRE team manages AWS platform operations and the infrastructure supporting applications across EDSE. As our service portfolio grows, we need an experienced individual contributor who can create clear, practical operating arrangements across SRE, application teams, Security, business stakeholders, and other service partners. This is a deliberately balanced role, with responsibilities split approximately: 50% service governance, operational readiness, resilience, and process leadership; 50% hands‑on AWS site reliability engineering. The balance will be measured over a typical quarter and may vary during major incidents, application handovers, continuity exercises, audits, and delivery periods. This is not a policy‑only, PMO, or ITSM process‑administration role. You will remain hands‑on with AWS infrastructure, automation, observability, incident response, and reliability improvement while creating lightweight, enforceable operational controls.

Roles & Responsibilities
  • Service governance and operating model − approximately 50%Establish and operate a consistent application handover process, with readiness criteria covering ownership, architecture, dependencies, monitoring, access, security findings, runbooks, recovery, known risks, knowledge transfer, hyper‑care, and formal acceptance. Identify incomplete or unsafe handovers and recommend deferring acceptance. Ensure exceptions are documented, time‑bound, and approved by the appropriate business, service, security, or risk owner. Maintain a service catalogue and operating profile for every supported application, including criticality, owners, recurring tasks, dependencies, support hours, access requirements, backup coverage, recovery objectives, and escalation paths. Define clear responsibilities, decision rights, support boundaries, and segregation of duties across SRE, application engineering, Product, Security, business owners, vendors, and other stakeholders. Agree service indicators and internal objectives, proposed SLAs, support hours, severity definitions, response and restoration expectations, maintenance windows, recovery expectations, and escalation paths.
  • Hands‑on AWS site reliability engineering − approximately 50%Operate and improve secure, reliable, scalable, and cost‑effective application infrastructure on AWS. Build and maintain infrastructure as code, delivery pipelines, and automation that removes repetitive operational effort. Improve observability through meaningful metrics, logs, traces, dashboards, alerts, and service‑health reporting. Participate in on‑call and incident response, including diagnosis, mitigation, recovery, communication, escalation, and blameless post‑incident reviews. Improve availability, performance, capacity, resilience, backup and recovery, patching, infrastructure lifecycle management, and security posture. Partner with application teams on architecture, production readiness, deployment safety, application resilience, and operational quality. Apply SRE practices such as SLIs, SLOs, error budgets, automation, and learning from failure to reduce repeat incidents, alert fatigue, manual toil, and recovery time.
What we expect of you
Basic Qualifications and Experience
  • Master's or Bachelor’s degree in Computer Science, Engineering, Information Systems, or related field
  • 5–8years of relevant experience
Must‑Have Skills
  • Strong understanding of AWS infrastructure, networking, identity and access management, security, monitoring, backup, and resilience patterns with depth in the services most relevant to data platforms: IAM, VPC, PrivateLink, S3, EC2, KMS, CloudWatch, CloudTrail, Secrets Manager, and STS.
  • Working knowledge of EKS, Lambda, Glue, EMR, RDS will be good to have.
  • Experience in operating business‑critical production services on AWS.
  • Strong hands‑on experience in Site Reliability Engineering, Platform Engineering, Cloud Operations, or Infrastructure Engineering.
  • Practical experience with infrastructure as code, CI/CD, operational automation, monitoring, observability, and incident response.
  • Demonstrated experience establishing or improving service handovers, support models, operational ownership, service levels, business continuity, disaster recovery, or security‑control processes.
  • Experience defining and testing recovery objectives, recovery plans, and technical recovery procedures.
  • Understanding of security controls, audit evidence, access governance, vulnerability management, risk acceptance, and segregation of duties.
  • Ability to convert unclear cross‑team responsibilities into practical, agreed, and measurable operating arrangements.
  • Strong facilitation, negotiation, documentation, and stakeholder‑management skills.
  • Ability to influence technical, business, security, and leadership stakeholders without relying on direct reporting authority.
  • A pragmatic approach to governance that improves accountability and reliability without introducing unnecessary bureaucracy.
  • Strong analytical and problem‑solving ability, with a record of turning complex or ambiguous platform challenges into scalable and reusable solutions.
  • Experience working in Agile delivery environments and using planning and delivery tools such as Jira or Jira Align.
Good‑to‑Have Skills
  • Experience supporting a portfolio of applications with different criticality and support requirements.
  • Experience in a regulated, security‑sensitive, or audit‑intensive environment.
  • Familiarity with frameworks such as AWS Well‑Architected, ITIL, and business‑continuity standards.
  • Experience with containers and orchestration platforms such as Docker, Amazon ECS, Amazon EKS, or Kubernetes.
Relevant Certifications
  • AWS Certified Solutions Architect
  • AWS Certified Security – Specialty
  • SAFe Agilist or another SAFe certification
Functional Skills
  • Excellent written and verbal communication, with the ability to explain complex platform concepts, architecture decisions, risks, and trade‑offs in clear, business‑relevant language.
  • Strong influencing and consensus‑building skills, including the ability to establish standards and drive adoption across teams without relying solely on formal authority.
  • A platform‑product mindset focused on reusable capabilities, paved roads, developer experience, measurable outcomes, and long‑term platform health.
  • Strong systems‑thinking and structured problem‑solving skills, with the ability to diagnose issues across application, platform, cloud, governance, security, and operating‑model boundaries.
  • High degree of ownership, initiative, and follow‑through, with the ability to move ambiguous topics from exploration through decision, implementation, adoption, and continuous improvement.
  • Collaborative and globally minded, with experience working effectively across architecture, governance, cybersecurity, operations, engineering, product, and business teams.
  • Strong planning, estimation, prioritization, and execution skills, with the ability to manage multiple initiatives while maintaining high standards for security, reliability, quality, and reusability.
  • Ability to balance innovation with enterprise risk, distinguishing between experimentation, limited preview adoption, and production‑ready capabilities.
  • Strong coaching and enablement skills, with the ability to create clear documentation, facilitate technical workshops, mentor engineers, and build an active platform community.
  • A growth mindset and commitment to continuous learning, modern engineering practices, constructive challenge, and responsible adoption of AI.
What you can expect of us

In addition to the base salary, Amgen offers competitive and comprehensive Total Rewards Plans that are aligned with local industry standards.

EQUAL OPPORTUNITY STATEMENT

Amgen is an Equal Opportunity employer and will consider all qualified applicants for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, protected veteran status, disability status, or any other basis protected by applicable law. We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, to perform essential job functions, and to receive other benefits and privileges of employment. Please contact us to request accommodation.

GCF Level

GCF Level GCF Level 04 Career Category Engineering Position Type Full time . Amgen is committed to unlocking the potential of biology for patients suffering from serious illnesses by discovering, developing, manufacturing and delivering innovative human therapeutics. This approach begins by using tools like advanced human genetics to unravel the complexities of disease and understand the fundamentals of human biology. Amgen focuses on areas of high unmet medical need and leverages its biologics manufacturing expertise to strive for solutions that improve health outcomes and dramatically improve people’s lives. A biotechnology pioneer since 1980, Amgen has grown to be one of the world's leading independent biotechnology companies, has reached millions of patients around the world and is developing a pipeline of medicines with breakaway potential. For more information, visit www.amgen.com and follow us on www.twitter.com/amgen

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Manager, Site Reliability Engineer - Data Platforms
Manager, Site Reliability Engineer - Data Platforms

Amgen Inc. (IR) • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Site Reliability Engineer – Data Platforms
Site Reliability Engineer – Data Platforms

Amgen SA • Hyderabad

On-site
INR 3,500,000 - 7,000,000
Manager, Site Reliability Engineer - Data Platforms
Manager, Site Reliability Engineer - Data Platforms

Amgen SA • Hyderabad

On-site
INR 3,000,000 - 4,500,000
Sr Associate Software Engineer Cloud Platform DevOps
Sr Associate Software Engineer Cloud Platform DevOps

Amgen • Hyderabad

On-site
INR 1,800,000 - 3,000,000
Senior Data Engineer
Senior Data Engineer

Amgen Inc. (IR) • Hyderabad

On-site
INR 3,500,000 - 6,000,000
Senior Data Engineer
Senior Data Engineer

Amgen • Hyderabad

On-site
INR 1,200,000 - 2,000,000
Competitive benefits
Collaborative culture
Professional development opportunities
Specialist IS Engineer
Specialist IS Engineer

Amgen Inc. (IR) • Hyderabad

On-site
INR 3,000,000 - 5,000,000
Specialist Software Engineer (Full Stack)
Specialist Software Engineer (Full Stack)

Amgen Inc. (IR) • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Sr Associate- Secure Data Exchange Services
Sr Associate- Secure Data Exchange Services

Amgen SA • Hyderabad

On-site
INR 1,200,000 - 2,100,000
Sr. Associate Data Engineer
Sr. Associate Data Engineer

Amgen • Hyderabad

On-site
INR 1,200,000 - 1,800,000