Turn this role into an interview — a resume and cover letter built around what this employer wants.
REED ELSEVIER SHARED SERVICES (PHILIPPINES) INC. is seeking a Site Reliability Engineer (AWS) to enhance resilience and reliability of cloud platforms.
You will partner with architecture and cloud teams to design robust recovery and testing programs and to implement resilience controls across AWS environments. The role requires strong hands-on AWS expertise, experience in disaster recovery, and a proactive approach to continuous improvement, reporting into an SRE leadership team in a hybrid
Flexible hybrid work setup
IT Equipment provided
HMO coverage starting from Day 1 for you and FOUR FREE dependents
Attractive retirement package with company matching
Life and Accident Insurance starting Day 1
24 Annual PTOs, additional 6 once you reach your 5th year with us
Competitive benefits with annual merit increase and incentives
Continuous improvement for our employees (workshops, certification programs, learning sessions, etc.)
Set-Up: Hybrid (2/month onsite reporting)
Schedule: Midshift 3:00 PM – 11:00 PM, Monday to Friday
Location: UP AyalaLand Technohub, Commonwealth, Quezon City
The Site Reliability Engineer - AWS will play a key role in improving the resilience, recoverability, and reliability of AWS-based platforms and services. This role requires strong hands‑on AWS expertise combined with experience in cloud infrastructure, reliability engineering, disaster recovery, and resilience testing. The successful candidate will work closely with architecture, cloud engineering, platform engineering, and Site Reliability Engineering (SRE) teams to design, validate, and continuously improve resilient cloud solutions that support critical business services.
Partner with architecture, cloud engineering, and platform teams to improve the resiliency, availability, and recoverability of AWS-hosted services.
Design, review, and validate disaster recovery and resilience strategies for AWS environments.
Develop and execute resilience testing activities, including failover testing, recovery validation, disaster recovery exercises, and resilience assessments.
Evaluate cloud architectures for reliability risks, single points of failure, and opportunities to improve resilience and availability.
Support implementation of resilience controls and recovery mechanisms across AWS environments.
Collaborate with engineering teams to embed resilience and reliability principles into solution designs and operational processes.
Analyze resilience incidents, recovery outcomes, and testing results to identify and implement continuous improvement opportunities.
Track recovery objectives (RTO/RPO), resilience metrics, and testing outcomes to ensure recovery requirements are met.
Provide technical leadership and guidance on AWS reliability, disaster recovery, high availability, and resilience best practices.
Contribute to the development and adoption of resilience standards, patterns, and reusable solutions across the organization.
Bachelor’s degree.
5 years of hands‑on experience as Site Reliability Engineer for AWS.
10+ years of hands‑on experience with AWS services (e.g., EC2, EKS, RDS, IAM, VPC, Multi-AZ architectures).
2–5 years of hands‑on experience in resilience testing, disaster recovery, and recovery strategies.
Strong analytical and problem‑solving skills, with the ability to assess complex systems and risks.
Average to above‑average communication skills with the ability to explain technical concepts to stakeholders.
AWS Certification (e.g., Solutions Architect).
Knowledge of infrastructure-as-code (e.g., Terraform, CloudFormation).
Knowledge of chaos engineering tools and practices.
Familiarity with observability, monitoring, and incident management tools.
Understanding of regulatory and compliance frameworks relevant to resilience.