Senior SRE Lead - Warehousing IT Operations

Procter & Gamble

Manila

On-site

PHP 1,451,000 - 2,120,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Procter & Gamble in Taguig City is seeking a Site Reliability Engineer (SRE) for the Warehousing IT Operations – Incident Response Team. The role leads incident response, enhances reliability, scalability, and performance of critical systems, and collaborates with software engineers, DevOps, and stakeholders to automate monitoring and preventive measures.

This is a managerial position requiring leadership of teams and end-to-end processes, driving data-driven decisions, and delivering

Qualifications

  • Knowledge or familiarity in system administration, including Linux/Unix environments, cloud platforms (such as AWS, Azure, or GCP) and SAP.
  • Experience with configuration management tools and infrastructure-as-code frameworks (e.g., Terraform).
  • Proficiency in at least one programming language (e.g., Python, C#) and experience with scripting for automation tasks.
  • Understanding of networking protocols, network infrastructures, load balancing, and DNS management.
  • Familiarity with containerization and orchestration technologies (e.g., Docker, Kubernetes).
  • Familiarity with databases and proficiency in writing SQL queries.
  • Experience or familiarity with monitoring and observability tools (e.g., Prometheus, Grafana).
  • Knowledge of incident response methodologies, root cause analysis, and implementing preventive measures.
  • Understanding of security best practices and experience with implementing secure systems.
  • Experience in Warehousing Management Systems (e.g. RTCIS, PrIME) or Warehousing Operations is a plus.

Responsibilities

  • Incident Response: Lead incident response efforts, swiftly resolving critical incidents to minimize downtime and user impact. Implement effective incident management processes, ensuring clear communication, coordination, and documentation. Conduct root cause analysis, implementing preventive measures and driving continuous improvement.
  • Reliability: Ensure high system availability through robust monitoring, alerting, and automated incident response systems. Optimize system architecture and configurations for improved performance, scalability, and fault tolerance. Collaborate cross-functionally to design and implement resilient systems using industry best practices. Implement comprehensive monitoring solutions, providing real-time insights into system performance and health. Configure and manage monitoring tools, ensuring accurate and actionable alerts for proactive incident response. Continuously evaluate and enhance monitoring strategies to improve system visibility and resource optimization.
  • Upskilling: Stay updated with industry trends, technologies, and best practices in Site Reliability Engineering. Continuously develop technical skills in system architecture, automation, cloud technologies, and incident response. Share knowledge, mentor team members, and foster a culture of learning and upskilling.
  • Managing Users/Customers’ Needs and Expectations: Collaborate directly with users and customers to understand their needs and pain points. Proactively address customer/user concerns, ensuring reliable and performant systems. Provide exceptional customer support, communicate updates, resolutions, and gather feedback for continuous improvement.
  • Monitoring: Implement comprehensive monitoring solutions, providing real-time insights into system performance and health. Configure and manage monitoring tools, ensuring accurate and actionable alerts for proactive incident response. Continuously evaluate and enhance monitoring strategies to improve system visibility and resource optimization.
  • Upskilling: Stay updated with industry trends, technologies, and best practices in Site Reliability Engineering. Continuously develop technical skills in system architecture, automation, cloud technologies, and incident response. Share knowledge, mentor team members, and foster a culture of learning and upskilling.
  • Managing Users/Customers’ Needs and Expectations: Collaborate directly with users and customers to understand their needs and pain points. Proactively address customer/user concerns, ensuring reliable and performant systems. Provide exceptional customer support, communicate updates, resolutions, and gather feedback for continuous improvement.

Skills

Linux/Unix
Cloud platforms (AWS/Azure/GCP)
Python/C#
Incident response
Automation
Team leadership
SRE fundamentals
Warehousing systems knowledge

Tools

Terraform
Docker
Kubernetes
Prometheus
Grafana

Job description

Procter & Gamble in Taguig City is seeking a Site Reliability Engineer (SRE) for the Warehousing IT Operations – Incident Response Team. The role leads incident response, enhances reliability, scalability, and performance of critical systems, and collaborates with software engineers, DevOps, and stakeholders to automate monitoring and preventive measures.

This is a managerial position requiring leadership of teams and end-to-end processes, driving data-driven decisions, and delivering

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

SRE Manager — Incident Response & Reliability Leader
SRE Manager — Incident Response & Reliability Leader

Procter & Gamble • Philippines

On-site
PHP 1,200,000 - 1,800,000
Senior SRE Lead — Incident Response & Reliability (Remote)
Senior SRE Lead — Incident Response & Reliability (Remote)

Procter & Gamble • Philippines

On-site
PHP 1,200,000 - 1,500,000
Performance bonus (STAR)
Flexible work schedule with work-from-
Health insurance
+2
SRE Lead: Incident Response & Reliability (Remote)
SRE Lead: Incident Response & Reliability (Remote)

Procter & Gamble • Manila

On-site
PHP 2,400,000 - 3,600,000
Performance bonus (STAR program)
Flexible work schedule with work-from-
Site Reliability Engineer — Incidents, Monitoring & Scale
Site Reliability Engineer — Incidents, Monitoring & Scale

Procter & Gamble • Philippines

On-site
PHP 1,000,000 - 1,400,000
Performance bonus
Flexible work schedule / work from hom
Health insurance
+2
Site Reliability Engineer – Incident Response & Resilience
Site Reliability Engineer – Incident Response & Resilience

Procter & Gamble • Manila

On-site
PHP 1,500,000 - 1,900,000
Performance bonus
Flexible work schedule / work fromhome
Health insurance
+2
Site Reliability Engineer - Warehousing IT Operations
Site Reliability Engineer - Warehousing IT Operations

Procter & Gamble • Philippines

On-site
PHP 1,200,000 - 1,800,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Procter & Gamble • Philippines

On-site
PHP 1,200,000 - 1,500,000
Performance bonus (STAR)
Flexible work schedule with work-from-
Health insurance
+2
Site Reliability Engineer - Warehousing IT Operations
Site Reliability Engineer - Warehousing IT Operations

Procter & Gamble • Manila

On-site
PHP 1,451,000 - 2,120,000
Identity & Access SRE — Site Reliability Engineer
Identity & Access SRE — Site Reliability Engineer

Procter & Gamble Philippines • Manila

On-site
PHP 900,000 - 1,300,000
Performance bonus
Flexible work schedule / work fromHome
Health insurance & wellness programs
Site Reliability Engineer
Site Reliability Engineer

Procter & Gamble • Philippines

On-site
PHP 1,000,000 - 1,400,000
Performance bonus
Flexible work schedule / work from hom
Health insurance
+2