An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Accenture in the Philippines invites a Senior Site Reliability Engineer to lead deep technical diagnostics across enterprise platforms. You will own root cause analysis, automate reliability improvements, and bridge customer impact with engineering teams.
You will design tools, dashboards, and playbooks to speed incident response, contribute to postmortems, and help define SLOs that reflect customer pain. 4–8+ years in SRE/DevOps, cloud-native experience, and strong automation skills are
Act as software detectives, provide a dynamic service identifying and solving issues within multiple components of critical business systems.
Handle complex customer issues, going deep into root cause analysis that spans customer configuration, product bugs, and underlying infrastructure.
Analyse patterns of customer issues, support cases, and outage impacts to identify systemic weaknesses.
Act as a liaison between the customer-facing support teams and the core Support/Development teams.
Develop tools, playbooks, and dashboards to help front-line support and diagnose and resolve issues more quickly.
Contribute to defining and refining SLOs to better reflect actual customer pain and perceived performance, not just server-side metrics.
Participate in incident response, bringing a strong understanding of customer impact.
Engage with key customers to understand their critical workloads, review their architecture, and provide guidance on reliability best practices.
Infrastructure & Platforms Kubernetes (essential) Linux (high priority) Networking (high priority) Service Mesh Basics (Istio) (high priority) PKI (high priority) AI / ML (high priority) Google Compute Storage Technologies (optional) Observability & Monitoring Prometheus (high priority) Grafana (high priority) Loki (high priority) Splunk Minimum 3 year(s) of experience is required