Senior Site Reliability Engineer SRE Kubernetes

Accenture in the Philippines

Cebu City

On-site

PHP 900,000 - 1,500,000

Full time

5 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Accenture in the Philippines seeks an experienced Site Reliability/DevOps engineer to identify and resolve issues across complex business systems. You will analyze root causes, drive reliability improvements, and collaborate with core engineering teams to prevent outages.

You’ll build tools and dashboards to aid front-line support and influence product reliability through postmortems and engineering input. The role emphasizes cloud-native tech, Kubernetes, Linux, observability, and strong

Qualifications

  • 4–8+ years in SRE/DevOps/Platform/Systems Engineering.
  • Experience supporting enterprise platforms and cloud-native tech.
  • Automation, observability and incident response skills.
  • Experience with AI/ML is a plus.

Responsibilities

  • Act as software detectives, identifying and solving issues in critical systems.
  • Perform deep root cause analysis across configuration, bugs, and infra.
  • Design, implement, and monitor reliability improvements and automation.
  • Bridge customer impact with engineering and support teams.
  • Develop tools, playbooks, and dashboards to speed resolution.
  • Contribute to postmortems and define preventative actions.

Skills

Root cause analysis
Incident management
SRE / DevOps experience
Customer engagement
Automation & tooling
SLO / metrics understanding

Tools

Kubernetes
Linux
Networking
Istio
Prometheus
Grafana
Loki
Splunk
Google Compute

Job description

Job Description:

Act as software detectives, provide a dynamic service identifying and solving issues within multiple components of critical business systems.

Deep Technical Support & Root Cause Analysis
  • Handle complex customer issues, going deep into root cause analysis that spans customer configuration, product bugs, and underlying infrastructure.
Proactive Reliability Improvement
  • Analyse patterns of customer issues, support cases, and outage impacts to identify systemic weaknesses.
  • Design and implement solutions, automation, or monitoring to prevent future occurrences (involving coding, configuration changes, or proposing architectural improvements).
Bridging Customer Impact and Engineering
  • Act as a liaison between the customer-facing support teams and the core Support/Development teams.
  • Translate customer pain into technical requirements and SLOs, and explain technical constraints and incident impacts back to support teams.
Improving Supportability
  • Develop tools, playbooks, and dashboards to help front-line support and diagnose and resolve issues more quickly.
  • Feedback into the product development lifecycle to ensure new features are designed with supportability and reliability in mind.
Customer-Centric Service Level Objectives (SLOs)
  • Contribute to defining and refining SLOs to better reflect actual customer pain and perceived performance, not just server-side metrics.
Incident Management & Postmortems
  • Participate in incident response, bringing a strong understanding of customer impact.
  • Contribute significantly to postmortems, ensuring preventative actions address both the technical root cause and the customer's experience.
Proactive Customer Engagement (for Key Customers)
  • Engage with key customers to understand their critical workloads, review their architecture, and provide guidance on reliability best practices.
Essential Experience & Qualifications
  • 4-8+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Systems Engineering.
  • Experience supporting enterprise platforms.
  • Experience with cloud-native technologies and automation.
  • Experience with AI / ML
Desirable
  • Certified Kubernetes Administrator (CKA) / Certified Kubernetes Application Developer (CKAD)
  • Google Professional Cloud DevOps Engineer
  • Google Professional Cloud Network Engineer / VMware Certified Professional - Network Virtualization (2V0-41.24)
  • Linux Foundation Certified System Administrator (LFCS) / Linux Foundation Certified IT Associate (LFCA) / Linux - ------- Professional Institute LPIC-3 Mixed Environments
  • Prometheus Certified Associate (PCA)
Required Technical Skills
Infrastructure & Platforms
  • Kubernetes (essential)
  • Linux (high priority)
  • Networking (high priority)
  • Service Mesh Basics (Istio) (high priority)
  • PKI (high priority)
  • AI / ML (high priority)
  • Google Compute
  • Storage Technologies (optional)
  • Observability & Monitoring
  • Prometheus (high priority)
  • Grafana (high priority)
  • Loki (high priority)
  • Splunk
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Devops Engineer
Senior Devops Engineer

V2 Solutions • Hinoba-an

Hybrid
PHP 1,200,000 - 2,400,000
Site Reliability / Cloud Platform Engineer
Site Reliability / Cloud Platform Engineer

Global Recruitment and Consultancy OPC • Cebu City

On-site
PHP 1,200,000 - 2,400,000
Site Reliability Engineer
Site Reliability Engineer

V2 Solutions • Hinoba-an

On-site
PHP 900,000 - 1,500,000
Staff SRE Engineer
Staff SRE Engineer

Stellar Cyber • España

On-site
PHP 5,528,000 - 7,372,000
VS01700 - SRE & Production Reliability Engineer
VS01700 - SRE & Production Reliability Engineer

E4 Software Services Pvt Ltd. • Hinoba-an

On-site
PHP 893,000 - 1,674,000
Lead DevOps Engineer
Lead DevOps Engineer

V2 Solutions • Hinoba-an

On-site
PHP 1,004,000 - 1,451,000
Site Reliability Engineer
Site Reliability Engineer

Philtech Inc. • Taguig

On-site
PHP 670,000 - 1,339,000
Health insurance
Retirement plans
Career growth opportunities
Site Reliability Engineer
Site Reliability Engineer

Pyramid Consulting, Inc • Mexico

On-site
PHP 5,618,000 - 8,115,000
Senior Platform Engineer (OpenShift, Kubernetes & GitOps)
Senior Platform Engineer (OpenShift, Kubernetes & GitOps)

Ford • Hinoba-an

On-site
PHP 1,800,000 - 2,400,000
Software Engineer- AI-Driven SRE & Cloud SRE
Software Engineer- AI-Driven SRE & Cloud SRE

Keka Technologies Private Limited • Mexico

On-site
PHP 2,215,000 - 3,322,000