Site Realibility Engineer

Gratitude Philippines

Quezon

On-site

PHP 914,000 - 1,095,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Gratitude Philippines is seeking a Site Reliability Engineer for a hybrid setup (2 WFH, 3 RTO) in Manila area. Night shift, immediate start, 1 headcount.

The role focuses on enhancing reliability across platforms using Azure Monitor, Grafana, and ServiceNow ITOM; require strong SRE background and hands-on observability experience. You will collaborate with IT Ops, platform, cyber, network, and product teams to design standards, reduce alert noise, and embed self-healing patterns where they

Qualifications

  • Bachelor’s degree in IT/CS/Engineering required.
  • 3+ years in monitoring/SRE with Azure Monitor/App Insights (KQL) and ServiceNow Event Management.
  • 5+ years in SRE with strong monitoring and application performance management.

Responsibilities

  • Define standards, patterns, and automations for monitoring and reliability focusing on Azure Monitor, ServiceNow ITOM, Grafana.
  • Partner with product teams to implement SLO/SLI driven operations and reduce alert noise.
  • Engineer monitoring patterns and runbooks with actionable alerts and incident routing.
  • Contribute to Monitoring strategy, governance and tooling.
  • Communicate to drive continuous improvement by reducing MTTR and embedding postmortem learnings.

Skills

Azure Monitor
ServiceNow ITOM
Grafana
KQL
SRE practices
Monitoring & Observability
Network security

Education

Bachelor's degree in IT/CS/Engineering

Tools

Azure Monitor
App Insights
Log Analytics
Grafana
Prometheus
App Dynamics
ThousandEyes
ServiceNow

Job description

Job title: Site Realibility Engineer
Work set up: Hybrid: 2 WFH & 3 RTO (Location: Manila (Eton Centris, Quezon Avenue, Quezon City))
Work shift: Night Shift
Salary: P90,000
Start date: ASAP
Headcount:1

Qualifications:
  • Bachelor’s degree in IT /Computer science/Engineering, or related field.
  • 3+ years in monitoring/observability/SRE roles with hands‑on experience in Azure Monitor/App Insights (KQL) and ServiceNow Event Management.
  • Strong knowledge in Azure Log Analytics, KQL, Telemetry, APM implementations
  • Demonstrated ability to collaborate across IT Operations team, platform, cyber, network, and product teams, strong written verbal communication for standards and enablement.
  • 5+ years of experience with SRE role and deep understanding of monitoring and application performance management
  • Knowledge of SLO platforms (e.g., Nobl9) and experience contributing to standards/governance artifacts.
  • Knowledge of proactive monitoring using Azure monitor services, telemetry, and synthetic transactions.
  • Understanding of network architecture and security: WAN/LAN, TCP/IP, PKI.
  • must have ITSM processes and tools ( ServiceNow), and compliance processes hands on experience
  • Have AIOps vision and awareness
Responsibilities:
  • You will design and define standards, patterns, and automations opportunities that elevate monitoring and reliability across platforms and applications, with a strong focus on Azure Monitor, ServiceNow ITOM Event Management, Grafana, and APM/Synthetics tooling
  • You’ll partner with product teams to implement SLO/SLI‑driven operations, reduce alert noise, accelerate incident response, and embed self‑healing where it matters most.
  • Engineer enterprise monitoring & event patterns by authoring and maintaining reference architectures, runbooks, and event management models (alert → event → incident) with actionable alerts and incidents routing.
  • Contribute to Monitoring and Observability & Event Management Strategy and tooling intake/governance checkpoints and coach product teams
  • Excellent communication skills to drive continuous improvement by reducing alert noise, shorten MTTR, and improve change success by embedding postmortem learnings into patterns, rules, and pipelines.
Must have Skills:
  • Cloud Observability: Azure Monitor/App Insights/Log Analytics (KQL)
  • Knowledge of Grafana, Prometheus, App Dynamics, ThousandEyes
  • Communication & Teaming – Able to translate complex reliability patterns into consumable standards and coach IT operations team via office hours/CoP sessions.
  • Technical Depth in Monitoring and Observability Stack – Hands‑on in ServiceNow Event Management, Azure Monitor/KQL, and automation.
  • Analytical & Systems Thinking – Uses SLI/SLOs, postmortems, and CMDB context to reduce noise, drive self‑healing, and measurably improve MTTR and KPIs.
Recruitment Process:
  • Paper screening (endorsing profile to Operations team to identify if candidate meets the basic requirement/qualification for the position)
  • L1 Interview (Interview process with Practice/Operations Team)
  • L2 Interview (Optional)
  • Technical assessment/Final Interview (Conducted by Customer’s Operations team)
Pre-screening Notes:
  • Highest Educational Attainment?
  • How many years of relevant experience do you have monitoring/observability/SRE roles with hands‑on experience in Azure Monitor/App Insights (KQL) and ServiceNow Event Management?
  • How many years of relevant experience do you have with SRE role and deep understanding of monitoring and application performance management?
  • How much is your last drawn salary?
  • How much is your salary expectation?
  • Are you amenable to work Hybrid in Quezon City with a Night Shift schedules?
  • When are you available to start once hired?
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Azure Site Reliability Engineer
Azure Site Reliability Engineer

GSS HR Solutions Private Limted • Quezon City

Hybrid
PHP 600,000 - 900,000
Site Reliability Engineer | Hybrid - Centris/Makati
Site Reliability Engineer | Hybrid - Centris/Makati

TASQ • Makati

Hybrid
PHP 1,000,000 - 1,800,000
Now Hiring: Site Reliability Engineer | Hybrid Quezon City
Now Hiring: Site Reliability Engineer | Hybrid Quezon City

Gratitude Philippines • Manila

On-site
PHP 900,000 - 1,200,000
Principal Site Reliability Engineer (SRE)
Principal Site Reliability Engineer (SRE)

Lewis Personnel Management • Makati

Hybrid
PHP 1,800,000 - 3,200,000
Site Realibility Engineer
Site Realibility Engineer

Gratitude Philippines • Manila

On-site
PHP 600,000 - 1,200,000
Monitoring, Observability & Event Management Architect
Monitoring, Observability & Event Management Architect

Gratitude Philippines • Quezon

Hybrid
PHP 2,100,000 - 3,600,000
Site Reliability Engineer- (Monitoring and Event Management)
Site Reliability Engineer- (Monitoring and Event Management)

Pan Asia Resources • Philippines

On-site
PHP 1,200,000 - 2,000,000
Azure SRE & Observability Engineer — Night Shift (Hybrid)
Azure SRE & Observability Engineer — Night Shift (Hybrid)

GSS HR Solutions Private Limted • Quezon City

Hybrid
PHP 600,000 - 900,000
Azure Observability SRE - Night Shift, Hybrid Manila
Azure Observability SRE - Night Shift, Hybrid Manila

Gratitude Philippines • Quezon

Hybrid
PHP 914,000 - 1,095,000
IT Solution Architect (ServiceNow ITOM) | Hybrid - Centris/Makati
IT Solution Architect (ServiceNow ITOM) | Hybrid - Centris/Makati

TASQ • Makati

Hybrid
PHP 1,800,000 - 3,200,000