Site Reliability Engineer | Hybrid - Centris/Makati

Tasq Work

Makati

Hybrid

PHP 1,500,000 - 2,300,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Hybrid work arrangement
Night shift

Job summary

Tasq Work is hiring a Monitoring & Observability Engineer to ensure the reliability of client monitoring and event management systems, with a strong emphasis on Azure Monitor, Grafana, and SLO/SLI-driven operations. You will design standards and automation to reduce noise and accelerate incident response.

You will partner with product teams to implement operational patterns, develop runbooks, and influence governance around observability, incident routing, and self-healing initiatives.

Qualifications

  • Bachelor’s degree in IT / Computer science / Engineering or related field.
  • 3+ years in monitoring/observability/SRE with hands-on Azure Monitor/App Insights (KQL) and ServiceNow Event Management.
  • Strong knowledge in Azure Log Analytics, KQL, Telemetry, and APM implementations.
  • Demonstrated ability to collaborate across IT Operations, platform, cyber, network, and product teams.

Responsibilities

  • Design standards, patterns, and automation opportunities to elevate monitoring and reliability with focus on Azure Monitor, Grafana, and APM tools.
  • Partner with product teams to implement SLO/SLI-driven operations and reduce alert noise.
  • Engineer enterprise monitoring patterns by authoring reference architectures, runbooks, and incident routing models.
  • Contribute to Monitoring and Observability strategy and governance checkpoints and coach product teams.
  • Communicate effectively to drive continuous improvement by embedding postmortem learnings into patterns and pipelines.

Skills

Excellent communication
Team collaboration
Analytical thinking

Education

Bachelor’s degree in IT / Computer Science / Engineering

Tools

Azure Monitor
App Insights
Log Analytics (KQL)
Grafana
Prometheus
App Dynamics
ThousandEyes
ServiceNow ITOM Event Management

Job description

Non-negotiable skills we are looking for:
  • Excellent communication skills to drive continuous improvement by reducing alert noise, shorten MTTR, and improve change success by embedding postmortem learnings into patterns, rules, and pipeline:
  • Cloud Observability: Azure Monitor/App Insights/Log Analytics (KQL)
  • Knowledge of Grafana, Prometheus, App Dynamics, ThousandEyes
  • Uses SLI/SLOs, postmortems, and CMDB and other context to reduce noise, drive self-healing, and measurably improve MTTR and KPIs.
High Level Summary:

Ensure the reliability of a client's critical monitoring and event management systems and services, provide standards and governance around monitoring and observability which enable Chevron business-critical processes to operate safely, efficiently, and reliably.

Key Responsibilities:
  • You will design and define standards, patterns, and automations opportunities that elevate monitoring and reliability across platforms and applications, with a strong focus on Azure Monitor, ServiceNow ITOM Event Management, Grafana, and APM/Synthetics tooling
  • You’ll partner with product teams to implement SLO/SLI-driven operations, reduce alert noise, accelerate incident response, and embed self-healing where it matters most.
  • Engineer enterprise monitoring & event patterns by authoring and maintaining reference architectures, runbooks, and event management models (alert → event → incident) with actionable alerts and incidents routing.
  • Contribute to Monitoring and Observability & Event Management Strategy and tooling intake/governance checkpoints and coach product teamsInternal - General Use
  • Excellent communication skills to drive continuous improvement by reducing alert noise, shorten MTTR, and improve change success by embedding postmortem learnings into patterns, rules, and pipelines.
Technologies and Tools:
  • SRE Practices: Observability and Monitoring
  • Cloud Observability: Azure Monitor/App Insights/Log Analytics (KQL)
  • Grafana/Prometheus for metrics visualization where applicable
  • ServiceNow ITOM Event Management
  • Azure Fundamentals, Azure Monitor
  • DevOps and Automation Tools
  • Grafana, Prometheus, App Dynamics, ThousandEyes
  • Application Performance Monitoring and Digital User Experience tools
Required Qualifications:
  • Bachelor’s degree in IT /Computer science/Engineering, or related field.
  • 3+ years in monitoring/observability/SRE roles with hands-on experience in Azure Monitor/App Insights (KQL) and ServiceNow Event Management.
  • Strong knowledge in Azure Log Analytics, KQL, Telemetry, APM implementations
  • Demonstrated ability to collaborate across IT Operations team, platform, cyber, network, and product teams, strong written verbal communication for standards and enablement.
Preferred Qualifications:
  • 5+ years of experience with SRE role and deep understanding of monitoring and application performance management
  • Knowledge of SLO platforms (e.g., Nobl9) and experience contributing to standards/governance artifacts.
  • Knowledge of proactive monitoring using Azure monitor services, telemetry, and synthetic transactions.
  • Understanding of network architecture and security: WAN/LAN, TCP/IP, PKI.Internal - General Use
  • Familiarity with ITSM processes and tools (e.g., ServiceNow), and compliance processes
  • Have AIOps vision and awareness
Critical Selection Criteria:
  1. Communication & Teaming – Able to translate complex reliability patterns into consumable standards and coach IT operations team via office hours/CoP sessions.
  2. Technical Depth in Monitoring and Observability Stack – Hands-on in ServiceNow Event Management, Azure Monitor/KQL, and automation.
  3. Analytical & Systems Thinking – Uses SLI/SLOs, postmortems, and CMDB context to reduce noise, drive self-healing, and measurably improve MTTR and KPIs.
Additional Details:
  • Work set-up: Hybrid 3x / RTO 2x per week | Eton, Centris
  • Work shift: Nightshift
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer | Hybrid - Centris/Makati
Site Reliability Engineer | Hybrid - Centris/Makati

TASQ • Makati

Hybrid
PHP 900,000 - 1,500,000
(ACTIVE) Site Reliability Engineer | Hybrid Makati city
(ACTIVE) Site Reliability Engineer | Hybrid Makati city

Gratitude Philippines • Manila

On-site
PHP 1,800,000 - 3,000,000
Gratitude Philippines benefits
Site Reability Engineer
Site Reability Engineer

Stafflink Express • Philippines

On-site
PHP 914,000 - 1,095,000
Site Realibility Engineer
Site Realibility Engineer

Gratitude Philippines • Quezon City

Hybrid
PHP 914,000 - 1,095,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Gratitude Philippines • Quezon City

Hybrid
PHP 1,000,000 - 1,800,000
Monitoring & Event Management (SRE)
Monitoring & Event Management (SRE)

Lancesoft APAC • Makati

Hybrid
PHP 1,000,000 - 1,600,000
Site Reliability Engineer
Site Reliability Engineer

Gratitude Philippines • Quezon City

Hybrid
PHP 1,200,000 - 1,800,000
Hybrid work arrangement
Night shift
IT Solution Architect (ServiceNow ITOM) | Hybrid - Centris/Makati
IT Solution Architect (ServiceNow ITOM) | Hybrid - Centris/Makati

TASQ • Makati

Hybrid
PHP 1,800,000 - 2,400,000
Hybrid work model
Night shift
IT Solution Architect (ServiceNow ITOM) | Hybrid - Centris/Makati
IT Solution Architect (ServiceNow ITOM) | Hybrid - Centris/Makati

Tasq Work • Makati

Hybrid
PHP 1,800,000 - 2,400,000
Azure SRE: Reliability Engineer (Hybrid, Night Shift, Manila)
Azure SRE: Reliability Engineer (Hybrid, Night Shift, Manila)

Stafflink Express • Philippines

Hybrid
PHP 914,000 - 1,095,000