Site Reability Engineer

Stafflink Express

Philippines

Hybrid

PHP 914,000 - 1,095,000

Full time

5 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Stafflink Express is seeking a Site Reliability Engineer to design and implement robust monitoring across platforms with emphasis on Azure Monitor, Grafana, and ServiceNow ITOM. The role collaborates with product teams to define SLO/SLI driven operations, reduce alert fatigue, and accelerate incident response.

Night shift with hybrid work setup in Manila, offering hands-on exposure to telemetry and AIOps, and opportunities to mature enterprise patterns and governance artifacts.

Qualifications

  • Bachelor's degree in IT/CS/Engineering or related field.
  • 3+ years in monitoring/observability/SRE roles with hands‑on Azure Monitor/App Insights (KQL) and ServiceNow Event Management.
  • Strong knowledge in Azure Log Analytics, KQL, Telemetry, and APM implementations.
  • Collaborates across IT Operations, platform, cyber, network, and product teams; strong written and verbal communication.
  • 5+ years in SRE role with deep understanding of monitoring and application performance management.
  • Knowledge of Nobl9 SLO platforms and governance artifacts.
  • Knowledge of proactive monitoring using Azure Monitor services, telemetry, and synthetic transactions.
  • Understanding of network architecture and security: WAN/LAN, TCP/IP, PKI.
  • Familiarity with ITSM tools (ServiceNow) and compliance processes.
  • Have AIOps vision and awareness.
  • Not a job hopper.

Responsibilities

  • Design and define standards, patterns, and automation opportunities elevating monitoring and reliability with Azure Monitor, ServiceNow ITOM Event Management, Grafana, and APM/Synthetics tooling.
  • Partner with product teams to implement SLO/SLI-driven operations, reduce alert noise, accelerate incident response, and embed self-healing where it matters most.
  • Engineer enterprise monitoring & event patterns by authoring and maintaining reference architectures, runbooks, and event management models.
  • Contribute to Monitoring and Observability strategy and governance checkpoints and coach product teams.
  • Excellent communication skills to drive continuous improvement by reducing alert noise, shorten MTTR, and embed postmortem learnings into patterns and pipelines.

Skills

Azure Monitor / App Insights (KQL)
ServiceNow Event Management
Azure Log Analytics
Telemetry
APM implementations
Nobl9 / SLO platforms
ITSM processes (ServiceNow)
AIOps vision
Networking basics (WAN/LAN/TCP/IP)
Proactive monitoring & telemetry

Education

Bachelor’s degree in IT /Computer science/Engineering, or related field

Tools

Grafana
Prometheus
App Dynamics
ThousandEyes
ServiceNow

Job description

Job title: Site Realibility Engineer

Work set up: Hybrid: 2 WFH & 3 RTO (Location: Manila (Eton Centris, Quezon Avenue, Quezon City))

Work shift: Night Shift

Salary: P90,000

Start date: ASAP

Qualifications:
  • Bachelor’s degree in IT /Computer science/Engineering, or related field.
  • 3+ years in monitoring/observability/SRE roles with hands‑on experience in Azure Monitor/App Insights (KQL) and ServiceNow Event Management.
  • Strong knowledge in Azure Log Analytics, KQL, Telemetry, APM implementations
  • Demonstrated ability to collaborate across IT Operations team, platform, cyber, network, and product teams, strong written verbal communication for standards and enablement.
  • 5+ years of experience with SRE role and deep understanding of monitoring and application performance management
  • Knowledge of SLO platforms (e.g., Nobl9) and experience contributing to standards/governance artifacts.
  • Knowledge of proactive monitoring using Azure monitor services, telemetry, and synthetic transactions.
  • Understanding of network architecture and security: WAN/LAN, TCP/IP, PKI.
  • Familiarity with ITSM processes and tools (e.g., ServiceNow), and compliance processes
  • Have AIOps vision and awareness
  • Not a job hopper
Responsibilities:
  • You will design and define standards, patterns, and automations opportunities that elevate monitoring and reliability across platforms and applications, with a strong focus on Azure Monitor, ServiceNow ITOM Event Management, Grafana, and APM/Synthetics tooling
  • You’ll partner with product teams to implement SLO/SLI‑driven operations, reduce alert noise, accelerate incident response, and embed self‑healing where it matters most.
  • Engineer enterprise monitoring & event patterns by authoring and maintaining reference architectures, runbooks, and event management models (alert → event → incident) with actionable alerts and incidents routing.
  • Contribute to Monitoring and Observability & Event Management Strategy and tooling intake/governance checkpoints and coach product teams
  • Excellent communication skills to drive continuous improvement by reducing alert noise, shorten MTTR, and improve change success by embedding postmortem learnings into patterns, rules, and pipelines.
Must have Skills:
  • Cloud Observability: Azure Monitor/App Insights/Log Analytics (KQL)
  • Knowledge of Grafana, Prometheus, App Dynamics, ThousandEyes
  • Communication & Teaming – Able to translate complex reliability patterns into consumable standards and coach IT operations team via office hours/CoP sessions.
  • Technical Depth in Monitoring and Observability Stack – Hands‑on in ServiceNow Event Management, Azure Monitor/KQL, and automation.
  • Analytical & Systems Thinking – Uses SLI/SLOs, postmortems, and CMDB context to reduce noise, drive self‑healing, and measurably improve MTTR and KPIs.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Azure SRE: Reliability Engineer (Hybrid, Night Shift, Manila)
Azure SRE: Reliability Engineer (Hybrid, Night Shift, Manila)

Stafflink Express • Philippines

Hybrid
PHP 914,000 - 1,095,000
Solutions Architect (ServiceNow/Azure)
Solutions Architect (ServiceNow/Azure)

Lewis Personnel Management • Makati

Hybrid
PHP 1,200,000 - 2,000,000
Site Reliability Engineer
Site Reliability Engineer

MicroSourcing • Manila

On-site
PHP 900,000 - 1,500,000
Healthcare coverage
Paid time-off with cash conversion
Group life insurance
+3
Site Reliability Engineer
Site Reliability Engineer

IDEMIA PHILIPPINES INC. • Philippines

On-site
PHP 900,000 - 1,350,000
IT Solutions Architect – ServiceNow ITOM
IT Solutions Architect – ServiceNow ITOM

Avensys Consulting • Makati

On-site
PHP 1,200,000 - 2,100,000
Site Reliability Engineer - BGC - Hybrid - Up to 180K
Site Reliability Engineer - BGC - Hybrid - Up to 180K

weSource Management Consultancy Firm • Taguig

Hybrid
PHP 150,000 - 180,000
SRE: Cloud, Observability & Automation
SRE: Cloud, Observability & Automation

Trinity Workforce Solutions, Inc. • Makati

On-site
Site Reliability Engineer - Hybrid BGC - Up to 155K
Site Reliability Engineer - Hybrid BGC - Up to 155K

weSource Management Consultancy Firm • Taguig

Hybrid
Site Reliability Engineers
Site Reliability Engineers

Trinity Workforce Solutions, Inc. • Makati

On-site
Site Reliability Engineer
Site Reliability Engineer

AgileEngine • Mexico

Hybrid
PHP 8,734,000 - 13,100,000
Professional growth
Competitive USD-based pay
Exciting projects
+1