Robotics Reliability & Incident Response Engineer

Serve Robotics

Philippines

On-site

PHP 900,000 - 1,500,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Serve Robotics is seeking a Reliability Operations Engineer to support robotic and cloud systems. You will lead daytime incident investigations, triage escalations, and refine runbooks while collaborating with senior engineers and SREs.

The role requires 5+ years in reliability operations, strong Linux skills, and hands-on experience with GCP and observability tooling. Weekend on-call rotation is part of the role.

Qualifications

  • Bachelor’s degree in Computer Science, IT, Engineering, or equivalent hands-on experience.
  • 5+ years in Reliability Operations, SRE, DevOps, or related technical support.
  • Experience in Tier 1/2 investigations with log review and escalation.
  • Familiarity with cloud environments and incident response workflows.
  • Proficiency with Linux, logs, and basic diagnostics.
  • Ability to follow runbooks and document remediation steps.

Responsibilities

  • Lead incident investigations during regional daytime hours with timely updates.
  • Respond to Tier 1/2 escalations using runbooks, metrics, and diagnostics.
  • Update runbooks and docs for new issues and discoveries.
  • Run automations and improve tooling to streamline remediation tasks.
  • Use Grafana/Prometheus, GCP Monitoring, and OpenTelemetry to interpret metrics and logs.
  • Provide concise updates during incidents and coordinate with engineers.

Skills

Incident response
Linux
Cloud platforms (GCP)
Runbooks
Observability
Jira
On-call rotations
Collaboration with SREs

Education

Bachelor's degree

Tools

Grafana
Prometheus
Google Cloud Monitoring
OpenTelemetry

Job description

Serve Robotics is seeking a Reliability Operations Engineer to support robotic and cloud systems. You will lead daytime incident investigations, triage escalations, and refine runbooks while collaborating with senior engineers and SREs.

The role requires 5+ years in reliability operations, strong Linux skills, and hands-on experience with GCP and observability tooling. Weekend on-call rotation is part of the role.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Robotics Reliability & Incident Response Engineer
Robotics Reliability & Incident Response Engineer

Industrious Ventures • Philippines

On-site
PHP 1,200,000 - 1,800,000
Reliability Operations Engineer (Philippines)
Reliability Operations Engineer (Philippines)

Serve Robotics • Philippines

On-site
PHP 900,000 - 1,500,000
Reliability Operations Engineer (Malaysia)
Reliability Operations Engineer (Malaysia)

Industrious Ventures • Philippines

On-site
PHP 1,200,000 - 1,800,000
Remote Incident & Reliability Lead for AI Ops
Remote Incident & Reliability Lead for AI Ops

Ethos • Manila

On-site
PHP 5,765,000 - 7,799,000
Service Reliability Engineer
Service Reliability Engineer

Metrobank • Philippines

On-site
PHP 480,000 - 720,000
Staff SRE Engineer
Staff SRE Engineer

Stellar Cyber • España

On-site
PHP 5,528,000 - 7,372,000
Site Reliability Engineer
Site Reliability Engineer

NCS Philippines • Taguig

Hybrid
PHP 900,000 - 1,300,000
Lead Site Reliability Engineer (SRE)
Lead Site Reliability Engineer (SRE)

EPAM Systems • Mexico

On-site
PHP 5,846,000 - 8,616,000
Site Reliability Engineer
Site Reliability Engineer

PeoplePlusTech Inc. • Metro Manila

Hybrid
PHP 900,000 - 1,500,000
Senior AI-Driven Cloud Reliability Engineer
Senior AI-Driven Cloud Reliability Engineer

IgniteTech • Philippines

On-site
PHP 1,800,000 - 3,600,000