Site Reliability Engineer - Hybrid | Automation & Uptime

PeoplePlusTech Inc.

Metro Manila

Hybrid

PHP 700,000 - 1,100,000

Full time

3 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

PeoplePlusTech Inc. is seeking a Site Reliability Engineer (SRE) to join our growing technology team.

The ideal candidate will maintain the reliability, availability, scalability, and performance of critical applications and infrastructure, blending software engineering with systems administration to build automated, highly available platforms. The successful candidate will collaborate with development, infrastructure, security, and operations teams to drive automation, improve reliability, and

Qualifications

  • Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field.
  • Minimum of 3 years of experience in Site Reliability Engineering, DevOps, Cloud Engineering, Systems Engineering, or Infrastructure Operations.
  • Experience supporting production environments with high availability and uptime requirements.
  • Strong understanding of Linux and/or Windows server administration.
  • Experience with cloud platforms such as AWS, Microsoft Azure, or Google Cloud Platform (GCP).
  • Experience with containerization technologies such as Docker and orchestration platforms such as Kubernetes.
  • Hands-on experience with CI/CD tools such as Azure DevOps, GitHub Actions, Jenkins, GitLab CI/CD, or similar.
  • Experience with infrastructure-as-code tools such as Terraform, CloudFormation, or Ansible.
  • Familiarity with monitoring and observability platforms such as Prometheus, Grafana, Datadog, New Relic, Dynatrace, Splunk, or ELK Stack.
  • Strong scripting and automation skills using Python, Bash, PowerShell, or similar languages.
  • Experience in troubleshooting complex production incidents and conducting root cause analysis.
  • Knowledge of networking concepts, DNS, load balancing, firewalls, and security best practices.

Responsibilities

  • Design, implement, and maintain highly available and scalable infrastructure and application environments.
  • Monitor system health, application performance, and service availability using industry-standard monitoring and observability tools.
  • Develop and maintain automation scripts, tools, and workflows to improve operational efficiency.
  • Manage incident response, troubleshooting, root cause analysis (RCA), and post-incident reviews.
  • Collaborate with development teams to improve application reliability, deployment processes, and operational readiness.
  • Establish and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Service Level Agreements (SLAs).
  • Implement infrastructure-as-code (IaC) solutions for consistent and repeatable deployments.
  • Support CI/CD pipelines and deployment automation initiatives.
  • Perform capacity planning, performance tuning, and proactive system optimization.
  • Ensure compliance with security, governance, and operational best practices.
  • Create and maintain technical documentation, operational runbooks, and disaster recovery procedures.
  • Participate in on-call support and incident management activities as required.

Skills

Linux
Windows Server
AWS
Azure
GCP
Docker
Kubernetes
CI/CD
Terraform
Ansible
Python
Bash
PowerShell
Networking fundamentals
Prometheus
Grafana
Security best practices
Incident management
Root cause analysis

Education

Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field

Job description

PeoplePlusTech Inc. is seeking a Site Reliability Engineer (SRE) to join our growing technology team.

The ideal candidate will maintain the reliability, availability, scalability, and performance of critical applications and infrastructure, blending software engineering with systems administration to build automated, highly available platforms. The successful candidate will collaborate with development, infrastructure, security, and operations teams to drive automation, improve reliability, and

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

PeoplePlusTech Inc. • Metro Manila

Hybrid
PHP 700,000 - 1,100,000
Site Reliability Engineer - Remote/Hybrid, High Impact
Site Reliability Engineer - Remote/Hybrid, High Impact

Manatal • Philippines

Hybrid
PHP 800,000 - 1,400,000
Healthcare coverage on day one
Dependents coverage
Paid time-off with cash conversion
+2
Platform SRE: Scale, Automate & Reliability (Hybrid)
Platform SRE: Scale, Automate & Reliability (Hybrid)

Broadridge Financial Solutions • Manila, Hinoba-an

Hybrid
PHP 1,200,000 - 1,800,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Acquire Intelligence • Taguig

On-site
PHP 900,000 - 1,500,000
Senior Site Reliability Engineer - Scalable Cloud (Hybrid)
Senior Site Reliability Engineer - Scalable Cloud (Hybrid)

V2 Solutions • Hinoba-an

Hybrid
PHP 995,000 - 1,525,000
Senior Site Reliability Engineer - Cloud & Automation
Senior Site Reliability Engineer - Cloud & Automation

OpsWerks • Mandaluyong

On-site
PHP <100,000
Site Reliability Engineer – Scale & Automation
Site Reliability Engineer – Scale & Automation

Google Inc. • Hinoba-an

On-site
PHP 600,000 - 1,200,000
Site Reliability Engineer - Incident Response & Automation
Site Reliability Engineer - Incident Response & Automation

E-IT • Philippines

Remote
PHP 800,000 - 1,200,000
Site Reliability Engineer
Site Reliability Engineer

E-IT • Philippines

Remote
PHP 800,000 - 1,200,000
Hybrid Cloud SRE: Observability & Automation
Hybrid Cloud SRE: Observability & Automation

NICE • Manila

Hybrid
PHP 669,600 - 892,800