Site Reliability Engineer (SRE)

Astek

Santo Niño 1st

On-site

PHP 1,100,000 - 1,900,000

Full time

3 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Astek is seeking an experienced Production Support / Site Reliability Engineer (SRE) to support mission-critical, business-facing applications in cloud and containerised environments.

You will own incident triage, resolve complex issues across Linux/Unix, Kubernetes, and cloud stacks, and automate with Bash, scripts, and IaC tools.

The role requires a CS/IT degree, 5+ years of production or SRE experience, strong collaboration with stakeholders, and familiarity with Agentic AI use cases.

Qualifications

  • Degree in Computer Science, Information Technology, or related discipline.
  • At least 5 years of experience in Production/Application Support or SRE.
  • Hands-on experience supporting business-facing applications and users.
  • Proficiency in Control-M, Unix/Linux, Bash, and Shell scripting.
  • Experience with AWS and/or Azure and cloud-native environments.
  • Hands-on experience with Kubernetes and containerised applications.
  • Familiarity with AutoSys, Datadog, and Terraform.
  • Strong troubleshooting across application and infrastructure layers.
  • Experience with incident and problem management.
  • Strong analytical, communication, and stakeholder management skills.
  • Exposure to Agentic AI technologies and AI-enabled operational use cases.

Responsibilities

  • Provide day-to-day production and application support for mission-critical systems.
  • Investigate and resolve complex application and infrastructure issues across multiple layers.
  • Manage incident triage, incident management, problem management, and root cause analysis.
  • Monitor application health, system performance, and scheduled workloads to maintain reliability.
  • Troubleshoot issues across Linux/Unix, cloud, containerised, and application environments.
  • Develop and maintain Bash/Shell scripts to support operations and automation.
  • Collaborate with business users, engineering teams, and stakeholders to resolve production issues.
  • Identify opportunities to improve system reliability, monitoring, automation, and processes.

Skills

Analytical skills
Communication
Stakeholder management
Incident management
Troubleshooting
Automation mindset

Education

Bachelor's degree in Computer Science or Information Technology

Tools

Control-M
Unix/Linux
Bash
Shell scripting
AWS
Azure
Kubernetes
AutoSys
Datadog
Terraform

Job description

Production Support / Site Reliability Engineer (SRE)
Role Overview

We are looking for an experienced Production Support / Site Reliability Engineer (SRE) with experience in Agentic AI to support mission-critical, business-facing applications. This is a hands‑on, techno-functional role covering production operations, incident resolution, system reliability, and stakeholder support across modern cloud and containerised environments.

Key Responsibilities
  • Provide day‑to‑day production and application support for mission‑critical, business‑facing systems.
  • Investigate and resolve complex application and infrastructure issues across multiple technology layers.
  • Manage incident triage, incident management, problem management, and root cause analysis.
  • Monitor application health, system performance, batch processes, and scheduled workloads to maintain service reliability.
  • Troubleshoot issues across Linux/Unix, cloud, containerised, and application environments.
  • Develop and maintain Bash/Shell scripts to support operational activities and automation.
  • Work closely with business users, engineering teams, and other stakeholders to resolve production issues and minimise service disruption.
  • Identify opportunities to improve system reliability, monitoring, automation, and operational processes.
Requirements
  • Degree in Computer Science, Information Technology, or a related discipline.
  • At least 5 years of experience in Production/Application Support or Site Reliability Engineering (SRE).
  • Strong hands‑on experience supporting business‑facing applications and users.
  • Proficiency in Control‑M, Unix/Linux, Bash, and Shell scripting.
  • Experience with AWS and/or Azure and cloud‑native environments.
  • Hands‑on experience with Kubernetes and containerised applications.
  • Familiarity with operational and infrastructure tools such as AutoSys, Datadog, and Terraform.
  • Strong troubleshooting skills across application and infrastructure layers.
  • Experience with incident and problem management.
  • Strong analytical, communication, and stakeholder management skills.
  • Proactive, collaborative, and adaptable approach to working in fast‑paced environments
  • Exposure to Agentic AI technologies and AI‑enabled operational use cases.
Key Technologies

Control-M | Unix/Linux | Bash/Shell | AWS | Azure | Kubernetes | AutoSys | Datadog | Terraform | Cloud‑Native | Agentic AI

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Blackfort Consulting, Inc.. • Pateros

On-site
PHP 1,200,000 - 2,000,000
Site Reliability Engineer
Site Reliability Engineer

TymblHub • Hinoba-an

On-site
PHP 900,000 - 1,500,000
Agentic AI SRE: Cloud, Kubernetes & Incident Mastery
Agentic AI SRE: Cloud, Kubernetes & Incident Mastery

Astek • Santo Niño 1st

On-site
PHP 1,100,000 - 1,900,000
VS01700 - SRE & Production Reliability Engineer
VS01700 - SRE & Production Reliability Engineer

E4 Software Services Pvt Ltd. • Hinoba-an

On-site
PHP 893,000 - 1,674,000
Site Reliability & Cloud Operations Engineer (SRE)
Site Reliability & Cloud Operations Engineer (SRE)

CloudRaiden • Metro Manila

On-site
PHP 900,000 - 1,700,000
Site Reliability Application Engineers (AI Platform)
Site Reliability Application Engineers (AI Platform)

Astek • Santo Niño 1st

On-site
PHP 600,000 - 1,000,000
DevOps Engineer / Site Reliability Engineer (SRE)
DevOps Engineer / Site Reliability Engineer (SRE)

Concentrix • Mexico

On-site
PHP 1,000,000 - 1,600,000
Site Reliability Engineer
Site Reliability Engineer

Alsons/AWS Information Systems Inc. • Cebu City

On-site
PHP 600,000 - 1,000,000
Associate Site Reliability Engineer
Associate Site Reliability Engineer

Railway Corp • Mexico

On-site
PHP 5,474,000 - 7,908,000
Application Support Engineer
Application Support Engineer

TASQ • Pateros

On-site
PHP 1,200,000 - 2,000,000