Senior Kubernetes Engineer

OpsWerks

Mandaluyong

On-site

PHP 892,800 - 1,116,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A managed services provider in Mandaluyong is seeking a Senior Kubernetes Engineer (EKS) to manage production workloads and support incident response. The role requires 3+ years of hands-on Kubernetes experience, strong knowledge of Amazon EKS, and the ability to mentor junior engineers. Key responsibilities include troubleshooting, enhancing platform reliability, and coordinating with clients. Excellent communication skills and Linux fundamentals are a must. Experience with observability tools and scripting is a plus.

Qualifications

  • 3+ years hands-on Kubernetes experience supporting production environments.
  • Strong experience with Amazon EKS and Kubernetes components.
  • Proven experience in incident response and troubleshooting.

Responsibilities

  • Handle incident response and troubleshooting for Kubernetes/EKS production environments.
  • Act as a senior escalation point for complex cluster issues.
  • Mentor and guide junior engineers.

Skills

Kubernetes support
Incident response
Troubleshooting
CloudWatch
Communication skills
Linux fundamentals
Basic scripting

Tools

Amazon EKS
Prometheus/Grafana
Terraform

Job description

As a Senior Kubernetes Engineer (EKS), you will be part of a managed services team supporting Kubernetes clusters running production workloads. You will handle incident response and troubleshooting, drive operational improvements, and mentor junior engineers—while coordinating closely with internal teams and client stakeholders.

  • Provide hands‑on incident response and troubleshooting for Kubernetes/EKS production environments, including investigation, mitigation, and follow‑through actions.
  • Act as a senior escalation point for complex cluster and application platform issues (networking, DNS, ingress, autoscaling, scheduling, node issues).
  • Support and maintain EKS platform operations such as cluster upgrades, add‑ons management, node group/launch template changes, security patching, and capacity planning—following client change management processes.
  • Improve platform reliability by enhancing observability (metrics/logs/traces), alert quality, and runbook maturity.
  • Identify recurring issues and implement preventative actions through automation, standardization, and documentation.
  • Create and maintain runbooks, troubleshooting guides, operational checklists, and platform standards.
  • Participate in post‑incident reviews (RCA) and ensure corrective and preventive actions are tracked to completion.
  • Mentor and guide junior engineers through reviews, pair troubleshooting, knowledge sharing, and operational best practices.
  • Demonstrate leadership during incidents and projects by coordinating tasks, communicating clearly, and keeping teams aligned on priorities.
  • Participate in an on‑call rotation as part of a 24/7 operations model, with proper handoffs and team support.
Your Qualifications
  • 3+ years hands‑on Kubernetes experience supporting production environments.
  • Strong experience with Amazon EKS and common Kubernetes components (CoreDNS, kube-proxy, CNI, Ingress controllers, autoscaling).
  • Proven experience in incident response, troubleshooting, and production operations (debugging pods, networking/DNS issues, node problems, resource constraints, rollout failures).
  • Working knowledge of Kubernetes fundamentals: deployments, services, ingress, configmaps/secrets, RBAC, namespaces, quotas/limits, PDBs, and upgrade readiness.
  • Familiarity with observability and troubleshooting tools (CloudWatch, Prometheus/Grafana, Splunk/ELK, kubectl debugging, etc.).
  • Basic scripting/automation ability (e.g., Python or Bash) to reduce repetitive operational tasks.
  • Solid Linux and networking fundamentals (TCP/IP basics, DNS, TLS, load balancing concepts).
  • Excellent communication skills (written and oral)—can write clear incident updates, documentation, and explain technical issues to stakeholders.
  • Preferably has leadership skills (formal or informal), with the ability to guide others, lead discussions, and influence improvements.
Plus points if you have:
  • Kubernetes certifications (CKA/CKAD) or AWS certifications
  • Terraform or Crossplane
  • Envoy / Nginx / Proxy concepts (Ingress, service routing, L7 behavior, TLS)
  • Experience with service mesh (Istio/Linkerd) and advanced traffic management
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Kubernetes Engineer - Production EKS & Incident Lead
Senior Kubernetes Engineer - Production EKS & Incident Lead

OpsWerks • Mandaluyong

On-site
PHP 800,000 - 1,000,000
Platform DevOps Engineer — Kubernetes & Cloud
Platform DevOps Engineer — Kubernetes & Cloud

Opswerks • Philippines

On-site
DevOps Engineer (Developer Platform)
DevOps Engineer (Developer Platform)

OpsWerks • Cebu City

On-site
PHP 600,000 - 900,000
DevOps Engineer (Developer Platform)
DevOps Engineer (Developer Platform)

Opswerks • Philippines

On-site
PHP 800,000 - 1,200,000
Google Kubernetes Engine (GKE) Engineer
Google Kubernetes Engine (GKE) Engineer

Bravissimo Resourcing Inc. • Manila

On-site
Senior OpenShift / Kubernetes Engineer
Senior OpenShift / Kubernetes Engineer

Avensys Consulting • Philippines

On-site
PHP 1,800,000 - 3,200,000
Cloud Engineer - AWS - Hybrid - 00587
Cloud Engineer - AWS - Hybrid - 00587

Hunter's Hub Inc. • Taguig

On-site
Site Reliability / Cloud Platform Engineer
Site Reliability / Cloud Platform Engineer

Global Recruitment and Consultancy OPC • Cebu City

On-site
PHP 1,200,000 - 2,400,000
Senior Platform Engineer - Remote Kubernetes & EKS
Senior Platform Engineer - Remote Kubernetes & EKS

SOLANA FOUNDATION • España

On-site
PHP 15,451,000 - 30,903,000
Remote flexibility
Challenging projects
Collaborative culture
Senior Site Reliability Engineer
Senior Site Reliability Engineer

AgileEngine • Mexico

Hybrid
MXN 1,525,000 - 2,034,000
Mentorship
TechTalks
Growth roadmaps
+7