Infrastructure & Platform Operations Engineer - Philippines

DysrupIT Pty

Hinoba-an

Hybrid

PHP 900,000 - 1,200,000

Full time

12 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Hybrid/Remote work arrangement
HMO coverage
Career growth opportunities
Competitive compensation

Job summary

DysrupIT Pty in the Philippines is seeking an Infrastructure & Platform Operations Engineer to maintain and improve our production platforms across cloud and on‑premise environments. You will focus on BAU support, service stability and knowledge transfer, with opportunities to drive automation and platform upgrades.

The role requires strong hands‑on Linux, AWS, PostgreSQL and MongoDB experience, plus Docker/Kubernetes expertise and scripting skills.

Qualifications

  • Typically 5+ years of experience in infrastructure engineering, platform operations, systems engineering or production support.
  • Strong BAU and production support experience with ownership of live incidents and remediation.
  • Strong hands‑on AWS experience in production environments.
  • Very strong Linux administration and troubleshooting, especially Ubuntu and Red Hat.
  • Deep networking knowledge including TCP/IP, DNS, routing, firewalls, VPNs, TLS and cloud networking.
  • Strong MongoDB production administration and troubleshooting.
  • Operational PostgreSQL experience including backup, recovery and monitoring.
  • Hands‑on Docker and Kubernetes with container lifecycle, networking and logging.
  • Python and/or PowerShell scripting for automation; Bash also valuable.
  • Working knowledge of Git and GitHub for version control.

Responsibilities

  • Work as part of Service Operations to support enterprise infrastructure and production platforms.
  • Administer, support and troubleshoot customer environments including core infra and databases.
  • Support Linux environments (Ubuntu, Red Hat) and Windows where required.
  • Diagnose complex network connectivity issues across TCP/IP, DNS, routing, VPNs, TLS and cloud networking.
  • Monitor and administer PostgreSQL and MongoDB production environments.
  • Manage containerised environments using Docker and Kubernetes and related tooling.
  • Support cloud services in AWS with EKS/ECS and related platform health activities.
  • Use scripting to automate repeatable tasks and improve operational efficiency.
  • Collaborate with DevOps, Data Services and Product teams to deliver end‑to‑end outcomes.

Skills

Linux admin
Networking
AWS
Docker
Kubernetes
SQL PostgreSQL
MongoDB
Python/PowerShell
Git/GitHub
Incident management

Tools

AWS
Kubernetes
Docker
GitHub
Terraform
Jira/JSM

Job description

ABOUT DYSRUPIT

DysrupIT is a consulting‑led technology firm. We help mid‑market to enterprise businesses solve business problems through technology — whether that’s consulting, execution, managed services, or staff augmentation — and we take accountability for the outcome. We are dedicated to making a positive impact in the communities we serve.

COMPANY CULTURE

At DysrupIT, success isn’t measured in headcount placed or hours billed — it’s measured in outcomes delivered. We’re a team that takes ownership of the work, stays curious about problems beyond our immediate scope, and builds relationships meant to grow, not just renew or end. We invest in our people with the training and support they need to grow their careers, and we back a culture where everyone, regardless of role, is encouraged to notice opportunities, ask one more question, and help lead the story for our clients, not just deliver it.

JOB SUMMARY

The Infrastructure & Platform Operations Engineer supports, maintains and improves the infrastructure and production platforms used to deliver Nephos services to customers. The jobholder works across cloud / on‑premise infrastructure, Linux, networking, databases, containers, monitoring and enterprise platforms, with a strong initial focus on BAU production support, service stability and knowledge transfer. As capability develops, the role will also contribute to platform change, upgrades, automation and project delivery.

JOB RESPONSIBILITIES

Working hours: 9–5 UK hours during onboarding and transition, with flexibility for occasional out‑of‑hours upgrades, changes and major incidents.

  • Work as part of Service Operations to support enterprise infrastructure and production platforms, including health monitoring, issue resolution, proactive maintenance, resilience and continuous improvement.
  • Administer, support and troubleshoot customer environments, including core infrastructure, storage, networking, IAM awareness, compute resources and managed services relevant to supported platforms.
  • Support Linux environments, primarily Ubuntu and Red Hat, using command‑line administration and troubleshooting techniques; provide appropriate support for Windows where required.
  • Diagnose complex network and connectivity issues across TCP/IP, DNS, routing, firewalls, VPNs, TLS, proxies, load balancers and cloud networking.
  • Administer and support PostgreSQL production environments, including configuration, roles and permissions, monitoring, backup and restore, recovery, performance investigation and upgrades.
  • Administer and support MongoDB production environments, including health, logs, performance, backup and recovery, upgrades and troubleshooting.
  • Support containerised and orchestrated environments using Docker and Kubernetes, including container lifecycle, logs, configuration, volumes, networking, registries, health checks, resource constraints and failure diagnosis.
  • Support cloud container services and clusters, including AWS with EKS and ECS where applicable, and contribute to platform configuration, health and capacity management.
  • Monitor infrastructure and applications using enterprise monitoring and observability tools, including alert investigation, log analysis, threshold review, root‑cause identification and reduction of unnecessary alert noise.
  • Use scripting and automation, particularly Python and PowerShell, to improve repeatability, reduce manual effort and support operational tasks; Bash or other transferable scripting experience is also valuable.
  • Use Git and GitHub to manage scripts, configuration and operational content, following appropriate version‑control practices.
  • Support enterprise platform installations, configuration, upgrades, testing, validation and recovery, including data platforms such as BigID and internally developed solutions such as Nephos‑developed platforms where customer access is approved.
  • Learn and develop operational capability in new or unfamiliar enterprise platforms quickly, using structured knowledge transfer, documentation, practical shadowing and independent task completion.
  • Take ownership of Incident and Problem records, applying structured troubleshooting to identify issues, restore service as quickly and safely as possible, complete root‑cause analysis and maintain accurate records.
  • Take ownership of Service Requests and technical operational tasks, ensuring work is completed accurately, securely and within agreed priorities.
  • Manage Change records relating to infrastructure, databases, platforms, upgrades and improvement activity, attending CAB where required and ensuring implementation, validation and rollback plans are appropriate.
  • Create and maintain clear knowledge articles, runbooks and work instructions, and actively participate in cross‑training so that critical operational knowledge is not concentrated with one individual.
  • Work directly with customer and internal technical teams during incidents, changes, upgrades and investigations, explaining findings clearly and maintaining a professional, customer‑focused approach.
  • Collaborate with DevOps, Data Services, Product, Delivery and other cross‑functional teams to resolve issues and deliver end‑to‑end platform outcomes.
  • Contribute technical input to platform improvements, resilience, capacity, monitoring and operational design. Final architecture and governance decisions remain shared internal responsibilities.
  • Support backup, restore and disaster‑recovery activities, including practical validation and recovery exercises.
  • Work in accordance with security principles including least privilege, secure credential handling, secrets, certificates, auditability, patching and vulnerability awareness.
  • Access to individual customer environments remains subject to the relevant customer contractual, security and access approval processes.
  • The role is focused on operating, supporting, deploying and improving production platforms rather than developing application functionality.
QUALIFICATIONS
  • Typically 5+ years of experience in infrastructure engineering, platform operations, systems engineering or production support; demonstrable capability is more important than a fixed number of years.
  • Strong, practical BAU and production support experience, including ownership of live incidents, service restoration, monitoring, failed changes or deployments, root‑cause analysis and controlled remediation.
  • Strong hands‑on AWS experience in production environments. AWS is the primary cloud requirement for this role.
  • Very strong demonstrable Linux administration and troubleshooting experience, particularly Ubuntu and Red Hat, including confident command‑line use.
  • Deep practical networking and connectivity troubleshooting experience covering TCP/IP, DNS, routing, firewalls, VPNs, TLS and cloud networking.
  • Strong production MongoDB administration and troubleshooting experience. This is a critical capability for the role.
  • Meaningful PostgreSQL operational experience, including administration, backup and recovery, monitoring and troubleshooting; deeper production DBA experience is highly valued.
  • Hands‑on production experience with Docker and a strong working knowledge of container lifecycle, networking, logging, health and troubleshooting.
  • Strong hands‑on Kubernetes experience, including deployment health, pods, services, configuration, logs and failure diagnosis.
  • Experience supporting enterprise applications and platforms through installation, configuration, version upgrades, health monitoring and technical troubleshooting.
  • Strong scripting and automation capability, particularly Python and/or PowerShell; experience with Bash or other transferable scripting languages is also relevant.
  • Strong monitoring and observability fundamentals, including infrastructure metrics, log investigation, alert analysis and distinguishing symptoms from root cause.
  • Excellent issue analysis, troubleshooting and fault‑finding skills, with a methodical approach to complex production problems.
  • Working knowledge of Git and GitHub, including repositories, version control and collaborative working practices.
  • Practical understanding of backup, restore, recovery and service‑resilience principles.
  • Practical IT service management experience covering Incident, Problem, Change, Service Request and major‑incident processes. Formal ITIL certification is not mandatory.
  • Experience using a service‑management platform such as Jira/JSM, ServiceNow, Remedy, Freshservice or equivalent.
  • Good security awareness, including least privilege, IAM concepts, secrets, certificates, MFA, privileged access, patching, vulnerability awareness and secure credential handling.
  • Strong customer‑centric attitude with the ability to communicate clearly with technical and non‑technical stakeholders during normal operations and incidents.
  • Excellent written and verbal communication skills, including the ability to create clear technical documentation and operational procedures.
  • Demonstrable ability to learn complex unfamiliar technology and become independently effective following structured onboarding and knowledge transfer.
  • Able to work independently following onboarding, manage multiple priorities and recognise when escalation or wider technical input is required.
  • Flexible, organised and self‑motivated, with willingness to participate in occasional out‑of‑hours upgrades, changes and major incidents where required.
NICE TO HAVE
  • Hands‑on Azure experience. Azure is expected to become increasingly relevant, although AWS remains the primary current requirement.
  • Experience with Google Cloud Platform and services such as Cloud Run or Cloud SQL.
  • Experience with infrastructure‑as‑code principles and tooling such as Terraform, CloudFormation, Ansible or equivalent.
  • Familiarity with CI/CD platforms and deployment pipelines such as GitHub Actions, Jenkins, GitLab CI or Azure DevOps.
  • Previous experience with BigID or another complex enterprise data‑governance platform. BigID experience is not required, but the ability to learn it is important.
  • Experience with RabbitMQ, Redis or similar messaging/cache technologies.
  • Experience with LogicMonitor or comparable enterprise monitoring platforms.
  • Experience using privileged‑access and secrets‑management technologies such as Delinea and HashiCorp tooling.
  • Experience with REST APIs and API troubleshooting.
  • Experience in applying, renewing and troubleshooting TLS certificates.
  • Experience with disaster‑recovery planning, RPO/RTO considerations and practical recovery testing.
  • Knowledge of data governance, information security or regulated customer environments.
  • Relevant AWS, Azure, Kubernetes, PostgreSQL, MongoDB or ITIL certifications. Certifications are advantageous but not mandatory.
  • Experience supporting microservices, serverless or distributed platform architectures.
WHAT WE OFFER
  • Competitive compensation package commensurate with experience.
  • Government‑mandated benefits plus supplemental HMO coverage.
  • Collaborative and professional work environment.
  • Career growth opportunities within a growing technology organization.
  • Hybrid/Remote work arrangement (where applicable).
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Infrastructure & Platform Operations Engineer
Infrastructure & Platform Operations Engineer

DysrupIT • Philippines

On-site
PHP 1,200,000 - 2,000,000
DevOps Engineer
DevOps Engineer

DysrupIT • Philippines

On-site
PHP 893,000 - 1,562,000
Competitive compensation
HMO coverage
Collaborative work environment
+2
Platform & Infrastructure Ops Engineer (Hybrid/Remote)
Platform & Infrastructure Ops Engineer (Hybrid/Remote)

DysrupIT Pty • Hinoba-an

Hybrid
PHP 900,000 - 1,200,000
Hybrid/Remote work arrangement
HMO coverage
Career growth opportunities
+1
Cloud Infrastructure Operations and DevOps Platform Lead
Cloud Infrastructure Operations and DevOps Platform Lead

RecruitGo Careers • Philippines

On-site
PHP 2,000,000 - 3,500,000
Office in BGC
US shift coverage
Competitive compensation
DevOps Engineer | Python | BASH | Terraform | Hybrid | RTO 1x In Month
DevOps Engineer | Python | BASH | Terraform | Hybrid | RTO 1x In Month

AVENSYS CONSULTING INC. • Philippines

Hybrid
PHP 1,500,000 - 2,100,000
Cloud & Infrastructure Specialist
Cloud & Infrastructure Specialist

Valsoft Corporation • Philippines

Remote
PHP 1,000,000 - 1,500,000
Senior IT Engineer
Senior IT Engineer

OpsWerks • Mandaluyong

On-site
Cloud Infrastructure Operations And DevOps Platform Lead
Cloud Infrastructure Operations And DevOps Platform Lead

RecruitGo • Philippines

On-site
PHP 1,000,000 - 2,000,000
Office-based role in BGC, Taguig
Senior IT Infrastructure Engineer
Senior IT Infrastructure Engineer

Straive • Philippines

Hybrid
PHP 1,200,000 - 2,400,000
Senior Platform Engineer
Senior Platform Engineer

Our Clients • Pasay

On-site
PHP 1,200,000 - 1,800,000