Systems Engineer - ELK Observability (Infra Delivery)

Synapxe

Singapore

On-site

SGD 180,000 - 260,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Synapxe is seeking a senior leader to own and evolve the Observability Platform, driving strategy, architecture, and operations across monitoring, logging, metrics, tracing, visualization, and AI/ML capabilities.

You will collaborate with vendors and internal teams to onboard infrastructure and applications, advance anomaly detection, predictive analytics, and self-healing automation, while mentoring engineers in a fast-paced HealthTech environment in Singapore.

Qualifications

  • Degree in Computer Science or Computer Engineering or equivalent.
  • 7+ years in enterprise infrastructure, platform engineering, or IT operations.
  • 2+ years hands-on in enterprise monitoring, observability, and logging platforms.
  • 3+ years experience with server, virtualization, cloud, network, or security infra.
  • 1 year experience in programming, scripting, and automation (Python, PowerShell, Bash, Ansible, Terraform).

Responsibilities

  • Lead and manage the Synapxe Observability Platform strategy, architecture, and operations.
  • Onboard and integrate infrastructure, network, security, and application systems into the observability platform.
  • Drive initiatives in anomaly detection, event correlation, root cause analysis, predictive analytics, self-healing automation, capacity optimization, noise reduction, and outage prediction.
  • Design dashboards, reports, and visualization for infrastructure, security, application, and end-user experience metrics.
  • Establish monitoring standards, alert strategies, service health indicators, and operational KPIs to improve reliability and reduce MTTD/MTTR.
  • Mentor engineers and provide technical guidance to the team.

Skills

Analytical thinking
Troubleshooting
Problem-solving
Pattern recognition
Communication skills
Stakeholder management
Documentation skills
Mentoring
Leadership
Growth mindset

Education

Bachelor's degree in Computer Science/Engineering

Tools

Splunk
Elastic
Dynatrace
Datadog
Grafana
Prometheus
OpenTelemetry
Terraform
Ansible
PowerShell

Job description

Company description:

Synapxe is the national HealthTech agency inspiring tomorrow's health. The nexus of HealthTech, we connect people and systems to power a healthier Singapore.
Together with partners, we create intelligent technological solutions to improve the health of millions of people every day, everywhere. Reimagine the future of health together with us at www.synapxe.sg

Job description:
Position Overview

Lead and manage the Synapxe Observability Platform, overseeing its strategy, architecture, operations, and continuous enhancement across monitoring, logging, metrics, tracing, visualization, and AIOps capabilities. Drive platform reliability, performance, and operational excellence through technical leadership, governance, automation, and adoption of emerging observability technologies, while collaborating with cross-functional teams and mentoring engineers to foster technical excellence and innovation.

Role & Responsibilities
  • Partner with vendors and internal stakeholders to design, develop, and enhance the Synapxe Observability Platform, including monitoring, logging, metrics, distributed tracing, visualization, and AI/ML functionalities.
  • Lead the onboarding and integration of infrastructure, network, security, and application systems into the observability platform.
  • Drive initiatives in anomaly detection, event correlation, root cause analysis, predictive analytics, self-healing automation, capacity optimization, noise reduction, and outage prediction.
  • Design and develop dashboards, reports, and visualization capabilities that provide a unified view of infrastructure, security, application, and end-user experience metrics.
  • Establish monitoring standards, alert strategies, service health indicators, and operational KPIs to improve service reliability and reduce Mean Time to Detect (MTTD) and Mean Time to Resolve (MTTR).
  • Provide technical consultancy and support to stakeholders in report customization, dashboard development, and observability adoption.
  • Lead incident analysis, trend analysis, and continual service improvement initiatives based on observability insights.
  • Manage, mentor, and provide technical guidance to team members and junior engineers. Partner with vendors and internal stakeholders to design, develop, and enhance the Synapxe Observability Platform, including monitoring, logging, metrics, distributed tracing, visualization, and AI/ML functionalities.
  • Lead the onboarding and integration of infrastructure, network, security, and application systems into the observability platform.
  • Drive initiatives in anomaly detection, event correlation, root cause analysis, predictive analytics, self-healing automation, capacity optimization, noise reduction, and outage prediction.
  • Design and develop dashboards, reports, and visualization capabilities that provide a unified view of infrastructure, security, application, and end-user experience metrics.
  • Establish monitoring standards, alert strategies, service health indicators, and operational KPIs to improve service reliability and reduce Mean Time to Detect (MTTD) and Mean Time to Resolve (MTTR).
  • Provide technical consultancy and support to stakeholders in report customization, dashboard development, and observability adoption.
  • Lead incident analysis, trend analysis, and continual service improvement initiatives based on observability insights.
  • Manage, mentor, and provide technical guidance to team members and junior engineers.
Requirements
  • Degree in Computer Science or Computer Engineering or equivalent
  • Minimum 7 years of experience in enterprise-level infrastructure, platform engineering, or IT operations environments.
    • Minimum 2 years of hands-on experience in enterprise monitoring, observability, and logging platforms.
    • Minimum 3 years of experience in server, virtualization, cloud, network, or security infrastructure technologies.
    • Minimum 1 year of experience in programming, scripting, and automation technologies (e.g., Python, PowerShell, Bash, Ansible, Terraform).
  • Strong understanding of observability concepts, including metrics, logs, distributed tracing, alerting, dashboards, and AIOps.
  • Experience with enterprise observability platforms such as Splunk, Elastic, Dynatrace, Datadog, Grafana, Prometheus, OpenTelemetry, or equivalent technologies.
  • Experience with RHEL
  • Strong analytical, troubleshooting, problem-solving, and pattern-recognition skills.
  • Strong communication, stakeholder management, and documentation skills.
  • Proven ability to lead technical teams and mentor engineers.
  • Possess a growth mindset with a strong interest in emerging technologies and industry best practices.
  • 2 years contract
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Systems Engineer - Observability, ELK (Infra Delivery)
Senior Systems Engineer - Observability, ELK (Infra Delivery)

Synapxe • Singapore

On-site
SGD 90,000 - 130,000
Systems Engineer (ELK Observability) - Infra Operations
Systems Engineer (ELK Observability) - Infra Operations

SYNAPXE PTE. LTD. • Singapore

On-site
SGD 90,000 - 150,000
Senior Systems Engineer - Observability & Infra Automation (Central Infra Services)
Senior Systems Engineer - Observability & Infra Automation (Central Infra Services)

Synapxe • Singapore

Hybrid
SGD 80,000 - 110,000
Systems Engineer (ELK Observability) - Infra Operations
Systems Engineer (ELK Observability) - Infra Operations

Synapxe • Singapore

On-site
SGD 120,000 - 180,000
Senior Observability Engineer (ELK Platform & Automation)
Senior Observability Engineer (ELK Platform & Automation)

Synapxe • Singapore

On-site
SGD 90,000 - 130,000
Observability Platform Lead — ELK & AIOps
Observability Platform Lead — ELK & AIOps

Synapxe • Singapore

On-site
SGD 180,000 - 260,000
Systems Engineer - Observability (Infra Deliivery)
Systems Engineer - Observability (Infra Deliivery)

Synapxe • Singapore

On-site
SGD 90,000 - 150,000
Observability Engineer
Observability Engineer

SCIENTE INTERNATIONAL PTE. LTD. • Singapore

On-site
SGD 90,000 - 130,000
Site Reliability Engineer
Site Reliability Engineer

NTT Data Singapore • Singapore

On-site
SGD 110,000 - 190,000
Observability Solutions Architect
Observability Solutions Architect

Cisco • Singapore

On-site
SGD 80,000 - 120,000