Senior Production Engineer

863 Parameta Solutions (Singapore) Pte. Limited

Taguig

On-site

PHP 558,000 - 781,200

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

863 Parameta Solutions (Singapore) Pte. Limited seeks a Support Engineer in Taguig. The role involves supporting global trading applications and ensuring stability in a complex environment.

Candidates should have experience with AWS environments, incident management, and observability tooling. The position requires a strong technical background and the ability to communicate effectively with various teams.

Qualifications

  • Solid understanding of AWS environments, including load balancers and managed databases.
  • Experience supporting business-critical front/mid office applications.
  • Strong customer-focus and ability to communicate complex technical issues.

Responsibilities

  • Own and lead the investigation of complex production issues.
  • Champion improvements to observability and service health visibility.
  • Lead technical response during major incidents and champion post-incident reviews.

Skills

AWS operational environments
Incident management
Root-cause analysis
Observability tooling (Grafana, CloudWatch, etc.)
Customer focus

Education

Degree level or equivalent education

Tools

CloudWatch
Grafana
Splunk

Job description

Role Overview

The Support Engineer will be part of a small team responsible for supporting global trading applications. The role requires deep technical and functional expertise to resolve escalated incidents, support operational workflows across multiple regions, and maintain the stability, performance and reliability of a complex global trading platform. Responsibilities include monitoring system health, analysing incidents, coordinating changes and releases, troubleshooting electronic trading workflows, and collaborating with technical and business teams. The role is both hands‑on and operational, requiring a structured mindset, situational awareness and the ability to operate confidently in a fast‑moving, globally distributed environment. The position will increasingly contribute to a Site Reliability Engineering (SRE)-aligned support model, focusing on automation, observability, reliability metrics and reduction of operational toil.

Responsibilities
  • Own and lead the investigation of complex production issues, ensuring clear root‑cause analysis and permanent remediation.
  • Own production health for assigned services, overseeing monitoring, alerts and runtime behaviour.
  • Own operational readiness for assigned platforms and trading‑day stability.
  • Champion improvements to observability, alerting quality and service health visibility.
  • Own production impact assessment for changes and releases, ensuring risk is understood and mitigated.
  • Lead technical response during major incidents, acting as the production authority for diagnosis and recovery.
  • Champion post‑incident reviews, ensuring preventative actions are defined and delivered.
  • Own and drive systemic reliability improvements based on incident trends and failure patterns.
  • Lead cross‑team collaboration, influencing engineering decisions from a production perspective.
  • Champion automation and tooling enhancements to reduce toil and improve mean time to detect and mean time to restore.
Essential Qualifications
  • Educated to degree level or equivalent combination of education and experience.
  • Solid understanding of AWS operational environments, including load balancers, regional failover behaviour, instance lifecycles, and managed databases.
  • Experience supporting business‑critical front/mid office applications.
  • Deep knowledge of market data flows, instrument definitions, pricing mechanisms and session‑based connectivity.
  • Ability to interpret complex application logs and diagnose backend issues with accuracy and speed.
  • Strong root‑cause analysis capability with the ability to evaluate symptoms, isolate faults and determine remediation paths.
  • Solid experience in incident management, major‑incident coordination and structured problem‑solving.
  • Demonstrated ability to work across regions, managing concurrent issues, escalations and stakeholder communications.
  • Clear understanding of change‑management disciplines including risk assessment and deployment validation.
  • Familiarity with observability tooling such as Grafana, CloudWatch, ELK and Splunk, including metrics, logs, dashboards and alerting.
  • Proven ability to work across multidisciplinary teams (Business, Operations, Developers and DevOps).
  • Strong customer‑focus and ability to communicate complex technical issues in a business‑friendly manner.
  • Comfortable supporting global operations and adapting to multi‑region workflows.
  • Demonstrated interest or experience in applying SRE principles such as reliability metrics, automation and continuous improvement within a support or operations role.
  • Experience contributing to improved mean time to detect (MTTD) and mean time to restore (MTTR) through better observability, tooling or process.
  • Understanding of the balance between feature delivery and operational stability in business‑critical systems.
Desired Qualifications
  • Solid experience supporting trading platforms, financial exchanges or real‑time transactional systems.
  • Reasonable exposure to FIX‑based workflows, messaging pipelines or market‑connectivity architectures.
  • Experience working within AWS‑native or hybrid‑cloud financial environments.
  • Reasonable experience collaborating with front‑office trading desks or broker support teams.
  • Familiarity with CI/CD pipelines, DevOps practices or automated deployment frameworks.
  • Exposure to SRE concepts such as SLIs, SLOs, error budgets or reliability metrics within a support or operations context.
  • Experience improving platform resilience through automation, monitoring enhancements or operational tooling.
  • Familiarity with failure scenarios, recovery patterns or high‑availability strategies in distributed systems.
Location

Philippines – Ecoprime Building, Taguig City

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Production Engineer
Production Engineer

TP ICAP • Manila

On-site
Production Engineer
Production Engineer

768 TP ICAP Management Services Ltd (Philippines Branch) • Taguig

On-site
PHP 446,400 - 669,600
Team Lead - Production Engineering
Team Lead - Production Engineering

768 TP ICAP Management Services Ltd (Philippines Branch) • Taguig

On-site
PHP 1,200,000 - 1,500,000
Production Engineer
Production Engineer

ECLARO • Taguig

On-site
PHP 900,000 - 1,200,000
Sr. Production Engineer (Trading Applications) – Hybrid | Taguig City, Philippines
Sr. Production Engineer (Trading Applications) – Hybrid | Taguig City, Philippines

Charterhouse Pte Ltd • Taguig

On-site
PHP 600,000 - 800,000
Production Engineer, SRE‑Driven Trading Platform
Production Engineer, SRE‑Driven Trading Platform

768 TP ICAP Management Services Ltd (Philippines Branch) • Taguig

On-site
PHP 446,400 - 669,600
Senior Production Engineer
Senior Production Engineer

TP ICAP • Manila

On-site
PHP 900,000 - 1,800,000
Senior Production Engineer
Senior Production Engineer

TP ICAP Group • Taguig

On-site
PHP 1,200,000 - 1,500,000
Inclusive work environment
Opportunities for growth
Employee networking programs
Team Lead - Production Engineering
Team Lead - Production Engineering

TP ICAP • Manila

On-site
PHP 1,800,000 - 2,400,000
Global Production Engineer - SRE & Incident Response
Global Production Engineer - SRE & Incident Response

TP ICAP • Manila

On-site
PHP 900,000 - 1,800,000