Production Engineer

ECLARO

Taguig

On-site

PHP 900,000 - 1,200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

ECLARO is seeking a Production Engineer to join their team in Metro Manila, Taguig. This role involves supporting global trading applications, ensuring operational reliability and performance. The successful candidate will handle escalated issues, monitor system performance, and drive continuous improvements in the support function.

The ideal applicant should possess a solid understanding of AWS and experience with business-critical applications. Collaboration across technical and operational teams is essential for this position, which is aligned with evolving Site Reliability Engineering practices.

Qualifications

  • Solid understanding of AWS operational environments.
  • Experience supporting business-critical front/mid-office applications.
  • Ability to interpret complex application logs.

Responsibilities

  • Investigate and resolve escalated production issues.
  • Monitor the health and performance of global trading applications.
  • Support critical operational processes for trading.
  • Participate in change and release activities.
  • Analyze recurring incidents to identify systemic issues.

Education

Degree level education or equivalent experience

Tools

AWS operational environments
Grafana
CloudWatch
ELK
Splunk

Job description

Production Engineer

The individual will form part of a small team responsible for the support of global trading applications. The Support Engineer plays a critical role in maintaining the stability, performance, and operational reliability of a complex global trading platform. The role requires deep technical and functional expertise to resolve escalated issues, supports global operational workflows across multiple regions, and ensures the platform runs consistently throughout daily trading cycles.

This position involves monitoring system health, analyzing incidents, coordinating changes and releases. It requires strong troubleshooting skills, an understanding of electronic trading workflows, and the ability to collaborate effectively with a wide range of technical and business teams.

The Support Engineer ensures seamless daily operations, rapid incident response, accurate stakeholder communication, and proactive identification of risks or emerging issues. The role is both technically hands‑on and operationally engaged, requiring a structured mindset, situational awareness, and the ability to operate confidently in a fast‑moving, globally distributed environment.

As the platform and operating model continue to mature, this role will increasingly contribute to a Site Reliability Engineering (SRE)‑aligned support model, with greater emphasis on automation, observability, reliability metrics, and systemic reduction of operational toil. The successful candidate will help shape this evolution while continuing to ensure the stability of critical trading services. This role provides a strong pathway toward future SRE responsibilities as the platform and operating model continue to evolve.

Role Responsibilities
  • Investigate and resolve escalated production issues using structured troubleshooting, log analysis, and system behavior diagnostics.
  • Monitor the health and performance of global trading applications, responding to alerts, incidents, and operational risks across regions.
  • Support critical operational processes including Start of Day, End of Day, and environment readiness for global trading cycles.
  • Maintain and continuously improve monitoring, dashboards, and alerting to enhance visibility, early detection, and operational decision‑making.
  • Participate in change and release activities, including deployment support, impact assessment, validation, and post‑change review.
  • Coordinate and support major incidents, ensuring clear triage, effective communication, and controlled recovery.
  • Analyze recurring incidents and operational patterns to identify systemic issues and contribute to long‑term reliability improvements.
  • Contribute to the evolution of the support function toward an SRE‑aligned operating model, with increased focus on resilience, automation, and production readiness.
  • Identify opportunities to reduce manual operational effort through automation, tooling improvements, or process refinement.
  • Participate in post‑incident reviews, contributing to blameless root cause analysis and tracking preventative actions with engineering and platform teams.
Experience / Competences
Essential
  • Educated to degree level or equivalent combination of education and experience.
  • Solid understanding of AWS operational environments, including load balancers, regional failover behavior, instance lifecycles, and managed databases.
  • Experience supporting business‑critical front/mid‑office applications.
  • Deep knowledge of market data flows, instrument definitions, pricing mechanisms, and session‑based connectivity.
  • Ability to interpret complex application logs and diagnose backend issues with accuracy and speed.
Operational & Analytical Essentials
  • Strong root‑cause analysis capability with the ability to evaluate symptoms, isolate faults, and determine remediation paths.
  • Solid experience in incident management, major‑incident coordination, and structured problem‑solving.
  • Demonstrated ability to work across regions, managing concurrent issues, escalations, and stakeholder communications.
  • Clear understanding of change‑management disciplines including risk assessment and deployment validation.
  • Familiarity with observability tooling (Grafana, CloudWatch, ELK, Splunk), including metrics, logs, dashboards, and alerting used to assess system health and reliability.
Collaboration Essentials
  • Proven ability to work across multidisciplinary teams (Business, Operations, Developers, DevOps).
  • Strong customer‑focus and ability to communicate complex technical issues in a business‑friendly manner.
  • Comfortable supporting global operations and adapting to multi‑region workflows.
Reliability & Platform Engineering (Evolving)
  • Demonstrated interest or experience in applying SRE principles such as reliability metrics, automation, and continuous improvement within a support or operations role.
  • Experience contributing to improved mean time to detect (MTTD) and mean time to restore (MTTR) through better observability, tooling, or process.
  • Understanding of the balance between feature delivery and operational stability in business‑critical systems.
Desired
  • Solid experience supporting trading platforms, financial exchanges, or real‑time transactional systems.
  • Reasonable exposure to FIX‑based workflows, messaging pipelines, or market‑connectivity architectures.
  • Experience working within AWS‑native or hybrid‑cloud financial environments.
  • Reasonable experience collaborating with front‑office trading desks or broker support teams.
  • Familiarity with CI/CD pipelines, DevOps practices, or automated deployment frameworks.
  • Exposure to SRE concepts such as SLIs, SLOs, error budgets, or reliability metrics within a support or operations context.
  • Experience improving platform resilience through automation, monitoring enhancements, or operational tooling.
  • Familiarity with failure scenarios, recovery patterns, or high‑availability strategies in distributed systems.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Production Engineer
Production Engineer

TP ICAP • Manila

On-site
Senior Production Engineer
Senior Production Engineer

863 Parameta Solutions (Singapore) Pte. Limited • Taguig

On-site
PHP 558,000 - 781,200
Production Engineer
Production Engineer

768 TP ICAP Management Services Ltd (Philippines Branch) • Taguig

On-site
PHP 446,400 - 669,600
Team Lead - Production Engineering
Team Lead - Production Engineering

768 TP ICAP Management Services Ltd (Philippines Branch) • Taguig

On-site
PHP 1,200,000 - 1,500,000
Sr. Production Engineer (Trading Applications) – Hybrid | Taguig City, Philippines
Sr. Production Engineer (Trading Applications) – Hybrid | Taguig City, Philippines

Charterhouse Pte Ltd • Taguig

On-site
PHP 600,000 - 800,000
Senior Production Engineer
Senior Production Engineer

TP ICAP • Manila

On-site
PHP 900,000 - 1,800,000
Lead Application Support Analyst
Lead Application Support Analyst

Morgan McKinley • Philippines

Hybrid
PHP 1,440,000 - 2,880,000
Senior Production Engineer - Trading Platform Reliability
Senior Production Engineer - Trading Platform Reliability

Charterhouse Pte Ltd • Taguig

On-site
PHP 600,000 - 800,000
Team Lead - Production Engineering
Team Lead - Production Engineering

TP ICAP • Manila

On-site
PHP 1,800,000 - 2,400,000
Lead Application Support Engineer
Lead Application Support Engineer

Morgan McKinley • Taguig

On-site
PHP 1,200,000 - 1,600,000