Get more replies from employers
Send a job-specific resume in minutes.
TechTiera in Malaysia is seeking a senior infrastructure/production support engineer who can diagnose complex production issues and drive automation to improve stability.
You will own deep triage across OS, databases, logs and observability, design automation scripts, lead root-cause analysis, and coordinate with cross-functional teams to ensure service resilience and timely releases.
Jora Malaysia will close on 9th September 2026. Thank you for being with us, we are cheering you on as you continue your career journey.
This role is intended for senior engineers who can independently troubleshoot complex production issues, improve stability through automation and guide junior support resources.
Own deeper technical triage across OS, database, batch, middleware/application logs, observability and platform dependencies.
Design and enhance automation scripts for recurring support tasks, health checks, log extraction, alert enrichment and manual process reduction.
Lead problem analysis, trend identification and preventative actions for recurring incidents.
Support containerized or modernized environments where Docker/Kubernetes are part of the production or transformation landscape.
Support RBC communication surveillance and related upstream / downstream applications through incident triage, root-cause analysis and service restoration.
Maintain production stability through monitoring, alert analysis, capacity awareness, runbook execution and risk escalation.
Work with CTB, RTB, vendor and cross-functional technology teams to support production fixes, enhancements, transition readiness and release/change activities.
Use Jira and Confluence to maintain traceability of incidents, problems, changes, risks, user stories, knowledge articles and operational procedures.
Minimum 4 to 8 years of infrastructure/application production support experience, preferably in banking, capital markets or regulated technology.
Strong scripting depth in at least one language with ability to write, maintain and troubleshoot automation scripts independently.
Ability to write and optimize SQL queries, investigate database-related production issues and work with DBAs when deeper administration is required.
Hands-on experience in ITIL-based Incident, Problem and Change Management in production environments.
Strong operating system exposure across Linux / Unix / RHEL and Windows, with ability to troubleshoot application and infrastructure issues.
Scripting capability in at least one relevant language, preferably Bash, Python, PowerShell, Perl or batch scripting, with evidence of automation or support tooling.
Working knowledge of relational databases such as MS SQL and/or Sybase, including query execution, query writing and production issue investigation.
Experience using enterprise monitoring and observability tools such as Splunk, ITRS Geneos, AppDynamics, Dynatrace, Nagios, ELK or Grafana.
Ability to document support procedures, incident notes, change records and runbooks using Jira and Confluence or equivalent tools.
Banking, financial services, capital markets or regulated production support exposure.
Exposure to communication surveillance, conduct surveillance or regulatory surveillance platforms such as Smarsh Enterprise Conduct, Csurv or Theta Lake.
Experience with batch scheduling tools such as BMC Control-M, Autosys or Tidal.
Exposure to Docker and Kubernetes, especially troubleshooting deployed containers or supporting application platforms in production.
Understanding of sales tooling, CRM, trading or capital markets technology workflows.
Experience supporting disaster recovery, resilience testing, configuration management and production readiness reviews.