A complete application in a minute — tailored resume and cover letter, ready to send.
Tribus in Sydney is seeking a hands-on Site Reliability Engineer (Trading Infrastructure) to design, build and maintain Linux-based production systems powering a global electronic trading platform.
You will implement IaC with Terraform and Ansible, build observability with Prometheus, Grafana and Splunk, and collaborate with software engineers across C++, Java, C# and Python components to improve latency, automation and incident response.
Site Reliability Engineer (Trading Infrastructure) Multiple organisations
Sponsorship Available
Sydney or Hong Kong | Onsite | Global Trading Environment
Join a high-performance engineering team responsible for the reliability, scalability and operational excellence of the infrastructure powering a global electronic trading platform.
This is a hands-on Site Reliability Engineering role sitting close to the trading stack, where you'll help build resilient production systems, improve automation, and work alongside software engineers to support latency-sensitive applications operating across global financial markets.
Unlike traditional SRE environments focused primarily on SLI/SLO metrics, this team takes an event-driven approach to reliability engineering. You'll design intelligent monitoring and automated operational workflows that identify abnormal system behaviour, infrastructure anomalies and production events before they impact trading. The focus is on actionable signals, rapid diagnosis and engineering-led remediation rather than simply measuring service health.