Site Reliability Engineer

HUB24 Limited

Sydney

Hybrid

AUD 120,000 - 180,000

Full time

18 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Flexible hybrid work
Employee Share Scheme
Leave & wellbeing support
Enhanced parental leave
Discounts & financial wellbeing

Job summary

HUB24 Limited is expanding its Site Reliability Engineering team in Sydney. You will support reliable production environments, lead incident triage and RCA, and drive continuous service improvements across cloud and on-premises platforms.

You will work closely with engineering, infrastructure, and operations to implement SRE best practices, automate processes, and maintain high availability and security in a hybrid work setting.

Qualifications

  • 3-5 years of hands-on SRE/DevOps or Infrastructure Engineering.
  • Experience in incident triage, troubleshooting, and post-incident reviews.
  • Working knowledge of AWS and/or GCP cloud environments.
  • Experience using observability tools for monitoring and performance analysis.
  • System administration across Linux and Windows.
  • Exposure to Docker and Kubernetes containerisation/orchestration.
  • Experience with automation and IaC tools (Terraform, Ansible).
  • Strong communication and cross-team collaboration.

Responsibilities

  • Demonstrate hands-on SRE/DevOps expertise to support reliable, scalable production environments.
  • Lead incident triage and RCA to drive permanent fixes.
  • Monitor performance, availability, and security across HUB platforms.
  • Apply SRE practices such as SLOs/SLIs and incident post-mortems.
  • Support AWS/GCP cloud environments and containerised workloads.
  • Contribute to automation and runbooks to reduce manual effort.
  • Participate in on-call rotations for critical systems.

Skills

SRE/DevOps experience
Incident management
Cloud platforms (AWS/GCP)
Observability tools (Dynatrace)
Linux/Windows administration
Containerisation (Docker) & Orchestr (

Tools

Dynatrace
Docker
Kubernetes
Terraform
Ansible
AWS
GCP

Job description

About HUB24

At HUB24, we’re rethinking the way wealth management works, combining platform, technology and data to create better outcomes for financial professionals and their clients.


Our purpose is simple: Empower better financial futures, together.


What sets us apart is how we work. We back bold thinking, move with pace, and turn ideas into action. You'll have the opportunity to make a real impact across your team, the business, and for the clients we support every day.


HUB24 Limited is an ASX-listed company (ASX: HUB) and part of the ASX100. We have over 1,100 employees across Australia, with offices in Sydney, Melbourne, Brisbane, Perth and the Gold Coast.


Why you’ll enjoy working here

We create an environment where you can do your best work and see the impact of it.



  • Work with smart, collaborative people who get things done.

  • Your ideas won’t sit in a backlog, they'll be heard, tested and actioned.

  • Grow your career your way, with support to learn, stretch and explore new opportunities.


We also offer benefits to support you inside and outside of work:



  • Genuinely flexible and hybrid ways of working.

  • Employee Share Scheme.

  • Additional leave and wellbeing support.

  • Enhanced parental leave and support through different life stages.

  • Everyday benefits, including discounts and financial wellbeing support.


Why this is an exciting opportunity

HUB24 is expanding its Site Reliability Engineering function and investing in Dynatrace as our core monitoring and observability platform. This is an opportunity to be part of that team which will play a critical role in ensuring the reliability, performance and scalability across our technology platforms.


Working closely with engineering, infrastructure and operations teams, you will embed SRE best practices, proactively manage system health and help design resilient, high-availability services that support our customers and business growth.


What You’ll Be Doing


  • Demonstrated hands-on experience in Site Reliability Engineering (SRE), DevOps, or Infrastructure Engineering, supporting reliable, high-performing, and scalable production environments.

  • Provide technical support for complex production incidents, ensuring timely resolution and minimising customer impact.

  • Lead incident triage, troubleshooting, root cause analysis, and Post Incident Reviews (PIRs), driving permanent fixes and continuous service improvements.

  • Monitor system performance, availability, and security across HUB platforms, proactively identifying and escalating risks.

  • Apply SRE best practices, including SLIs, SLOs, error budgets, and reliability engineering principles, to improve service resilience and operational excellence.

  • Support and troubleshoot AWS and GCP cloud environments, ensuring stability, performance, and operational efficiency.

  • Administer and maintain Linux and Windows servers, including performance tuning, configuration management, and operational support.

  • Support containerised environments using Docker and Kubernetes, helping maintain scalable and resilient workloads.

  • Manage server patching and vulnerability remediation activities in line with compliance, security, and risk management requirements.

  • Contribute to automation initiatives, developing operational scripts and runbooks to reduce manual effort and improve consistency.

  • Participate in an on-call roster, providing support for critical systems, platforms, and services as required.


What You Bring

You don’t need to tick every box, but experience in the below will set you up for success.


Experience & Skills


  • 3-5 years of hands-on experience in Site Reliability Engineering, Observability Engineering, DevOps, or Infrastructure Engineering, including support for production systems and incident response.

  • Practical experience in incident management, including triage, troubleshooting, root cause analysis, and participation in on-call support for high-availability environments.

  • Working knowledge of cloud platforms such as AWS and/or GCP, including infrastructure troubleshooting, operational support, and optimisation activities.

  • Experience using observability tools, preferably Dynatrace, for monitoring, alerting, dashboarding, and performance analysis.

  • Solid system administration skills across Linux and Windows environments, including troubleshooting and basic performance tuning.

  • Exposure to containerisation and orchestration technologies such as Docker and Kubernetes is preferred.

  • Experience supporting server patching and vulnerability remediation, including prioritisation based on risk and compliance requirements.

  • Experience with automation and Infrastructure as Code tools such as Terraform, Ansible, or similar is desirable.

  • Strong communication skills, with the ability to document findings, explain technical issues clearly, and collaborate effectively across teams.


Our process

We aim to keep the process simple and respectful of your time:



  • You’ll receive an acknowledgement after applying.

  • Our Talent team will review your application and keep you updated.

  • If shortlisted, we’ll connect to learn more about you.

  • Interviews may be virtual or in person.

  • You’ll receive an outcome and feedback.


If you need any adjustments, please let us know - we’re here to support you.


Our commitment

We’re committed to building an inclusive environment where everyone feels valued and supported to do their best work. We welcome applications from people of all backgrounds, identities and experiences.


Agencies, we work with a panel of preferred suppliers and are not accepting any unsolicited CVs.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Systems Security Engineer
Systems Security Engineer

HUB24 • Sydney

On-site
AUD 150,000 - 210,000
Flexible work
Employee share scheme
Additional leave and wellbeing support
+2
Senior Software Engineer (Full stack)
Senior Software Engineer (Full stack)

HUB24 Limited • City of Melbourne

Hybrid
AUD 140,000 - 180,000
Flexible & hybrid work
Employee Share Scheme
Additional leave
+2
Software Engineer
Software Engineer

HUB24 • Sydney

On-site
AUD 100,000 - 150,000
DevSecOps Engineer
DevSecOps Engineer

HUB24 Limited • Australia

Hybrid
AUD 120,000 - 170,000
Flexible hybrid working
Employee Share Scheme
Additional leave and wellbeing support
+2
HUB24 Automation Centre (HAC) Team Lead
HUB24 Automation Centre (HAC) Team Lead

HUB24 • Sydney

Hybrid
AUD 140,000 - 190,000
Flexible & hybrid work
Employee Share Scheme
Extra leave & wellbeing
+2
Systems Security Engineer
Systems Security Engineer

HUB24 Limited • Sydney

On-site
AUD 120,000 - 180,000
Genuinely flexible and hybrid working
Employee Share Scheme
Additional leave and wellbeing support
+2
Systems Security Engineer
Systems Security Engineer

Hub24 Management Services Pty Ltd • Sydney

On-site
AUD 140,000 - 190,000
Flexible/hybrid work
Employee Share Scheme
Leave & wellbeing
+2
Senior Test Analyst
Senior Test Analyst

HUB24 • City of Melbourne

Hybrid
AUD 120,000 - 150,000
Flexible and hybrid working
Employee Share Scheme
Extra leave and wellbeing support
+2
Insights Analyst
Insights Analyst

Hub24 Management Services Pty Ltd • Sydney

Hybrid
AUD 110,000 - 140,000
Senior Test Analyst
Senior Test Analyst

Hub24 Management Services Pty Ltd • City of Melbourne

On-site
AUD 110,000 - 150,000
Hybrid work model
Employee Share Scheme
Additional leave and wellbeing support
+1