Cloud Site Reliability Engineer

Australian Payments Plus

Sydney

On-site

AUD 120,000 - 180,000

Full time

5 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

AP+ is seeking a Cloud Site Reliability Engineer to keep our cloud platform reliable and scalable. You’ll operate AWS/Kubernetes environments, automate tasks and strengthen observability to reduce toil and incidents.

You’ll work across EKS, EC2, EFS, VPC and IAM, using IaC tools like Terraform or CDK, and collaborate with Security and Service Management to ensure resilience.

Qualifications

  • Hands-on experience in Site Reliability, Cloud or Platform Engineering.
  • Strong AWS expertise across EKS, EC2, EFS, VPC and IAM.
  • Kubernetes and containerised workloads management and scaling.
  • IaC and automation using Terraform, CloudFormation, CDK or Ansible with Python/TypeScript.
  • CI/CD pipelines on GitHub/GitLab/Bitbucket with strong observability.

Responsibilities

  • Operate, monitor and improve cloud and platform infrastructure for high availability, resilience, performance and security.
  • Respond to incidents as first responder, troubleshooting reliability issues and reducing MTTR.
  • Engineer and optimise AWS and Kubernetes environments (EKS, EC2, EFS, VPC, IAM).
  • Build and improve CI/CD pipelines and automate operational tasks.
  • Strengthen observability across infra, apps and cloud services.
  • Build resilience via disaster recovery and controlled changes in collaboration with teams.

Skills

Site Reliability
Cloud Engineering
Kubernetes
AWS
EKS
EC2
EFS
VPC
IAM
Kubernetes management
IaC
Terraform
CloudFormation
CDK
Ansible
Python
TypeScript
CI/CD
Observability
Automation
Disaster Recovery

Tools

GitHub
GitLab
Bitbucket

Job description

Life @ AP+:

We are one connected team in pursuit of one inspiring purpose - to unite people and technology to power better experiences. Each of us has a part to play in making that happen. You’ll be encouraged to bring your big ideas forward and make a difference through your work. Taking steps forward in your career whilst still having room for fun, friendships, and flexibility in your daily life.

We’re driven by our core values: lead with heart, learn for tomorrow and live our legacy. A purpose like ours takes the inspired impact of an incredible team. Ready to change the game? We’re ready to help you do it.

The Purpose:

Reliability matters when you’re supporting technology that Australians depend on every day.

As a Cloud Site Reliability Engineer, you’ll help keep AP+’s cloud and platform services reliable, available, secure and resilient. This is a hands-on engineering role where you’ll combine cloud infrastructure, automation, observability and SRE practices to improve the performance and reliability of critical technology services.

Working primarily across AWS and Kubernetes environments, you’ll engineer for resilience, respond to complex incidents, reduce operational toil through automation and continually improve how our platforms perform at scale.

Key Responsibilities the Role Owns:
  • Operate, monitor and continually improve cloud and platform infrastructure, engineering for high availability, resilience, performance and security.
  • Act as a first responder to cloud and platform incidents, troubleshooting complex reliability, capacity and performance issues while improving MTTD and MTTR through lessons learned.
  • Engineer and optimise AWS and Kubernetes environments, including EKS, EC2, EFS, VPC and IAM, supporting scalable and resilient workloads.
  • Build and improve CI/CD pipelines and automate operational tasks to reduce manual effort, minimise risk and enable safe, frequent and reliable delivery.
  • Strengthen observability across infrastructure, applications and cloud services, using meaningful metrics to identify issues and drive improvements in availability and performance.
  • Build resilience into our platforms through disaster recovery, controlled change, configuration management and close collaboration with Engineering, Security and Service Management teams.
You’ll likely be a strong fit for this role if:
  • You bring hands-on experience in Site Reliability, Cloud or Platform Engineering, supporting complex and highly available cloud environments.
  • You have strong AWS expertise, ideally across EKS, EC2, EFS, VPC and IAM, underpinned by a solid understanding of cloud networking.
  • You’re experienced with Kubernetes and containerised workloads, including cluster management, scaling, upgrades and optimisation.
  • Infrastructure-as-code and automation are central to how you work, with experience using tools such as Terraform, CloudFormation, CDK or Ansible, along with Python, TypeScript or similar scripting languages.
  • You’ve designed and operated CI/CD pipelines using platforms such as GitHub, GitLab or Bitbucket and understand the importance of strong observability and operational monitoring.
  • You bring an engineering mindset focused on reliability and resilience, with experience across cloud security, incident response, disaster recovery and controlled change, ideally within a regulated or high-availability environment.
What happens next:

At AP+, we believe in the power of passion, pride and purpose. Our team is driven by a shared mission to make a difference in the world of payments, and we're proud to work together towards this common goal.

If you’re an SRE or Cloud Engineer who gets excited about automation, observability and engineering platforms that need to be there when it matters, we’d love to hear from you.

We want to remove all barriers to inclusion, so if you need advice or support with your application, we’re here to help. Please reach out to recruitment@auspayplus.com.au. We also encourage you to let us know your pronouns at any point during the recruitment process.

AP+ are not partnering with Recruitment agencies for this role
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Cloud Site Reliability Engineer
Cloud Site Reliability Engineer

Australian Payments Plus Limited • Sydney

Hybrid
AUD 140,000 - 210,000
Cloud Engineer
Cloud Engineer

Australian Payments Plus • Sydney

Hybrid
AUD 120,000 - 170,000
Cloud Engineer
Cloud Engineer

Australian Payments Plus Limited • Sydney

Hybrid
AUD 110,000 - 170,000
Application Support Analyst
Application Support Analyst

Australian Payments Plus • Council of the City of Sydney

On-site
AUD 90,000 - 120,000
ITSM Analyst: Drive Reliable Services
ITSM Analyst: Drive Reliable Services

Australian Payments Plus Limited • Sydney

Hybrid
AUD 90,000 - 120,000
ITSM Analyst
ITSM Analyst

Australian Payments Plus Limited • Sydney

On-site
AUD 90,000 - 120,000
Workplace Tech Support Champion
Workplace Tech Support Champion

Australian Payments Plus Limited • Sydney

Hybrid
AUD 65,000 - 90,000
Payments Platform Support Analyst (24/7)
Payments Platform Support Analyst (24/7)

Australian Payments Plus • Australia

On-site
AUD 90,000 - 120,000
Application Support Analyst
Application Support Analyst

auspayplus.com.au • Sydney

On-site
AUD 90,000 - 120,000
Payments Platform Support Analyst (24/7)
Payments Platform Support Analyst (24/7)

auspayplus.com.au • Sydney

On-site
AUD 90,000 - 120,000