Site Reliability Engineer - Appian

Luxoft

Sydney

On-site

AUD 140,000 - 190,000

Full time

13 days ago
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Luxoft in Sydney, Australia is seeking a Site Reliability Engineer with 10+ years of experience to improve service reliability and operational resilience across client banking platforms. You will own monitoring, resilience testing, disaster recovery planning, and manage SLIs and SLOs.

Responsibilities include designing observability capabilities, implementing automation and IaC, contributing to major incident response, and advancing AI-driven operational solutions using LLMs and AI-assisted

Qualifications

  • SRE/Production Engineering or Platform Engineering background with distributed systems
  • Experience with cloud-native tech and Kubernetes/OpenShift
  • Observability responsibilities across metrics, logs, traces with tooling like Elastic/Splunk/Grafana/Prometheus/OpenTelemetry
  • Familiarity with IaC, DevSecOps, and zero trust security principles
  • CI/CD automation using Ansible or Chef
  • Programming/scripting in Python, Java, Go, PowerShell
  • Ability to triage incidents, runbooks, RCAs and improvements

Responsibilities

  • Improve reliability, availability, and resilience via robust monitoring, resilience testing, disaster recovery planning, and SLO/SLI management
  • Design and enhance observability across customer journeys and cloud environments
  • Develop automation, self-healing tooling, IaC and CI/CD practices to reduce manual work
  • Participate in major incident response, root cause analysis, and continuous improvement
  • Design and implement AI-driven operational solutions and AI-assisted engineering practices
  • Drive capacity management and performance optimization to meet current and future demand

Skills

SRE/Production Engineering
Cloud-native architectures
Kubernetes/OpenShift
Observability & monitoring
Appian logs & performance
Appian environment On Prem & Cloud
Elastic/ Splunk/ Grafana/ Prometheus/
IaC & DevSecOps
CI/CD tooling (Ansible/Chef)
Python
Java/Go/PowerShell
Troubleshooting & RCA

Tools

Appian

Job description

Project description

We are seeking a Site Reliability Engineer with more than 10+ years of experience to support our client in the Australian banking sector. The role centers on improving service reliability, availability, and operational resilience through comprehensive monitoring, resilience testing, disaster recovery planning, and management of SLIs and SLOs.

Responsibilities
  • Improve service reliability, availability, and operational resilience through robust monitoring, resilience testing, disaster recovery planning, and management of SLIs, SLOs, and error budgets.
  • Design and enhance observability capabilities, including monitoring, alerting, dashboards, synthetic monitoring, automated health checks, and visibility across customer journeys and cloud environments.
  • Develop automation, self-healing capabilities, operational tooling, Infrastructure as Code and CI/CD practices to reduce manual effort and improve platform reliability.Participate in major incident response, service restoration, root cause analysis, and continuous improvement initiatives to minimize customer impact and improve recovery times.
  • Design and implement AI-driven operational solutions, utilize LLMs for incident triage and remediation, and drive adoption of GitHub Copilot and AI-assisted engineering practices.
  • Drive capacity management, performance optimization and reliability improvements to ensure services meet current and future demand.

SKILLS
Must have
  • Demonstrated experience in Site Reliability Engineering (SRE), Production Engineering, DevOps, or Platform Engineering, with a strong understanding of distributed systems, cloud-native technologies, Kubernetes/OpenShift, and modern application architectures.
  • Appian – Ability to use Appian futures and use scripting & tools to implementation solutions to reads logs, analyses performance impact/issues and drives fixes to resolution.
  • Having experience & knowledge on Appian environment setups including On Prem & Cloud.
  • Observability - Ability to use scripting and tooling to implement observability solutions, enabling the collection, analysis, and visualization of metrics, logs, and traces to support incident detection, diagnosis, and continuous service improvement. Familiarity with tools such as Elastic, Splunk, Grafana, Prometheus, OpenTelemetry, or similar enterprise monitoring platforms, and drive preventative improvements that enhance application performance and availability.
  • Knowledge of enterprise application, middleware, and content management platforms, including FileNet, ICN, WAS, container platforms, databases and integration technologies.
  • Reliability and Scalability - Ability to design and operate systems for high availability, fault tolerance, and disaster recovery, while ensuring systems can scale to meet current and future demand.
  • Familiarity with Infrastructure as Code (IaC), DevSecOps practices and Zero Trust security principles.
  • Proficiency in DevOps toolkits such as Ansible and Chef to support and drive CI/CD and reliability-automation initiatives.
  • Programming and scripting languages including Python, Java, Go, PowerShell, or equivalent technologies to support automation and platform engineering initiatives.
  • Troubleshooting - Capability to systematically identify, diagnose, and resolve technical issues across systems, applications, and networks, using analytical methods and tools to restore functionality, minimize disruption, and ensure stable operations.

Nice to have

Python

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer - Appian
Site Reliability Engineer - Appian

Luxoft • Australia

On-site
AUD 140,000 - 190,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Peoplebank • Sydney

On-site
AUD 150,000 - 190,000
Site Reliability Engineer
Site Reliability Engineer

Tata Consultancy Services Limited • Sydney

On-site
AUD 90,000 - 120,000
Appian Developer or Sr. Developer or Solution Designer (Sydney or Melbourne)
Appian Developer or Sr. Developer or Solution Designer (Sydney or Melbourne)

Viable Solutions Pty Ltd • Sydney

On-site
AUD 120,000 - 160,000
Appian Developer
Appian Developer

Zone IT Solutions • Sydney

On-site
AUD 120,000 - 170,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Firesoft People • Sydney

On-site
AUD 230,000 - 320,000
Appian Developer
Appian Developer

Infosys • Sydney

On-site
AUD 109,000 - 120,000
Income Protection Insurance
Paid Parental and Volunteer leaves
Employee Assistance Program (EAP)
+4
Site Reliability Engineer - Data Engineering
Site Reliability Engineer - Data Engineering

Talenza • Sydney

On-site
AUD 120,000 - 160,000
Site Reliability Engineer
Site Reliability Engineer

Tribus • Sydney

On-site
AUD 150,000 - 190,000
Appian Technical Lead
Appian Technical Lead

Tata Consultancy Services • Sydney

On-site
AUD 150,000 - 210,000