Site Reliability Engineer

Leidos

City of Melbourne

On-site

AUD 120,000 - 160,000

Full time

4 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Flex Your Way
Wellbeing That Matters
Learn Without Limits
Work That Makes an Impact
Build Your Future
And Much More

Job summary

Leidos Australia is seeking a Site Reliability Engineer to design, build, operate and support mission-critical platforms for Defence, Intelligence and Government sectors. You will ensure reliability, performance and resilience in highly regulated environments, working across Level 2/3 incident response, monitoring, and automation initiatives.

The role offers impactful, federally-relevant work with strong development and platform collaboration.

Qualifications

  • Experience working as a Site Reliability Engineer, Systems Engineer or in a comparable infrastructure-focused role.
  • Experience in Linux/UNIX administration, including troubleshooting, performance analysis and operational support within enterprise environments.
  • Capability in developing automation using Python, Go, Bash or similar scripting languages to improve operational efficiency and reduce manual effort.
  • Experience implementing and maintaining Infrastructure as Code and configuration management solutions using tools such as Terraform, Ansible, Puppet or Chef.
  • Knowledge of monitoring, logging and observability practices, with the ability to use service metrics and operational data to improve reliability and performance.
  • Experience supporting virtualised and/or cloud-based environments, including technologies such as VMware, AWS, Azure or GCP.
  • Experience working with CI/CD pipelines and deployment automation tools such as Jenkins, GitLab CI or similar platforms.
  • Ability to investigate complex incidents, contribute to root cause analysis and implement sustainable corrective actions.

Responsibilities

  • Support the availability, capacity and reliability of complex production environments, applying software engineering principles to infrastructure and operational challenges.
  • Provide Level 2 and Level 3 operational support, leading incident response activities, contributing to root cause analysis, and implementing preventative improvements that enhance system resilience.
  • Develop and maintain monitoring, alerting and logging capabilities, using meaningful service metrics, SLOs and SLIs to drive performance insights and continuous improvement.
  • Enhance automation, infrastructure-as-code and CI/CD practices across the environment, partnering with development and platform teams to enable consistent, secure and reliable deployments.

Skills

Linux/UNIX administration
Monitoring & observability
Incident response & RCA
Automation scripting

Tools

Terraform
Ansible
Puppet
Chef
Jenkins
GitLab CI
VMware
AWS
Azure
GCP
Bash scripting
Python
Go

Job description

At Leidos Australia, our work goes beyond a job title. We design, build, operate and support the systems, services and capabilities that help protect and advance the Australian way of life.

Backed by a global organisation of more than 40,000 people, our Australian team of around 2,000 has partnered with government, defence and national security customers for more than 30 years. We deliver mission-critical solutions where reliability, trust and purpose matter.

Leidos is a destination for people who want their work to have purpose, their contribution to be recognised, and their career to grow without needing to fit a single mould. Meaningful work is supported by benefits, flexibility and career pathways that help you grow while balancing life outside work. This includes:

  • Flex Your Way - Tailor work to your life with options including Life Days (up to 12 extra days off per year), part-time arrangements, job sharing, and compressed hours.
  • Wellbeing That Matters - Free access to Headspace and MeQuilibrium, plus dedicated wellness spaces to support your mental, physical, and emotional wellbeing.
  • Learn Without Limits - Access over 120,000 on-demand courses, alongside technical, project management, leadership, and professional development opportunities.
  • Work That Makes an Impact - Contribute to nationally significant projects that matter, solve complex challenges, and accelerate your career through unique growth opportunities.
  • Build Your Future - Whether you're deepening technical expertise or preparing for leadership, we'll invest in your development every step of the way.
  • And Much More - These are just a few of the ways we invest in our people and help them thrive.
About the Team

You'll join a team where people work with integrity, inclusion, innovation, agility, collaboration and commitment. In practice, that means sharing ownership, solving complex problems together and staying connected to the customer context, the purpose behind the work and how each role contributes to meaningful outcomes.

The Mission Software Solutions business delivers customized and off-the-shelf software solutions to Defence, Intelligence and other Government sectors. The business also builds capabilities that can be used for a wide range of applications in both traditional and non-traditional environments; providing leadership in the advancement of technological modernisation of complex enterprises within Government.

The Opportunity

As a Site Reliability Engineer within our AGO GEO SPACES program, you'll play a key role in delivering and sustaining technology solutions that support critical national security outcomes. Working in a secure, highly regulated environment, you'll help ensure the reliability, performance and resilience of mission-critical platforms while collaborating with engineering and operational teams to maintain exceptional service standards.

  • Support the availability, capacity and reliability of complex production environments, applying software engineering principles to infrastructure and operational challenges.
  • Provide Level 2 and Level 3 operational support, leading incident response activities, contributing to root cause analysis, and implementing preventative improvements that enhance system resilience.
  • Develop and maintain monitoring, alerting and logging capabilities, using meaningful service metrics, SLOs and SLIs to drive performance insights and continuous improvement.
  • Enhance automation, infrastructure-as-code and CI/CD practices across the environment, partnering with development and platform teams to enable consistent, secure and reliable deployments.
About You

You are someone who enjoys solving complex operational challenges and takes a thoughtful, collaborative approach to improving system reliability. You bring a combination of infrastructure expertise, automation capability and a continuous improvement mindset, enabling you to contribute effectively within high-availability environments supporting critical business outcomes.

  • Experience working as a Site Reliability Engineer, Systems Engineer or in a comparable infrastructure-focused role supporting production environments.
  • Experience in Linux/UNIX administration, including troubleshooting, performance analysis and operational support within enterprise environments.
  • Capability in developing automation using Python, Go, Bash or similar scripting languages to improve operational efficiency and reduce manual effort.
  • Experience implementing and maintaining Infrastructure as Code and configuration management solutions using tools such as Terraform, Ansible, Puppet or Chef.
  • Knowledge of monitoring, logging and observability practices, with the ability to use service metrics and operational data to improve reliability and performance.
  • Experience supporting virtualised and/or cloud-based environments, including technologies such as VMware, AWS, Azure or GCP.
  • Experience working with CI/CD pipelines and deployment automation tools such as Jenkins, GitLab CI or similar platforms.
  • Ability to investigate complex incidents, contribute to root cause analysis and implement sustainable corrective actions.
Advantageous
  • Experience supporting systems within Defence, National Security or other highly regulated environments.
  • Familiarity with fault-tolerant, large-scale production systems and high-availability infrastructure architectures.
  • Knowledge of secure systems administration practices, compliance requirements and operational governance frameworks.
  • Exposure to DevOps, platform engineering or site reliability practices within mission-critical environments.

Due to the nature of this role, you must be an Australian Citizen holding an active TSPV security clearance, with the ability to successfully obtain and maintain a DISA/OSA.

Be careful - Don’t provide your bank or credit card details when applying for jobs. Don't transfer any money or complete suspicious online surveys. If you see something suspicious, report this job ad .

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Leidos • Bendigo

On-site
AUD 140,000 - 190,000
Flexible work options
Wellbeing programs
Learning opportunities
+1
Site Reliability Engineer
Site Reliability Engineer

Leidos Australia • Canberra

On-site
AUD 140,000 - 190,000
Life Days
Part-time options
Wellbeing programs
+3
Site Reliability Engineer
Site Reliability Engineer

Leidos Australia • Wollongong City Council

On-site
AUD 120,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

Leidos • Wollongong City Council

On-site
AUD 120,000 - 180,000
Flex Your Way
Wellbeing That Matters
Learn Without Limits
+3
Windows Administrator
Windows Administrator

Leidos • City of Melbourne

On-site
AUD 110,000 - 160,000
Senior DevSecOps Engineer
Senior DevSecOps Engineer

Leidos Australia Pty Ltd • City of Melbourne

On-site
AUD 140,000 - 190,000
Flexibility options
Wellbeing programs
Learning platform access
+2
Site Reliability Engineering and Platform Lead
Site Reliability Engineering and Platform Lead

Leidos • City of Melbourne

On-site
AUD 180,000 - 240,000
Senior Full-Stack Platform Integrations Engineer
Senior Full-Stack Platform Integrations Engineer

Visa Hunt • Canberra

On-site
AUD 120,000 - 150,000
Flexibility
Wellbeing
Learning opportunities
+2
Linux Systems Administrator
Linux Systems Administrator

Leidos Australia • Bendigo

On-site
AUD 90,000 - 130,000
Flexible work options
Wellbeing programs
Learning resources
+2
Systems Administrator | Kubernetes
Systems Administrator | Kubernetes

Leidos • City of Melbourne

On-site
AUD 120,000 - 180,000
Flexibility and wellbeing programs
Learning & development resources
Career growth opportunities
+1