SRE: AI-Driven Cloud Reliability & Automation

Cover Genius

Sydney

On-site

AUD 120,000 - 180,000

Full time

4 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Cover Genius is seeking a Site Reliability Engineer to lead reliability and infrastructure across multiple teams. You will shape system design, tooling, and processes for scalable production systems and ensure rapid, safe delivery of features.

You will bring strong cloud, observability, and automation skills, with a track record in Kubernetes, IaC, AWS/GCP, and AI-enabled development. The role emphasizes collaboration, security, and proactive risk reduction.

Qualifications

  • 3+ years of experience in SRE, Platform Engineering, DevOps or related roles.
  • Strong understanding of SRE and platform engineering principles with cross-team project leadership.
  • Experience with observability tools such as Datadog, Elasticsearch, Prometheus, Grafana.
  • Experience with cloud native tech: Docker, Kubernetes; infrastructure as code with Terraform.
  • Scripting and internal tooling with Bash and at least one language (Python/Go).
  • Fluency with AI-assisted development environments (Cursor, Claude Code, Codex) and applying them to production workflows.
  • Experience with Linux, networking, distributed systems and high-availability deployments.
  • Bachelor's degree in CS/Engineering or equivalent desirable.

Responsibilities

  • Lead reliability and infrastructure projects across teams and domains.
  • Architect and build reliable, highly-available cloud infrastructure across products.
  • Develop observability standards, tooling, and dashboards for multiple teams.
  • Champion SLOs and incident response; act as incident commander for major outages.
  • Reduce toil through automation and self-service tooling.
  • Develop and maintain runbooks and incident playbooks.
  • Apply AI-assisted development to infrastructure problems and guide teams.
  • Contribute to capacity planning and cost optimization.
  • Mentor engineers on systems thinking and production ownership.
  • Implement security best practices in pipelines and infra.

Skills

SRE
Platform Eng
DevOps
Observability
Kubernetes
Terraform
Python/Go
AI tools
Linux
Networking
AWS/GCP

Education

Bachelor's degree in Computer Science/Engineering

Tools

Datadog
Elasticsearch
Prometheus
Grafana
Docker
Kubernetes
Terraform
Bash
Python
Go
Cursor
Claude Code
Codex

Job description

Cover Genius is seeking a Site Reliability Engineer to lead reliability and infrastructure across multiple teams. You will shape system design, tooling, and processes for scalable production systems and ensure rapid, safe delivery of features.

You will bring strong cloud, observability, and automation skills, with a track record in Kubernetes, IaC, AWS/GCP, and AI-enabled development. The role emphasizes collaboration, security, and proactive risk reduction.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI-First Site Reliability Engineer — Hybrid
AI-First Site Reliability Engineer — Hybrid

King River Capital Group • New South Wales

Hybrid
AUD 140,000 - 180,000
Flexible Work Environment
AI-Driven SRE: On-Prem & Cloud Reliability Lead
AI-Driven SRE: On-Prem & Cloud Reliability Lead

Precisely • Australia

On-site
AUD 140,000 - 190,000
Senior SRE - AI-Driven Cloud Reliability, Multi-Cloud
Senior SRE - AI-Driven Cloud Reliability, Multi-Cloud

Mantel Group • Sydney

Hybrid
AUD 140,000 - 210,000
Senior Cloud SRE: Scale, Automation & Observability
Senior Cloud SRE: Scale, Automation & Observability

Elasticsearch B.V. • Australia

Remote
AUD 180,000 - 260,000
Competitive pay
Health coverage
Flexible location
+4
Site Reliability Engineer: Kubernetes, Terraform & Automation
Site Reliability Engineer: Kubernetes, Terraform & Automation

Future Secure AI • City of Melbourne

On-site
AUD 100,000 - 150,000
Competitive salary
Flexible work environment
High-performance culture
+1
Associate SRE: Build, Automate & Improve Reliability
Associate SRE: Build, Automate & Improve Reliability

Culture Amp Pty Ltd • City of Melbourne

On-site
AUD 77,000 - 104,000
Equity through Employee Share Option
Learning programs and coaching
Quarterly refresh days
+4
Associate SRE: Growth-Focused Reliability Engineer (Hybrid)
Associate SRE: Growth-Focused Reliability Engineer (Hybrid)

Culture Amp • City of Melbourne

Hybrid
AUD 81,000 - 99,000
Employee equity through option program
Learning programs and coaching
Quarterly refresh days
+6
Senior Cloud SRE: AWS, Observability & Automation
Senior Cloud SRE: AWS, Observability & Automation

CareCone Group • Sydney

On-site
AUD 180,000 - 240,000
AI-Driven SRE & Infrastructure Automation Engineer
AI-Driven SRE & Infrastructure Automation Engineer

Sky • Port Stephens Council

Hybrid
AUD 150,000 - 210,000
Free SkyTV/ NOW package
Pension up to 9%
Private healthcare
+4
Senior SRE: AI-Powered Cloud & Kubernetes (AUS)
Senior SRE: AI-Powered Cloud & Kubernetes (AUS)

Nucleus Security • City of Melbourne

Hybrid
AUD 150,000 - 210,000
Health insurance
Equity in startup
Flexible PTO
+1