Site Reliability Engineer: Data Platform at Scale

Talenza

Sydney

On-site

AUD 120,000 - 180,000

Full time

3 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Global Trading Firm in Sydney seeks an experienced Site Reliability Engineer to join a high‑performing Data Engineering team. You will operate a multi‑petabyte data platform, ensuring reliability, scalability and efficient deployments.

You will work hands-on with Linux, Kafka, HDFS, Dremio and in-house tooling, focusing on operational excellence and automation. This role offers genuine ownership and growth in a distributed, production environment.

Qualifications

  • 2–3 years in an SRE/platform/infrastructure/production engineering role.
  • Strong Linux troubleshooting across processes, filesystems, networking and memory pressure.
  • Hands-on operator experience with Kafka, HDFS or Kubernetes (deployed/configured/upgraded/tuned).
  • Experience operating self-managed infrastructure (bare metal, data centre, self-run VMs or self-managed Kubernetes).
  • Python experience for automation, health checks or deployment workflows.
  • Exposure to Docker, Kubernetes, Helm and infrastructure-focused CI/CD.
  • Curious, pragmatic mindset with interest in how complex systems behave under pressure.

Responsibilities

  • Run, monitor and improve large-scale data platforms including Kafka, HDFS, Dremio and in-house pipelines.
  • Troubleshoot production issues across Linux, storage, networking and distributed infrastructure.
  • Build automation and CI/CD to speed up deployments and improve safety and repeatability.
  • Support upgrades, capacity planning, incident response and long-term reliability improvements.
  • Collaborate with systems and network engineers, developers, researchers and end users to solve platform problems.
  • Evaluate and introduce new technologies as the environment evolves.

Skills

Linux troubleshooting
Python automation

Tools

Kafka
HDFS
Kubernetes
Docker
Helm
Terraform
Ansible
Puppet
Airflow
Presto
Dremio

Job description

Global Trading Firm in Sydney seeks an experienced Site Reliability Engineer to join a high‑performing Data Engineering team. You will operate a multi‑petabyte data platform, ensuring reliability, scalability and efficient deployments.

You will work hands-on with Linux, Kafka, HDFS, Dremio and in-house tooling, focusing on operational excellence and automation. This role offers genuine ownership and growth in a distributed, production environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer - Data Engineering
Site Reliability Engineer - Data Engineering

Talenza • Sydney

On-site
AUD 120,000 - 180,000
Site Reliability Engineer - Data Engineering
Site Reliability Engineer - Data Engineering

IMC B.V. • Sydney

On-site
AUD 120,000 - 180,000
Site Reliability Engineer - Data Engineering
Site Reliability Engineer - Data Engineering

IMC Trading • Sydney

On-site
AUD 120,000 - 170,000
Data Platform SRE: Scale & Automate Data Pipelines
Data Platform SRE: Scale & Automate Data Pipelines

IMC Trading • Sydney

On-site
AUD 120,000 - 170,000
Site Reliability Engineer
Site Reliability Engineer

Tribus • Sydney

On-site
AUD 150,000 - 190,000
Senior Site Reliability Engineer — Scale Global Infra
Senior Site Reliability Engineer — Scale Global Infra

Blinq • Sydney

On-site
AUD 140,000 - 200,000
Equity & ownership
Competitive salary
Generous paid time off
+2
Data - Site Reliability Engineer
Data - Site Reliability Engineer

Optiver • Sydney

On-site
AUD 90,000 - 130,000
Performance-based bonus
Training & mentorship
Daily breakfast & in-house barista
+3
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Firesoft People • Sydney

On-site
AUD 230,000 - 320,000
Data Platform SRE: Scale, Automation & Delivery
Data Platform SRE: Scale, Automation & Delivery

IMC B.V. • Sydney

On-site
AUD 120,000 - 180,000
Site Reliability Engineer: AI Infra & HPC
Site Reliability Engineer: AI Infra & HPC

Firmus Technologies • City of Melbourne

On-site
AUD 120,000 - 180,000