Senior Site Reliability Engineer, ANZ

JobSpace

Auckland

On-site

NZD 120,000 - 180,000

Full time

9 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Relocation allowance
Wellness allowance
Parental leave
Commuting/car park allowance
Office lunches

Job summary

Partly seeks a Senior Site Reliability Engineer to own the reliability of our cloud infrastructure and Kubernetes clusters. You will lead SRE practices, build scalable systems, and collaborate across teams to deliver robust software in production.

You will optimize costs, drive automation, and mentor engineers while improving production readiness and incident response capabilities. Auckland-based offices support flexible work arrangements.

Qualifications

  • Hands-on with modern SRE practices and tooling.
  • Strong systems programming and software engineering foundations.
  • Experience scaling cloud infrastructure and CI/CD systems.

Responsibilities

  • Ensure stability, scalability, and security of cloud infra and Kubernetes clusters.
  • Monitor costs and optimize allocation for reliability and value.
  • Collaborate with developers and leadership to plan infrastructure improvements.
  • Drive production readiness, automate deployments, and fix issues fast.
  • Lead incident responses and assist teams with debugging across stack.

Skills

SRE practices
Docker/Kubernetes
Terraform
GitOps
Python/Bash
Cloud platforms
Linux
Leadership
Communication
Incident response

Tools

ArgoCD
Kafka

Job description

Jora New Zealand will close on 9th September 2026. Thank you for being with us, we are cheering you on as you continue your career journey.

Partly is connecting the world's parts, and we're doing that by building the AI infrastructure layer for the global repair industry, starting with the $2tn automotive market. Our frontier model, Interpreter, is the world's first AI purpose-built to understand vehicle damage and the parts needed to fix it. Thousands of businesses across the global repair supply chain already rely on it.

Founded by ex-Rocket Lab engineers, we've tripled in size in the last 18 months and have recently raised a $50m Series B led by DST Global (Anthropic, Airbnb, Meta, TikTok, Spotify) and including Blackbird Ventures (Canva, CultureAmp etc.), WNDR, Activant Capital, Icehouse Ventures, Square Peg, Airtree, and Ecliptic Venture Capital. We're headquartered in Austin, with offices in New Zealand and London.

We're continuing to build a world-class team ensuring Partly is a place where people can do the best work of their lives. We're proud of the culture we've built, and our values are lived throughout every experience.

This Role

Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, distributed systems, ensuring that both internally critical and externally visible services have the reliability, uptime, and performance appropriate to clients' needs while enabling a fast rate of improvement. SREs maintain constant awareness of system capacity and performance, ensuring our networks, platforms, and tools are scalable, secure, and reliable so engineers can focus on delivering impactful software. This senior role demands high autonomy, leadership, and strategic thinking, making it ideal for those excited by the challenge of designing and supporting the infrastructure that connects the world's parts.

What Will You Do
Reliability Engineering:

Ensure the stability, scalability, and security of our cloud infrastructure, Partly & 3rd party applications in our Kubernetes powered clusters. Leverage Infrastructure-as-Code and automation (Terraform for GCP, GitOps with ArgoCD, Custom scripts in Python/Bash, etc.) to deploy and manage workloads and resources in a repeatable, automated way.

Cost Optimisation:

Monitor and optimise costs across our cloud and on-prem infrastructure, ensuring we get maximum value from our investments. Make recommendations for resource allocation or architecture changes to improve cost-efficiency without sacrificing reliability or performance.

Cross-Functional Collaboration:

Work closely with developers, data engineers, and leadership to plan infrastructure needs and improvements. Provide tooling, guidance and training to the engineering team on SRE practices, and collaborate during software delivery to ensure smooth integrations from code to production.

Software Engineering:

Make sure our software meets high production readiness standards. When you see a problem or an opportunity to improve, you drive the solution.

Troubleshooting:

Participate in incidents resolutions, give developers helping hand in debugging applications, networks, databases, compute systems.

Your Skills
Software Engineering:

You excel at developing and maintaining large, established software systems beyond simple scripts and utilities. You definitely know what makes software maintainable and you are able to write robust code.

Firmly Grounded Computer Science Fundamentals:

Including data structures, concurrency, architecture, APIs, testing, and design patterns.

System Engineering Fundamentals:

You most likely know how to deploy and use memory or stack sampling profiler, how to locate excessive lock contention, how to identify network issues, etc.

SRE Expertise:

Hands-on experience with modern SRE practices and tooling – for example, containerization (Docker/Kubernetes), infrastructure‑as‑code (Terraform), and GitOps workflows (ArgoCD or equivalent). You have designed, built, and maintained scalable infrastructure and CI/CD systems.

Cloud & Systems Knowledge:

Deep familiarity with at least one major cloud platform and Linux operating system. You can tune servers, manage databases/storage, and wrangle Kubernetes clusters.

Ownership & Leadership:

High degree of ownership and bias for action, with a proactive approach to solving problems. You take initiative and don't wait to be told what to do. You have demonstrated leadership through mentoring junior engineers or leading small teams/projects, even if not formally a manager. We're seeking a track record of ownership over critical systems and successful delivery of complex projects.

Collaboration & Communication:

Excellent communication skills (written and verbal) and a collaborative attitude. You can work across teams and departments – from explaining technical issues to non-technical colleagues, to coordinating with engineers on deployments. You value teamwork and knowledge sharing.

Adaptability:

Willingness to wear multiple hats and adapt to evolving needs. In a fast-growing startup environment, requirements can change – you're excited by the chance to learn new skills, take on new challenges, and grow with the role.

Bonus Points:

Experience in a high-growth startup environment, which means you're used to the pace and ambiguity.

Any prior experience maintaining security compliance and certifications in a company is a plus.

If you have used specific tools we use (GCP, ArgoCD, GitLab CI, Kafka, etc.), that's great – if not, you can learn quickly.

If you have significant experience running production workloads over Apache Cassandra and / or Postgres database

If you developed software in Rust programming language and can mentor other developers on the best practices in Rust.

Please note: if you don't have all the skills/experience listed above but believe you could be outstanding in this role, please still consider applying. Many folks, especially those from underrepresented or marginalised groups, often count themselves out. Please allow us to learn more about you and why you're exceptional!

Healthy, Catered Lunches

Enjoy fresh, healthy lunches every workday in our Auckland, Christchurch, London and San Francisco offices. With no meal prep needed, you can eat, connect, and refuel with your team. (And yes, snacks and drinks are always on hand.)

Healthy Body, Healthy Mind

We care about performing at our peak. Every team member gets a $1,500 annual wellness allowance (or local equivalent) on a Partly-branded card. Use it on things such gym memberships, rock climbing, physio, massage, GP visits, prescriptions; anything that you or your family, need!

Family Comes First

Primary caregivers receive 3 months of fully paid parental leave, plus a flexible return-to-work (four days on full pay for your first three months back).

Getting Here Is On Us

If you commute to a Partly office or co-working space, choose from a paid 24/7 car park or commute allowance. One less thing to think about!

Workspaces That Inspire

Our brand new, architecturally designed offices are built for collaboration and creativity, with great coffee, social spaces, and some of the best cafes a few steps away.

Office-First with Flexibility

In cities where we have an office (Christchurch, Auckland, London, San Francisco), we default there every day. This let's us move faster, make better decisions and build strong relationships. We also operate with a very high trust environment, so you can manage your time around your life, and flex your schedule to get your best work done.

We Celebrate Together

From weekly happy hours and monthly lunches to quarterly season openers and an annual global offsite, we make time to connect, celebrate, and have fun as one team.

Relocation

If you are relocating from overseas or domestically to Partly HQ, we offer a generous relocation allowance to support your move

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer, ANZ
Senior Site Reliability Engineer, ANZ

Partly • Christchurch

On-site
NZD 120,000 - 180,000
Healthy Lunches
Wellness Allowance
Parental Leave
+2
Senior Site Reliability Engineer, ANZ
Senior Site Reliability Engineer, ANZ

Partly • Auckland

On-site
NZD 120,000 - 170,000
Time off
Flat structure
Employee Experience Team
+6
Platform Security Engineer, NZ
Platform Security Engineer, NZ

JobSpace • Christchurch

On-site
NZD 140,000 - 190,000
Healthy lunches
Wellness allowance
Parental leave
+2
Platform Security Engineer
Platform Security Engineer

Partly Group • Christchurch, Auckland

Hybrid
NZD 90,000 - 130,000
Healthy, catered lunches
Annual wellness allowance
Paid parental leave
+4
Senior Software Engineer, NZ
Senior Software Engineer, NZ

Partly • Christchurch

On-site
NZD 120,000 - 180,000
Healthy lunches
Wellness allowance
Parental leave
+4
Senior Software Engineer, NZ
Senior Software Engineer, NZ

JobSpace • Auckland

On-site
NZD 120,000 - 180,000
Healthy lunches
Wellness allowance
Parental leave (3 months)
+3
Senior Software Engineer, NZ
Senior Software Engineer, NZ

Visa Hunt • Christchurch

Hybrid
NZD 110,000 - 150,000
Healthy lunches
Wellness allowance
Parental leave
+2
Senior Software Engineer, NZ
Senior Software Engineer, NZ

Partly • Auckland

On-site
NZD 110,000 - 140,000
Healthy, Catered Lunches
Annual wellness allowance
Paid parental leave
+4
Engineering Team Lead, NZ
Engineering Team Lead, NZ

Visa Hunt • Christchurch

Hybrid
NZD 120,000 - 180,000
Healthy, Catered Lunches
Healthy Body, Healthy Mind
Family Comes First
+4
Head of Platform
Head of Platform

Partly Group • Christchurch, Auckland

Hybrid
NZD 210,000 - 320,000
Flexible working hours
Offices in Christchurch CBD
Offices in Auckland
+2