Senior Site Reliability Engineer

Clera

United States

Remote

USD 150,000 - 210,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Clera is seeking a Senior Site Reliability Engineer to own cloud infrastructure and elevate DevOps and reliability practices. You will evolve our GCP infrastructure and maturity of the observability platform, driving incident management and partnering with Product & Engineering to ship reliable software.

You will champion SLOs, SLIs, and error budgets, and lead on-call rotations within a blameless culture while helping teams govern usage at scale across a distributed environment.

Qualifications

  • 3+ years of Site Reliability Engineering or production SRE experience.
  • Strong proficiency with Google Cloud Platform (GCP), including cost optimisation and governance.

Responsibilities

  • Design, evolve, and scale our cloud infrastructure on GCP.
  • Build tooling and automation that promote team autonomy and reduce toil.
  • Advance our observability platform, improving mean time to recovery (MTTR) and system visibility.
  • Build transparency into infrastructure costs and drive cost optimisation initiatives.
  • Champion reliability best practices including SLOs/SLIs, error budgets, and post-incident reviews.
  • Lead on-call rotations and incident management, fostering a blameless culture.
  • Help engineering teams leverage GCP effectively and govern usage at scale.

Skills

GCP
SRE experience
Python
Bash
Go
Observability
SLOs/SLIs
Incident management
On-call rotations
Cost governance

Tools

Kubernetes
Terraform
Deployment Manager
Prometheus
Grafana
OpenTelemetry

Job description

About the Role

We are a well-funded AI/ML company operating at the intersection of geospatial intelligence and infrastructure analytics. Our engineering team is distributed across Europe and North America, and we're looking for a Senior Site Reliability Engineer to take ownership of our cloud infrastructure and elevate our DevOps and reliability practices.

In this role, you'll evolve our Google Cloud Platform (GCP) infrastructure, mature our observability platform, drive incident management processes, and partner closely with Product & Engineering teams to ship reliable, high-quality software. You'll be a key voice in championing SLOs, error budgets, and DORA metrics across the organisation.

What You'll Do
  • Design, evolve, and scale our cloud infrastructure on GCP.

  • Build tooling and automation that promote team autonomy and reduce toil.

  • Advance our observability platform, improving mean time to recovery (MTTR) and system visibility.

  • Build transparency into infrastructure costs and drive cost optimisation initiatives.

  • Champion reliability best practices including SLOs/SLIs, error budgets, and post-incident reviews.

  • Lead on-call rotations and incident management, fostering a blameless culture.

  • Help engineering teams leverage GCP effectively and govern usage at scale.

What We're Looking For

Required:

  • 3+ years of Site Reliability Engineering or production SRE experience.

  • Strong proficiency with Google Cloud Platform (GCP), including cost optimisation and governance.

Nice to Have / Additional Skills:

  • Hands-on experience with Kubernetes for cluster and workload management.

  • Infrastructure as Code experience — Terraform, Deployment Manager, or similar.

  • Scripting and automation skills in Python, Bash, or Go.

  • Strong observability stack experience: Prometheus, Grafana, OpenTelemetry, logging, and tracing.

  • Proven ability to define and implement SLOs/SLIs and error budgets.

  • Experience with incident management, post-incident reviews, and on-call rotations.

Location

This is a fully remote role, open to candidates based in the EU, UK, or North America. The primary hub is in the Netherlands. Please note that visa sponsorship is not available for this position.

Compensation & Benefits

Compensation details were not provided for this role. Our team spans multiple countries and we offer competitive, location-adjusted packages. Further details will be discussed during the interview process.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

TalentDome Staffing • United States

On-site
USD 140,000 - 210,000
Staff SRE
Staff SRE

Selby Jennings • Chicago (IL)

On-site
USD 140,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

Harrison Clarke • New York (NY)

On-site
USD 120,000 - 160,000
Site Reliability Engineer
Site Reliability Engineer

Compunnel, Inc. • New Jersey

On-site
USD 120,000 - 150,000
Site Reliability Engineer
Site Reliability Engineer

Amiri Recruiting • Mountain View (CA)

On-site
USD 130,000 - 160,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Jobgether • United States

On-site
USD 150,000 - 200,000
Competitive salary
Comprehensive healthcare coverage
401(k) plan with company matching
+3
Site Reliability Engineer
Site Reliability Engineer

HostPapa, Inc. • United States

Remote
USD 120,000 - 180,000
Remote-first work
Professional development
Flexible hours
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Mission Staffing • New York (NY)

On-site
USD 140,000 - 200,000
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Apply • Tempe (AZ), Northern (KY)

Hybrid
USD 180,000 - 240,000
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Gen Digital Inc. • Tempe (AZ)

On-site
USD 180,000 - 230,000