Senior Platform Reliability Engineer — Observability & Cloud

Elasticsearch B.V.

Canada

On-site

CAD 138,000 - 186,000

Full time

6 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Competitive pay
Health coverage
Flexible schedules
Generous vacation days
Donations matching
Volunteer hours
Parental leave

Job summary

Elastic, the Search AI Company, is seeking a Platform Observability & Analytics engineer to own end-to-end delivery of complex projects and harden Elastic Cloud infrastructure. You will write Terraform, Python, and Go, participate in 24/7 on-call rotations, and review production changes to critical systems across multiple time zones.

You will mentor engineers, improve runbooks and documentation, and contribute to security-conscious infrastructure decisions while collaborating with a distributed

Qualifications

  • 5+ years of SRE, platform engineering, or infrastructure engineering experience
  • Proficiency with Terraform; comfortable owning large, multi-workspace configurations in a team setting
  • Strong software engineering fundamentals in Python; comfort with Go is a plus
  • Deep Linux systems knowledge and experience operating containerized workloads in production
  • Experience carrying a 24/7 on-call rotation, resolving incidents under pressure, and writing RCAs that hold up under review
  • A track record of delivering end-to-end projects of moderate-to-high complexity with minimal oversight and being a code/design reviewer
  • Security-conscious mindset when architecting infrastructure
  • Mentoring less experienced engineers and contributing ideas in team discussions
  • Clear written and verbal communication across audiences
  • Ability to work across time zones, in real-time and asynchronously

Responsibilities

  • Owning end-to-end delivery of moderate-to-high complexity projects on the team’s roadmap, with minimal day-to-day direction
  • Operating and hardening shared Elastic Cloud infrastructure (ECH, ECE, and ECK) as Infrastructure as Code - writing and reviewing the Terraform, Python, and Go that other engineers depend on
  • Carrying a 24/7 on-call rotation: responding to incidents, driving them to resolution, and writing clear RCAs/postmortems that lead to lasting fixes
  • Reviewing others' code and designs, and being a trusted second set of eyes on production changes to critical infrastructure
  • Mentoring less experienced engineers, and proactively raising risks, ideas, and improvements in team discussions
  • Improving runbooks, documentation, and operational processes so the on-call load gets lighter over time

Skills

Terraform
Python
Go
Linux
SRE
On-call
Code review
Security mindset
Mentoring
Documentation

Tools

Kubernetes
ArgoCD
Helm
Vault
Teleport
Puppet
Ansible

Job description

Elastic, the Search AI Company, is seeking a Platform Observability & Analytics engineer to own end-to-end delivery of complex projects and harden Elastic Cloud infrastructure. You will write Terraform, Python, and Go, participate in 24/7 on-call rotations, and review production changes to critical systems across multiple time zones.

You will mentor engineers, improve runbooks and documentation, and contribute to security-conscious infrastructure decisions while collaborating with a distributed

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE – Cloud Platform Reliability & On-Call Lead
Senior SRE – Cloud Platform Reliability & On-Call Lead

Engg • Canada

On-site
CAD 138,000 - 186,000
Health coverage
Flexible locations and schedules
Generous vacation days
+1
Platform Engineer
Platform Engineer

LanceSoft, Inc. • Montreal (administrative region)

On-site
CAD 80,000 - 120,000
Observability Engineer - SRE/DevOps (Hybrid)
Observability Engineer - SRE/DevOps (Hybrid)

Okta • Toronto

Hybrid
CAD 110,000 - 152,000
Equity
Bonus
Health insurance
+4
Senior Site Reliability Engineer (Observability & Analytics) – Platform Infra
Senior Site Reliability Engineer (Observability & Analytics) – Platform Infra

Engg • Canada

On-site
CAD 138,000 - 186,000
Health coverage
Flexible locations and schedules
Generous vacation days
+1
Senior Site Reliability Engineer (Observability & Analytics) – Platform Infra
Senior Site Reliability Engineer (Observability & Analytics) – Platform Infra

Elasticsearch B.V. • Canada

Hybrid
CAD 138,000 - 186,000
Competitive pay
Health coverage
Flexible schedules
+4
Software Development Manager, Platform Engineering
Software Development Manager, Platform Engineering

Autodesk • Toronto

On-site
CAD 120,000 - 150,000
Founding Engineer (Agentic Platform)
Founding Engineer (Agentic Platform)

Katalyze AI, Inc. • Toronto

On-site
CAD 130,000 - 200,000
Senior Platform Engineer: Kubernetes & Cloud Infra Lead
Senior Platform Engineer: Kubernetes & Cloud Infra Lead

Okta • Toronto

Hybrid
CAD 136,000 - 187,000
Equity (where applicable)
Bonus
Health insurance
+6
Platform Engineer: Observability & Cloud Telemetry
Platform Engineer: Observability & Cloud Telemetry

LanceSoft, Inc. • Montreal (administrative region)

On-site
CAD 80,000 - 120,000
Senior Java Core/Infra Engineer for Elasticsearch
Senior Java Core/Infra Engineer for Elasticsearch

Elastic • Canada

On-site
CAD 128,000 - 203,000
Health coverage for you and family
Flexible locations & schedules
Generous vacation days
+1