Senior Platform Reliability Engineer - Observability & Cloud

Zoom

San Jose (CA)

Hybrid

USD 99,000 - 229,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Zoom is seeking a hands-on Platform Engineer to own the reliability, scalability, and operational excellence of government-facing products. You will own the observability platform, work across Kubernetes, Terraform, and cloud infra, and respond to incidents with urgency, delivering permanent improvements.

You will design Terraform modules, runbooks, and SLOs; operate production Kubernetes workloads at scale; collaborate to reduce toil and raise resilience across teams.

Qualifications

  • US citizenship or lawful permanent resident status.
  • 4–5+ years of hands-on Platform Engineering, SRE, DevOps, or Production Engineering in live production.
  • Production-grade Kubernetes experience at scale.
  • Authors Terraform modules and providers from scratch.
  • Owns incidents end-to-end with on-call, triage, and root-cause analysis.
  • Deploys or improves observability/monitoring with Python or Bash.
  • Clear written communication; runbooks and incident reports.

Responsibilities

  • Architect end-to-end deployment lifecycle of the observability platform across cloud providers.
  • Design and build Terraform modules and providers; automate provisioning and config management.
  • Operate and improve production Kubernetes workloads; lead incident response and permanent fixes.
  • Develop reusable runbooks and standards for observability onboarding and alerting.
  • Collaborate to define and track SLOs/SLIs; reduce toil via automation.

Skills

Kubernetes operations
On-call incident experience
Terraform modules & providers
Python / Bash scripting
Technical writing / runbooks

Tools

Datadog
Oracle OCI
AWS GovCloud
GitLab CI/CD

Job description

Zoom is seeking a hands-on Platform Engineer to own the reliability, scalability, and operational excellence of government-facing products. You will own the observability platform, work across Kubernetes, Terraform, and cloud infra, and respond to incidents with urgency, delivering permanent improvements.

You will design Terraform modules, runbooks, and SLOs; operate production Kubernetes workloads at scale; collaborate to reduce toil and raise resilience across teams.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Platform Reliability Engineer - Kubernetes & Observability
Senior Platform Reliability Engineer - Kubernetes & Observability

PLP Group • New York (NY)

On-site
USD 75,000 - 130,000
Senior Platform Engineer — Cloud, CI/CD & Observability
Senior Platform Engineer — Cloud, CI/CD & Observability

S27a • Los Altos (CA), Northern (KY)

Hybrid
USD 150,000 - 210,000
Remote-friendly environment
Competitive compensation
Real-Time Platform Reliability Engineer
Real-Time Platform Reliability Engineer

Socket.dev • San Jose (CA)

Hybrid
USD 99,000 - 229,000
Senior Platform Engineer – Observability & Cloud Infra (Hybrid)
Senior Platform Engineer – Observability & Cloud Infra (Hybrid)

CJ • Santa Barbara (CA)

Hybrid
USD 112,000 - 172,000
401K matching
Wellness programs
Comprehensive medical, dental, vision
+2
Senior SRE: Cloud, Kubernetes & Observability Leader
Senior SRE: Cloud, Kubernetes & Observability Leader

L'Oréal • United States

On-site
USD 150,000 - 230,000
Remote Platform Reliability Engineer
Remote Platform Reliability Engineer

Redwood Logistics LLC • Northern (KY)

Hybrid
USD 160,000 - 175,000
Health insurance
401k with match
Paid time off
+1
Senior Platform Engineer: Observability & Kubernetes
Senior Platform Engineer: Observability & Kubernetes

Publicis Groupe Holdings B.V • Agoura Hills (CA)

Hybrid
USD 110,000 - 167,000
401K matching
Medical, dental, and vision coverage
Hybrid work arrangements
+1
Principal DevOps Architect - Cloud & Colocation Reliability
Principal DevOps Architect - Cloud & Colocation Reliability

Artha Nexgen • San Jose (CA)

Hybrid
USD 147,000 - 339,000
Senior Observability Platform Engineer – Scale & Reliability
Senior Observability Platform Engineer – Scale & Reliability

CVS Health • Tennessee

Hybrid
USD 83,000 - 222,000
Real-Time Cloud SRE & Reliability Engineer
Real-Time Cloud SRE & Reliability Engineer

Zoom • San Jose (CA)

Hybrid
USD 99,000 - 229,000