Senior SRE: Cloud Native Reliability & Kubernetes

Cisco Systems, Inc.

San Francisco (CA)

Hybrid

USD 180,000 - 240,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Cisco ThousandEyes is seeking a Senior Site Reliability Engineer to design and operate large‑scale, highly available distributed systems in the cloud. The role collaborates with application teams to boost reliability, performance, and security of the platform.

You will drive 24x7 incident response, on‑call rotations, and automation across Kubernetes, Prometheus, OpenTelemetry, and ArgoCD. Hybrid work with weekly in‑office attendance in SF, Seattle, Austin or New York is offered.

Qualifications

  • 5+ years of experience in a related role.
  • Proficiency in Python or Go.
  • Ability to build scalable, secure, well-tested solutions across dev and prod lifecycles.
  • Strong understanding of Unix/Linux systems and client‑server protocols.
  • Knowledge of Site Reliability principles: Incident Response, Change Management, Distributed Systems, Deployment Strategies, and SLOs.

Responsibilities

  • Collaborate with software engineers to optimize architecture for availability and latency.
  • Design and implement scalable operations tooling for multi‑region growth.
  • Deploy and maintain AWS cloud‑native services that are elastic and fault‑tolerant.
  • Participate in 24x7 incident response and on‑call rotation.
  • Leverage Kubernetes, Service Mesh, Prometheus, OpenTelemetry, and ArgoCD to improve system reliability.
  • Automate production operations and guardrails; enable scalable service deployment and testing.
  • Develop automation for scalable platform operations including chaos testing and scale testing.
  • Stay updated on industry best practices for scalability and reliability across the ThousandEyes platform.
  • Identify obstacles hindering operations excellence and provide solutions across engineering teams.
  • Generalize and standardize processes for repeatable success across microservices.

Skills

Python
Go
Unix/Linux
SRE Principles

Tools

Kubernetes
AWS
OpenTelemetry
Prometheus
ArgoCD

Job description

Cisco ThousandEyes is seeking a Senior Site Reliability Engineer to design and operate large‑scale, highly available distributed systems in the cloud. The role collaborates with application teams to boost reliability, performance, and security of the platform.

You will drive 24x7 incident response, on‑call rotations, and automation across Kubernetes, Prometheus, OpenTelemetry, and ArgoCD. Hybrid work with weekly in‑office attendance in SF, Seattle, Austin or New York is offered.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote Senior Network Reliability Engineer (SRE)
Remote Senior Network Reliability Engineer (SRE)

Gainbridge • Zionsville (IN), Northern (KY)

On-site
USD 135,000 - 190,000
Health Insurance
Dental Insurance
Vision Insurance
+4
Cloud-Native SRE Lead: Automation, Kubernetes & Resilience
Cloud-Native SRE Lead: Automation, Kubernetes & Resilience

Axiom Pursuits • San Francisco (CA)

On-site
USD 150,000 - 180,000
Senior SRE: Cloud-Native Reliability & Automation
Senior SRE: Cloud-Native Reliability & Automation

Crypto Pro Network • New York (NY)

On-site
USD 150,000 - 180,000
Senior Cloud SRE — Kubernetes & Reliability Lead
Senior Cloud SRE — Kubernetes & Reliability Lead

TP-Link Systems Inc. • Irvine (CA)

On-site
USD 140,000 - 180,000
Free snacks and drinks
Fully paid medical, dental, and vision insurance
401k contributions
+3
Senior Site Reliability Engineer – Cloud, Kubernetes & Automation
Senior Site Reliability Engineer – Cloud, Kubernetes & Automation

Socure • United States

On-site
USD 160,000 - 180,000
Senior SRE & DevOps Engineer | Cloud + Kubernetes
Senior SRE & DevOps Engineer | Cloud + Kubernetes

TechDigital Group • Framingham (MA)

On-site
USD 110,000 - 150,000
Senior SRE Engineer – Cloud, Kubernetes & Automation
Senior SRE Engineer – Cloud, Kubernetes & Automation

VBeyond Corporation • Dallas (TX)

On-site
USD 100,000 - 140,000
Senior SRE: Cloud, Kubernetes & 24x7 Reliability Lead
Senior SRE: Cloud, Kubernetes & 24x7 Reliability Lead

NTT DATA, Inc. • Baltimore (MD)

On-site
USD 88,000 - 110,000
Remote Senior Staff SRE — Kubernetes, CI/CD & Cloud Scale
Remote Senior Staff SRE — Kubernetes, CI/CD & Cloud Scale

Far Coder • Northern (KY)

Hybrid
USD 170,000 - 227,000
Senior SRE – AI Cloud Platform, Kubernetes Expert
Senior SRE – AI Cloud Platform, Kubernetes Expert

Socket.dev • San Francisco (CA)

On-site
USD 180,000 - 240,000
Health, dental, vision coverage for in
Wellness and commuter stipends
401k with 2% company match
+1