Senior SRE: AI Resilience & Platform Reliability Lead

Cisco

San Jose (CA)

Hybrid

USD 170,000 - 308,000

Full time

22 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Cisco is hiring a Staff Site Reliability Engineer to lead reliability, scalability, and operational architecture for the Splunk Agent Observability platform. You will shape long-term reliability strategy, drive major infrastructure initiatives, and improve deployment automation across cloud and on-prem environments.

You will mentor engineers, set best practices, and collaborate with customers and internal teams to build secure, scalable deployment architectures.

Qualifications

  • 8+ years’ experience in SRE, platform/cloud infra, or related fields.
  • 5+ years’ operating large-scale Kubernetes in production.
  • Experience with AWS, GCP, or other public clouds.
  • Strong skills in Python and/or Go for tooling and automation.

Responsibilities

  • Define and drive the reliability roadmap for platform services.
  • Lead architecture for deployment platforms across cloud and air-gapped environments.
  • Establish SLOs, capacity planning, and resiliency reviews.

Skills

Kubernetes
AWS
GCP
Python
Go
CI/CD
Infrastructure as Code
Observability

Education

Bachelor's degree
Master's degree
PhD

Tools

Terraform
CI/CD tools
Monitoring tools

Job description

Cisco is hiring a Staff Site Reliability Engineer to lead reliability, scalability, and operational architecture for the Splunk Agent Observability platform. You will shape long-term reliability strategy, drive major infrastructure initiatives, and improve deployment automation across cloud and on-prem environments.

You will mentor engineers, set best practices, and collaborate with customers and internal teams to build secure, scalable deployment architectures.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior SRE - AI Platform Reliability (Hybrid)
Senior SRE - AI Platform Reliability (Hybrid)

Cisco • New York (NY)

Hybrid
USD 187,000 - 308,000
Medical insurance
Dental insurance
Vision insurance
+5
Senior SRE - AI Resilience & Cloud Platform
Senior SRE - AI Resilience & Cloud Platform

Cisco • San Francisco (CA)

Hybrid
USD 187,000 - 268,000
Health insurance
401(k) with Cisco match
Paid parental leave
+2
Senior SRE (Hybrid) - AI Resilience Platform
Senior SRE (Hybrid) - AI Resilience Platform

020 Cisco Systems, Inc. • San Jose (CA)

Hybrid
USD 168,000 - 245,000
Senior SRE - AI Resilience & Kubernetes Platform
Senior SRE - AI Resilience & Kubernetes Platform

Cisco • San Francisco (CA)

Hybrid
USD 168,000 - 245,000
Medical insurance
401(k) with company match
Paid parental leave
+3
Senior AI Reliability Engineer
Senior AI Reliability Engineer

Cisco • New York (NY)

Hybrid
USD 149,000 - 282,000
Medical, dental, vision insurance
401(k) with Cisco matching
Paid parental leave
+1
Senior SRE — AI Resilience & Cloud Platforms
Senior SRE — AI Resilience & Cloud Platforms

Cisco • San Jose (CA)

Hybrid
USD 168,000 - 245,000
Senior SRE: Observability, Splunk & Automation (Hybrid)
Senior SRE: Observability, Splunk & Automation (Hybrid)

ISO New England Inc. • Holyoke (MA)

Hybrid
USD 134,000 - 170,000
Hybrid work environment (3 days/week)
Senior SRE: AI-Driven Cloud Reliability & Automation
Senior SRE: AI-Driven Cloud Reliability & Automation

Hidden Jobs • United States

Remote
USD 191,000 - 226,000
Equity incentive
Flexible PTO
Health insurance
+2
Senior SRE — AI-Driven Reliability & Oncall Leadership
Senior SRE — AI-Driven Reliability & Oncall Leadership

Block • San Francisco (CA)

On-site
USD 160,700 - 283,600
Healthcare coverage
Health Savings Account
Retirement Plans
+5
Senior SRE — Flexible, AI-Driven Reliability
Senior SRE — Flexible, AI-Driven Reliability

Salesforce, Inc. • San Francisco (CA)

Hybrid
USD 148,000 - 224,000