Senior SRE - AI Platform Reliability (Hybrid)

Cisco

New York (NY)

Hybrid

USD 187,000 - 308,000

Full time

13 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Medical insurance
Dental insurance
Vision insurance
401(k) with company match
Paid parental leave
Disability coverage
Life insurance
Vacation + holidays

Job summary

Cisco is seeking a Staff Site Reliability Engineer to provide technical leadership for the reliability, scalability, and operational architecture of Splunk Agent Observability's platform. This hybrid role focuses on deploying and scaling cloud and on-prem environments with strong emphasis on automation and best practices.

You will mentor engineers, define reliability roadmaps, and partner with leadership to steer platform architecture and production readiness in a complex, distributed system

Qualifications

  • 8+ years’ experience in Site Reliability Engineering or related fields.
  • 5+ years operating large-scale Kubernetes platforms in production.
  • Experience with AWS, GCP, or other public clouds.
  • Strong experience designing CI/CD platforms and deployment automation at scale.

Responsibilities

  • Define and drive the technical roadmap for platform reliability, scalability, and operational excellence.
  • Lead architecture and evolution of deployment platforms for cloud/on-prem environments.
  • Establish SRE standards including SLOs, capacity planning, and resiliency reviews.
  • Lead reliability initiatives across Kubernetes, deployment infra, databases, and networking.
  • Drive automation to reduce toil and improve productivity.
  • Design and build internal platforms and tooling for reliable operations at scale.
  • Lead incident response and drive long-term remediation and root-causes.

Skills

Kubernetes
Cloud platforms
CI/CD
Python/Go
Observability
Infrastructure as Code

Education

Bachelor's degree in CS/related
Master's degree preferred

Tools

Terraform

Job description

Cisco is seeking a Staff Site Reliability Engineer to provide technical leadership for the reliability, scalability, and operational architecture of Splunk Agent Observability's platform. This hybrid role focuses on deploying and scaling cloud and on-prem environments with strong emphasis on automation and best practices.

You will mentor engineers, define reliability roadmaps, and partner with leadership to steer platform architecture and production readiness in a complex, distributed system

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior SRE (Hybrid) - AI Resilience Platform
Senior SRE (Hybrid) - AI Resilience Platform

020 Cisco Systems, Inc. • San Jose (CA)

Hybrid
USD 168,000 - 245,000
Senior SRE: AI Resilience & Platform Reliability Lead
Senior SRE: AI Resilience & Platform Reliability Lead

Cisco • San Jose (CA)

Hybrid
USD 170,000 - 308,000
Senior SRE - AI Resilience & Cloud Platform
Senior SRE - AI Resilience & Cloud Platform

Cisco • San Francisco (CA)

Hybrid
USD 187,000 - 268,000
Health insurance
401(k) with Cisco match
Paid parental leave
+2
Senior SRE - AI Resilience & Kubernetes Platform
Senior SRE - AI Resilience & Kubernetes Platform

Cisco • San Francisco (CA)

Hybrid
USD 168,000 - 245,000
Medical insurance
401(k) with company match
Paid parental leave
+3
Senior AI Reliability Engineer
Senior AI Reliability Engineer

Cisco • New York (NY)

Hybrid
USD 149,000 - 282,000
Medical, dental, vision insurance
401(k) with Cisco matching
Paid parental leave
+1
Senior SRE — AI Resilience & Cloud Platforms
Senior SRE — AI Resilience & Cloud Platforms

Cisco • San Jose (CA)

Hybrid
USD 168,000 - 245,000
Senior SRE: Observability, Splunk & Automation (Hybrid)
Senior SRE: Observability, Splunk & Automation (Hybrid)

ISO New England Inc. • Holyoke (MA)

Hybrid
USD 134,000 - 170,000
Hybrid work environment (3 days/week)
Senior SRE - Hybrid, Observability & Reliability
Senior SRE - Hybrid, Observability & Reliability

Early Warning Services LLC • Chicago (IL)

Hybrid
USD 106,000 - 130,000
Healthcare Coverage
401(k) Plan with match
PTO and Holidays
+1
Senior SRE: AI-Driven, Cloud-Native Reliability (Hybrid)
Senior SRE: AI-Driven, Cloud-Native Reliability (Hybrid)

OutSystems • San Francisco (CA)

Hybrid
USD 140,000 - 210,000
Hybrid work model
Senior SRE - Hybrid, AWS & Observability
Senior SRE - Hybrid, AWS & Observability

Early Warning • San Francisco (CA)

Hybrid
USD 128,000 - 156,000
Healthcare coverage
401(k) plan
Paid time off
+2