Stand out for this role — generate a tailored resume and cover letter in about a minute.
Cisco is seeking an Observability-focused Site Reliability Engineer to design and operate enterprise-grade monitoring and tracing platforms for our cloud infrastructure in Boise. The role emphasizes building and maintaining Splunk, Elasticsearch, Grafana, Tempo, and OpenTelemetry pipelines.
You will collaborate across platform, application, security, and network teams, participate in on-call rotations, and drive reliability improvements across Cisco's cloud environment.
The application window is expected to close on: 12/17/2026
Join the Observability team within Cloud Network Platform Engineering's Site Reliability Engineering organization. You will help design, build, and operate enterprise observability platforms for logging, metrics, tracing, and alerting across large-scale cloud infrastructure.
This role will contribute to customer-facing production services, with an initial focus on Kubernetes-based platform workloads such as Nextunnel and broader SRE and reliability initiatives over time. You will lead efforts that improve visibility, reduce time to detect and resolve incidents, increase platform reliability, and strengthen operational excellence across Cisco's cloud environment.
The role includes participation in a shared SRE pager and on-call rotation, production support, and incident response.