Senior Staff SRE & AIOps Engineer

ServiceNow

California (MO)

Hybrid

USD 191,000 - 334,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

ServiceNow seeks a Senior Staff Reliability Engineer – SRE & AIOps to drive automation-first infrastructure across hybrid cloud and data center operations. You will architect tooling, auto-remediation, and patterns for reliable, scalable cloud platforms with follow-the-sun teams.

Role combines deep Kubernetes, cloud, and DevOps expertise with leadership to reduce toil, improve MTTR, and guide engineers in advanced reliability practices. Base pay ranges are disclosed for the location.

Qualifications

  • Hands-on experience designing and operating production Kubernetes clusters at scale.
  • Experience building closed-loop auto-remediation and self-healing systems.
  • Proficiency with AWS, Azure, and GCP for multi-region cloud deployments.
  • Strong IaC and GitOps practices across hybrid/multi-cloud environments.
  • Familiarity with observability, incident management, and log aggregation.

Responsibilities

  • Design, deploy, and operate enterprise-scale Kubernetes clusters across hybrid and multi-cloud environments with high availability.
  • Architect and implement auto-remediation systems to reduce MTTR and toil.
  • Evolve SRE tooling for monitoring, incident management, and observability.
  • Define SLOs, error budgets, and automated runbooks for on-call engineers.
  • Develop Infrastructure-as-Code frameworks and GitOps pipelines for reproducible deployments.
  • Manage hybrid cloud and data center operations including DR and cost optimization.
  • Drive CI/CD, service mesh, and network security patterns for rapid releases.
  • Create scalable on-call schedules and incident response processes across time zones.
  • Mentor junior engineers and promote blameless post-incident learning.

Skills

Kubernetes Mastery
Incident Auto-Remediation
Cloud Platform Experience
DevOps & IaC
SRE Tooling Fluency
Distributed Systems
On-Call Operations
Hybrid Cloud Operations
AI/ML in Ops
Influence Without Authority

Education

Bachelor's degree in CS/CE or related

Tools

Kubernetes
Terraform
GitOps
CI/CD
Observability Platforms

Job description

ServiceNow seeks a Senior Staff Reliability Engineer – SRE & AIOps to drive automation-first infrastructure across hybrid cloud and data center operations. You will architect tooling, auto-remediation, and patterns for reliable, scalable cloud platforms with follow-the-sun teams.

Role combines deep Kubernetes, cloud, and DevOps expertise with leadership to reduce toil, improve MTTR, and guide engineers in advanced reliability practices. Base pay ranges are disclosed for the location.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Staff Software Engineer – SRE & AIOps
Senior Staff Software Engineer – SRE & AIOps

ServiceNow • California (MO)

Hybrid
USD 191,000 - 334,000
Head of SRE & Service Enablement (AI-Driven)
Head of SRE & Service Enablement (AI-Driven)

ServiceNow • California (MO)

Hybrid
USD 221,000 - 387,000
Health plans
401(k) Plan with company match
ESPP
+3
Director of Site Reliability Engineering & Service Enablement
Director of Site Reliability Engineering & Service Enablement

ServiceNow • Santa Clara (CA)

On-site
USD 260,000 - 360,000
Generous family leave
Matched donations
Annual learning stipends
+3
Staff Site Reliability Engineer – Multi-Cloud Infra & Automation
Staff Site Reliability Engineer – Multi-Cloud Infra & Automation

ServiceNow • Minnesota

Hybrid
USD 115,000 - 195,000
Health plans
401(k) with company match
ESPP
+3
Staff Data Platform Engineer – Kubernetes & AI-Driven Cloud
Staff Data Platform Engineer – Kubernetes & AI-Driven Cloud

ServiceNow • California (MO)

Hybrid
USD 150,000 - 262,000
Health plans
401(k) Plan with company match
ESPP (Employee Stock Purchase Plan)
+3
Senior Cloud SRE: Reliability, Automation & Remote Work
Senior Cloud SRE: Reliability, Automation & Remote Work

Embedded Shishya • United States

Remote
USD 120,000 - 180,000
Base salary and permanent contract
Paid certifications
Career plan
+9
Senior Production SRE: Cloud & On-Prem Reliability
Senior Production SRE: Cloud & On-Prem Reliability

Weights & Biases • New York (NY)

On-site
USD 140,000 - 180,000
Medical Insurance
Dental Insurance
Vision Insurance
+15
Sr SRE Automation Engineer
Sr SRE Automation Engineer

Compunnel, Inc. • Austin (TX), Northern (KY)

On-site
USD 130,000 - 180,000
Director of Reliability & Service Enablement
Director of Reliability & Service Enablement

ServiceNow • Santa Clara (CA)

On-site
USD 260,000 - 360,000
Generous family leave
Matched donations
Annual learning stipends
+3
Lead Site Reliability Engineer (SRE) / Principal Site Reliability Engineer (SRE)
Lead Site Reliability Engineer (SRE) / Principal Site Reliability Engineer (SRE)

Mindlance • Irving (TX)

On-site
USD 120,000 - 160,000