Staff Reliability Engineer

ServiceNow

California (MO)

On-site

USD 167,000 - 291,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Health plans
401(k) Plan with company match
ESPP
Matching donations
Flexible time away plan
Family leave programs

Job summary

ServiceNow is seeking a senior SRE/Platform Engineer to design and operate cloud-native platforms for software validation, release validation, and production readiness. You will build automated test pipelines, observability signals, and quality gates integrated into CI/CD workflows.

Ideal candidates have 8+ years in SRE/DevOps or related fields, with hands-on Kubernetes expertise, and strong software engineering skills across Python, Go, Java, or Ruby.

Qualifications

  • 8+ years of experience in SRE/DevOps/Platform/Infra engineering
  • Hands-on Kubernetes across cluster operations
  • Experience building cloud-native platforms
  • CI/CD integration with Kubernetes & cloud-native workflows
  • Automation to improve productivity & release quality
  • Progressive delivery, canary, feature flags, automated rollback
  • Chaos engineering and resilience testing
  • Strong software engineering skills (Python/Go/Java/Ruby)
  • AI-assisted engineering experience is a plus
  • Observability, SLOs/SLIs, incident management
  • Ability to lead projects and collaborate across teams

Responsibilities

  • Design, build, and operate cloud-native platforms for validation and production readiness
  • Create automated test pipelines, observability, and quality gates in CI/CD
  • Develop deployment intelligence and reliability signals
  • Collaborate with engineering teams to improve release quality and efficiency

Skills

Python
Go
Java
Ruby
Automation
Observability
AI-assisted engineering
Distributed systems

Education

Bachelor's degree in Computer Science or related field
Master's degree
PhD (with relevant experience)

Tools

Kubernetes
CI/CD
GitOps

Job description

Role Overview

Join us to build the next generation of cloud-native reliability, release, and test platforms that enable engineering excellence, developer productivity, and high-confidence ServiceNow releases through automation, observability, and AI-driven operations.

What You Will Do

Design, build, and operate cloud-native engineering platforms for software validation, release validation, and production readiness. Build and integrate automated test pipelines, observability, reliability signals, deployment intelligence, and quality gates into CI/CD workflows.

Why It Might Be a Fit

To be successful in this role you have experience in leveraging or critically thinking about how to integrate AI into work processes, decision-making, or problem-solving. You have 8+ years of experience in Site Reliability Engineering (SRE), DevOps, Platform Engineering, Software Engineering, or Infrastructure Engineering with a Bachelor's degree; or 6 years and a Master's degree; or a PhD with 3 years experience; or equivalent experience.

Requirements
  • 8+ years of experience in Site Reliability Engineering (SRE), DevOps, Platform Engineering, Software Engineering, or Infrastructure Engineering
  • Hands-on experience with Kubernetes across cluster operations, networking, storage, security, autoscaling, and multi-cluster environments
  • Experience building and operating cloud-native platforms supporting scalable, highly available services
  • Experience integrating Kubernetes with CI/CD, GitOps, automated test pipelines, deployment validation, and cloud-native deployment workflows
  • Experience designing and implementing automation to improve developer productivity, release quality, and operational efficiency
  • Experience with progressive delivery practices, including canary deployments, feature flags, automated rollback, and deployment verification
  • Experience with chaos engineering, resilience testing, disaster recovery, and reliability validation
  • Strong software engineering skills with hands-on experience designing, developing, testing, and debugging applications using Python, Go, Java, or Ruby
  • Experience leveraging AI-assisted engineering for intelligent testing, release risk analysis, incident diagnostics, or operational automation is a plus
  • Strong understanding of observability, monitoring, SLI/SLOs, incident management, and production operations for distributed systems
  • Demonstrated ability to solve complex technical problems, drive projects independently, and collaborate effectively across engineering teams
Benefits
  • base pay of $166,500 - $291,400
  • equity
  • variable/incentive compensation
  • health plans, including flexible spending accounts
  • 401(k) Plan with company match
  • ESPP
  • matching donations
  • flexible time away plan
  • family leave programs
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Staff Software Engineer – SRE, Release & Test Platforms
Senior Staff Software Engineer – SRE, Release & Test Platforms

ServiceNow • California (MO)

On-site
USD 180,000 - 270,000
Health plans
401(k) Plan with company match
ESPP
+3
Staff Reliability Engineer
Staff Reliability Engineer

ServiceNow • Kirkland (WA)

On-site
USD 167,000 - 291,000
Health plans
401(k) with company match
ESPP
+3
Site Reliability Engineering (SRE)
Site Reliability Engineering (SRE)

Weekday (YC W21) • New York (NY)

On-site
USD 150,000 - 250,000
Equity or bonus opportunities
Health benefits
Paid time off
+2
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • North Carolina

On-site
USD 165,000 - 215,000
Pre‑IPO Stock Options
Medical, Dental & Vision care
401(k)
+2
Senior Cloud Reliability Engineer: AI-Driven Platforms
Senior Cloud Reliability Engineer: AI-Driven Platforms

ServiceNow • California (MO)

On-site
USD 167,000 - 291,000
Equity
Health plans
401(k) Plan with company match
+4
Staff Reliability Engineer
Staff Reliability Engineer

ServiceNow • Minnesota

On-site
USD 160,000 - 260,000
Senior Staff Platform Engineer — AI-Driven Reliability
Senior Staff Platform Engineer — AI-Driven Reliability

ServiceNow • California (MO)

On-site
USD 180,000 - 270,000
Health plans
401(k) Plan with company match
ESPP
+3
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • New York (NY)

Hybrid
USD 165,000 - 215,000
Pre-IPO Stock Options
Medical, Dental & Vision care
401(k)
+1
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • New Jersey

On-site
USD 165,000 - 215,000
Pre-IPO Stock Options
Medical, Dental & Vision care
401(k)
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Supio • San Francisco (CA)

On-site
USD 170,000 - 220,000