Senior SRE: AI-Driven Platform Reliability & Scale

Medallia

McLean (VA)

On-site

USD 129,000 - 190,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Health benefits
401(k) matching
Paid parental leave
Paid holidays

Job summary

Medallia is seeking a Senior Site Reliability Engineer to design, operate, and evolve our cloud-native platforms. You will collaborate across teams to boost reliability, scalability, and performance while driving automation and AI-assisted workflows at scale.

The role demands technical leadership, hands-on implementation of IaC, CI/CD, and observability, plus participation in on-call rotations to support production systems. Strong collaboration and strategic thinking are essential.

Qualifications

  • 5+ years of experience leading reliability, platform engineering, infrastructure, or cloud operations initiatives in production environments.

Responsibilities

  • Design, build, and operate highly available, scalable, and secure production platforms.
  • Partner with software engineering teams to improve application reliability, scalability, performance, and operational readiness.
  • Lead complex incident investigations, root cause analyses, and reliability improvement initiatives.
  • Design and implement automation, self-service capabilities, and platform solutions that reduce operational toil.
  • Leverage AI-assisted engineering tools and automation platforms to accelerate troubleshooting, improve productivity, and reduce operational overhead.
  • Identify opportunities to streamline operational processes through automation, AI-enabled workflows, and platform engineering practices.
  • Drive adoption of SRE principles, reliability standards, and operational best practices across engineering organizations.
  • Develop and maintain infrastructure-as-code, deployment automation, and operational tooling.
  • Support and improve CI/CD and GitOps-based deployment workflows.
  • Design observability strategies using monitoring, logging, tracing, and alerting platforms.
  • Participate in architecture reviews and provide guidance on scalability, resiliency, and operational excellence.
  • Mentor junior engineers and contribute to the technical growth of the broader engineering organization.
  • Act as a force multiplier by creating reusable solutions, self-service capabilities, and engineering standards that increase the effectiveness of multiple teams.
  • Drive adoption of AI-assisted engineering workflows and operational automation across the organization.
  • Drive engineering leverage initiatives that improve the productivity, reliability, and effectiveness of multiple engineering teams.
  • Influence the broader engineering organization through platform thinking, standardization, and operational simplification.

Skills

Leadership
Reliability
Platform engineering
Automation
AI-assisted workflows
SRE practices
Incident response
CI/CD
GitOps
Python/Go/Bash
Networking

Tools

Kubernetes
Terraform
ArgoCD
Prometheus
Grafana

Job description

Medallia is seeking a Senior Site Reliability Engineer to design, operate, and evolve our cloud-native platforms. You will collaborate across teams to boost reliability, scalability, and performance while driving automation and AI-assisted workflows at scale.

The role demands technical leadership, hands-on implementation of IaC, CI/CD, and observability, plus participation in on-call rotations to support production systems. Strong collaboration and strategic thinking are essential.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior SRE: Scale Resilient AI Platforms & Automation
Senior SRE: Scale Resilient AI Platforms & Automation

Relx Plc • Philadelphia

Hybrid
USD 95,000 - 159,000
Senior SRE: Reliability, Automation & AI Platforms
Senior SRE: Reliability, Automation & AI Platforms

RX Brasil • Philadelphia

On-site
USD 95,000 - 159,000
Annual incentive bonus
Senior SRE: AI-Driven Reliability & Cloud Automation
Senior SRE: AI-Driven Reliability & Cloud Automation

NDEAVOUR CONSULTING • United States

Hybrid
USD 120,000 - 150,000
Remote Office
Parking Space
Fun Office Space
+7
Senior SRE: AI-Driven Reliability & Automation (Hybrid)
Senior SRE: AI-Driven Reliability & Automation (Hybrid)

Namely • United States

Hybrid
USD 120,000 - 150,000
Senior SRE: Scale & Reliability for AI-Driven SaaS Platform
Senior SRE: Scale & Reliability for AI-Driven SaaS Platform

Instrumental Inc. • Palo Alto (CA)

On-site
USD 175,000 - 229,000
Health benefits
Commuter plans
Parental leave
Senior SRE: AI-Driven, Cloud-Native Reliability (Hybrid)
Senior SRE: AI-Driven, Cloud-Native Reliability (Hybrid)

OutSystems • San Francisco (CA)

Hybrid
USD 140,000 - 210,000
Hybrid work model
Senior SRE Platform Engineer – AI-Powered Reliability
Senior SRE Platform Engineer – AI-Powered Reliability

UiPath • Denver (CO)

Hybrid
USD 160,000 - 210,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Medallia • McLean (VA)

On-site
USD 129,000 - 190,000
Health benefits
401(k) matching
Paid parental leave
+1
Senior SRE, AI Platform: Scale Reliability & Kubernetes
Senior SRE, AI Platform: Scale Reliability & Kubernetes

United States Digital Space LLC • Paris (TX)

On-site
USD 80,000 - 111,000
Staff SRE: AI-Driven Reliability & Platform Architect
Staff SRE: AI-Driven Reliability & Platform Architect

Devopsroles • Northern (KY)

Remote
USD 150,000 - 225,000
Equity
Benefits program