Senior SRE: AI-Driven Platform Reliability & Scale

Medallia

McLean (VA)

On-site

USD 129,000 - 190,000

Full time

12 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health benefits
401(k) matching
Paid parental leave
Paid holidays

Job summary

Medallia is seeking a Senior Site Reliability Engineer to design, operate, and evolve our cloud-native platforms. You will collaborate across teams to boost reliability, scalability, and performance while driving automation and AI-assisted workflows at scale.

The role demands technical leadership, hands-on implementation of IaC, CI/CD, and observability, plus participation in on-call rotations to support production systems. Strong collaboration and strategic thinking are essential.

Qualifications

  • 5+ years of experience leading reliability, platform engineering, infrastructure, or cloud operations initiatives in production environments.

Responsibilities

  • Design, build, and operate highly available, scalable, and secure production platforms.
  • Partner with software engineering teams to improve application reliability, scalability, performance, and operational readiness.
  • Lead complex incident investigations, root cause analyses, and reliability improvement initiatives.
  • Design and implement automation, self-service capabilities, and platform solutions that reduce operational toil.
  • Leverage AI-assisted engineering tools and automation platforms to accelerate troubleshooting, improve productivity, and reduce operational overhead.
  • Identify opportunities to streamline operational processes through automation, AI-enabled workflows, and platform engineering practices.
  • Drive adoption of SRE principles, reliability standards, and operational best practices across engineering organizations.
  • Develop and maintain infrastructure-as-code, deployment automation, and operational tooling.
  • Support and improve CI/CD and GitOps-based deployment workflows.
  • Design observability strategies using monitoring, logging, tracing, and alerting platforms.
  • Participate in architecture reviews and provide guidance on scalability, resiliency, and operational excellence.
  • Mentor junior engineers and contribute to the technical growth of the broader engineering organization.
  • Act as a force multiplier by creating reusable solutions, self-service capabilities, and engineering standards that increase the effectiveness of multiple teams.
  • Drive adoption of AI-assisted engineering workflows and operational automation across the organization.
  • Drive engineering leverage initiatives that improve the productivity, reliability, and effectiveness of multiple engineering teams.
  • Influence the broader engineering organization through platform thinking, standardization, and operational simplification.

Skills

Leadership
Reliability
Platform engineering
Automation
AI-assisted workflows
SRE practices
Incident response
CI/CD
GitOps
Python/Go/Bash
Networking

Tools

Kubernetes
Terraform
ArgoCD
Prometheus
Grafana

Job description

Medallia is seeking a Senior Site Reliability Engineer to design, operate, and evolve our cloud-native platforms. You will collaborate across teams to boost reliability, scalability, and performance while driving automation and AI-assisted workflows at scale.

The role demands technical leadership, hands-on implementation of IaC, CI/CD, and observability, plus participation in on-call rotations to support production systems. Strong collaboration and strategic thinking are essential.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE: AI-Driven Reliability & Automation (Hybrid)
Senior SRE: AI-Driven Reliability & Automation (Hybrid)

Namely • United States

Hybrid
USD 120,000 - 150,000
Senior SRE: Scale & Reliability for AI-Driven SaaS Platform
Senior SRE: Scale & Reliability for AI-Driven SaaS Platform

Instrumental Inc. • Palo Alto (CA)

On-site
USD 175,000 - 229,000
Health benefits
Commuter plans
Parental leave
Senior SRE Platform Engineer – AI-Powered Reliability
Senior SRE Platform Engineer – AI-Powered Reliability

UiPath • Denver (CO)

Hybrid
USD 160,000 - 210,000
Senior SRE & Platform Engineer: Reliability & Automation
Senior SRE & Platform Engineer: Reliability & Automation

Mindtech Company • United States

On-site
USD 120,000 - 180,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Medallia • McLean (VA)

On-site
USD 129,000 - 190,000
Health benefits
401(k) matching
Paid parental leave
+1
Senior SRE Platform Engineer – AI‑Driven Reliability
Senior SRE Platform Engineer – AI‑Driven Reliability

Socket.dev • Denver (CO)

Hybrid
USD 140,000 - 190,000
Senior Cloud Platform SRE for Scalable AI Systems
Senior Cloud Platform SRE for Scalable AI Systems

Mistral AI • Germany (OH)

On-site
USD 130,000 - 210,000
Healthcare coverage
Relocation support
Retirement plans
+2
Senior SRE: AI Cloud Reliability & Scale
Senior SRE: AI Cloud Reliability & Scale

Neura Market • San Francisco (CA)

Hybrid
USD 150,000 - 210,000
Health, dental, vision coverage
401k with company match
Wellness stipend
+1
Senior SRE: Lead Reliability for AI Platform (Hybrid)
Senior SRE: Lead Reliability for AI Platform (Hybrid)

Docebo • United States

Hybrid
USD 140,000 - 210,000
Health benefits
Paid vacation days
Docebo Days
+3
Senior SRE: AI Cloud Platform & Kubernetes Expert
Senior SRE: AI Cloud Platform & Kubernetes Expert

Lambda • Bellevue (WA)

On-site
USD 180,000 - 260,000
Health insurance
Dental insurance
Vision insurance
+3