Senior Manager, AI-Driven SRE & Cloud Reliability

Intuit Inc.

Mountain View (CA)

On-site

USD 222,000 - 301,000

Full time

7 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Intuit Inc. in Mountain View seeks a Senior Manager, Site Reliability Engineering to lead a hands-on team of 10–15 engineers responsible for the availability and operational health of Fintech Platform services on AWS.

You will set technical direction, review designs, and coach engineers through complex production issues. A core priority is AI Ops: embedding autonomous operations to 3x impact, reducing toil, and accelerating value delivery.

Qualifications

  • 8+ years in systems engineering or SRE with 3+ years leading teams.
  • Proven AWS production infra experience at scale (EC2, EKS/ECS, VPC, RDS, etc).
  • Track record of high-availability outcomes for mission-critical, customer-facing systems.
  • Deep incident management experience with postmortems and prevention.
  • Strong foundation in distributed systems, networking, and IaC (Terraform/CloudFormation).
  • Experience with observability tools (Datadog, Splunk, Prometheus).
  • Ability to balance hands-on depth with people leadership and executive communication.

Responsibilities

  • Own end-to-end operational excellence for Fintech Platform services and achieve near-99.999% availability.
  • Lead and grow a team of 10–15 SREs, hiring, mentoring, and setting goals.
  • Participate in architecture reviews, contribute to design decisions, and write/read code or IaC as needed.
  • Define and execute an AI Ops roadmap to automate detection, diagnosis, and remediation.
  • Identify toil, replace with autonomous agents, and measure impact on velocity.
  • Drive incident management maturity and rigorous root-cause analysis.
  • Build scalable AWS infrastructure with resiliency, auto-remediation, and multi-region failover.
  • Define and report on SLOs/SLIs and quality metrics to guide investments.
  • Collaborate with product, security, and compliance teams on reliability.
  • Establish on-call playbooks and escalation paths to reduce MTTR.
  • Oversee capacity planning, cost optimization, and roadmap decisions.
  • Represent Infrastructure & SRE in leadership forums and executive reviews.
  • Foster a culture of operational rigor and continuous improvement.

Skills

Leadership
Incident management
AI Ops
Reliability engineering
Communication
Strategy

Education

Bachelor’s degree in CS/Engineering

Tools

Terraform
CloudFormation
Datadog
Prometheus/Grafana
Kubernetes
EC2
EKS

Job description

Intuit Inc. in Mountain View seeks a Senior Manager, Site Reliability Engineering to lead a hands-on team of 10–15 engineers responsible for the availability and operational health of Fintech Platform services on AWS.

You will set technical direction, review designs, and coach engineers through complex production issues. A core priority is AI Ops: embedding autonomous operations to 3x impact, reducing toil, and accelerating value delivery.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE Manager - AI Ops & 99.999% Availability
Senior SRE Manager - AI Ops & 99.999% Availability

Intuit • Mountain View (CA)

On-site
USD 222,000 - 301,000
Staff App Ops Engineer - Cloud Reliability & AI Ops
Staff App Ops Engineer - Cloud Reliability & AI Ops

Intuit • New York (NY)

On-site
USD 140,000 - 210,000
Senior Manager, Site Reliability Engineering
Senior Manager, Site Reliability Engineering

Intuit Inc. • Mountain View (CA)

On-site
USD 222,000 - 301,000
Staff App Ops Engineer - AI-Driven, Always-On Fintech
Staff App Ops Engineer - AI-Driven, Always-On Fintech

Intuit • Mountain View (CA)

On-site
USD 203,000 - 274,000
Senior Manager, Site Reliability Engineering
Senior Manager, Site Reliability Engineering

Intuit • Mountain View (CA)

On-site
USD 222,000 - 301,000
Senior Staff Software Engineer — AI-Driven, Global Impact
Senior Staff Software Engineer — AI-Driven, Global Impact

Intuit • Mountain View (CA)

On-site
USD 221,000 - 299,000
Staff App Ops Engineer — Cloud-Native Reliability
Staff App Ops Engineer — Cloud-Native Reliability

Intuit Inc. • New York (NY)

On-site
USD 180,000 - 240,000
Staff Backend Engineer: AI-Driven, Scalable Cloud Systems
Staff Backend Engineer: AI-Driven, Scalable Cloud Systems

Intuit Inc. • Mountain View (CA)

On-site
USD 203,000 - 274,000
Senior AI-Driven SRE for Cloud Reliability
Senior AI-Driven SRE for Cloud Reliability

Cerebras • Mountain View (CA)

Hybrid
USD 100,000 - 150,000
Competitive salary and benefits package
Opportunities for professional growth
Collaborative work environment
Senior Staff Architect — AI-Driven Fintech Platform
Senior Staff Architect — AI-Driven Fintech Platform

Intuit • Mountain View (CA)

On-site
USD 221,000 - 299,000