Senior Cloud Reliability Engineer: Multi-Cloud & AI

Palo Alto Networks

United States

On-site

USD 150,000 - 230,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Palo Alto Networks is seeking a Senior Site Reliability Engineer to design and operate cloud platforms across GCP, AWS, and global data centers. You will push AI-driven automation in incident detection, remediation, and capacity planning while improving reliability and performance.

You will mentor SREs, partner with development teams, and drive end-to-end automation across large-scale infrastructure, ensuring production-readiness and scalable operations.

Qualifications

  • 7+ years of experience in DevOps, Site Reliability, or infrastructure engineering.
  • Experience in multi-cloud environments — GCP, AWS, and familiarity with OCI.
  • Experience designing and operating infra across multiple cloud providers.
  • Infrastructure as Code with Terraform and Ansible.
  • Strong Python and shell scripting for automation.
  • Experience applying AI/ML to operational workflows is a strong plus.

Responsibilities

  • Design, build, and operate cloud infrastructure enabling reliable deployments and monitoring.
  • Leverage AI/ML to automate incident detection, root cause analysis, and remediation.
  • Build and integrate AI-powered tools into SRE workflows for intelligent alerting and capacity planning.
  • Write automation code for provisioning and operating infrastructure at massive scale.
  • Develop self-healing systems that diagnose issues and take corrective action with minimal human intervention.
  • Mentor other SREs on best practices in infrastructure orchestration and AI-augmented operations.
  • Represent SRE in design reviews and work cross-functionally with engineering teams on readiness.

Skills

DevOps & SRE
Python
Shell scripting
Linux
Distributed systems
CI/CD
Networking fundamentals
Problem solving
AI/ML in operations

Education

BS or MS in Computer Science or related field

Tools

Terraform
Ansible
GitLab
Artifactory
LLM-based tooling
AIOps platforms

Job description

Palo Alto Networks is seeking a Senior Site Reliability Engineer to design and operate cloud platforms across GCP, AWS, and global data centers. You will push AI-driven automation in incident detection, remediation, and capacity planning while improving reliability and performance.

You will mentor SREs, partner with development teams, and drive end-to-end automation across large-scale infrastructure, ensuring production-readiness and scalable operations.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Cloud SRE & AI-Driven Infra Architect
Lead Cloud SRE & AI-Driven Infra Architect

Palo Alto Networks, Inc. • Santa Clara (CA)

On-site
USD 130,000 - 170,000
Employee benefits
Diverse workplace
Senior AI Cloud Engineer - Multi-Cloud, IaC & Automation
Senior AI Cloud Engineer - Multi-Cloud, IaC & Automation

Palo Alto Networks, Inc. • Santa Clara (CA)

On-site
USD 145,000 - 235,500
Principal Cloud SRE & Automation Engineer
Principal Cloud SRE & Automation Engineer

Palo Alto Networks, Inc. • Santa Clara (CA)

On-site
USD 151,000 - 246,000
Senior Multi-Cloud Security Engineer
Senior Multi-Cloud Security Engineer

Palo Alto Networks, Inc. • California (MO)

On-site
USD 152,000 - 245,000
Senior AI-Native Cloud Software Engineer
Senior AI-Native Cloud Software Engineer

Palo Alto Networks, Inc. • Santa Clara (CA)

On-site
USD 145,000 - 235,500
Senior SRE Manager: Reliability, Automation & Leadership
Senior SRE Manager: Reliability, Automation & Leadership

Palo Alto Networks, Inc. • Santa Clara (CA)

On-site
USD 180,000 - 260,000
Cloud Infrastructure SRE: Reliability & Automation Engineer
Cloud Infrastructure SRE: Reliability & Automation Engineer

Robotics Prcocess Automation, LLC • Berkeley Heights (NJ)

On-site
Senior Product Manager, Cloud Reliability & AI Ops
Senior Product Manager, Cloud Reliability & AI Ops

Palo Alto Networks, Inc. • Santa Clara (CA)

On-site
USD 185,000 - 301,000
SRE Engineering Manager — Reliability & Incidents Lead
SRE Engineering Manager — Reliability & Incidents Lead

Palo Alto Networks • United States

On-site
USD 180,000 - 240,000
Senior PM Lead, Operations & Reliability — Cloud SaaS
Senior PM Lead, Operations & Reliability — Cloud SaaS

Palo Alto Networks • Santa Clara (CA)

On-site
USD 185,900 - 300,675