Principal Cloud Infrastructure Engineer (Advanced Threat Protection)

Palo Alto Networks, Inc.

Santa Clara (CA)

On-site

USD 130,000 - 170,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Employee benefits
Diverse workplace

Job summary

Palo Alto Networks, Inc. is seeking a Senior Site Reliability Engineer to design and operate cloud-native infrastructure across GCP and AWS. This role involves leveraging AI and automation to enhance operational efficiency and reliability.

The ideal candidate will have over 7 years of experience in SRE or DevOps, with expertise in multi-cloud environments and automation tools. Your contributions will include building resilient systems and mentoring other SREs while driving best practices.

Qualifications

  • 7+ years of experience in DevOps, Site Reliability, or infrastructure engineering.
  • Experience with GCP, AWS, and OCI.
  • Strong problem-solving and communication skills.

Responsibilities

  • Design and operate platforms that power applications across GCP and AWS.
  • Automate incident detection and root cause analysis.
  • Mentor other SREs on best practices and automation.

Skills

DevOps
Site Reliability Engineering
Multi-cloud environments (GCP, AWS)
Infrastructure as Code (Terraform, Ansible)
Python scripting
Linux proficiency
CI/CD pipelines
AI/ML in operations

Education

BS or MS in Computer Science or related field

Tools

Terraform
Ansible
GitLab
Artifactory
Docker

Job description

Job Summary:

Palo Alto Networks is at the forefront of cloud-native infrastructure, where reliability, scale, and intelligent automation define the future of operations. As a Senior Site Reliability Engineer, you will design and operate the platforms that power our applications across GCP, AWS, and global data centers - and you'll push the boundary of what's possible by leveraging AI and machine learning to transform how we approach SRE.

This isn't just about keeping the lights on. You'll build intelligent systems that predict incidents before they happen, automate root cause analysis, and continuously optimize our infrastructure. You'll be a critical bridge between engineering and our Infrastructure Platform, combining deep SRE expertise with AI-driven automation to deliver unprecedented levels of reliability and operational efficiency.

If you're excited about applying AI to real-world infrastructure challenges - and you thrive in an environment where automation isn't just a nice-to-have but a core philosophy - this is your next career.

Your Impact
  • Design, build, and operate cloud infrastructure that enables reliable, rapid deployment of microservices with resilient operations and effective monitoring
  • Leverage AI/ML to automate incident detection, root cause analysis, and remediation - reducing toil and accelerating mean time to resolution
  • Build and integrate AI-powered tools (LLM-based agents, AIOps platforms) into SRE workflows for intelligent alerting, log analysis, and capacity planning
  • Write automation code for provisioning and operating infrastructure at massive scale
  • Develop self-healing systems that can automatically detect anomalies, diagnose issues, and take corrective action with minimal human intervention
  • Work with development teams to ensure applications are production-ready, scalable, and reliable from the ground up
  • Identify and drive opportunities to improve automation for code deployment, management, and observability of application services
  • Establish end-to-end monitoring and alerting on all critical components, incorporating AI-driven anomaly detection and predictive analytics
  • Participate in the on-call rotation supporting the platform and production applications
  • Lead root cause analysis of critical business and production issues, building runbooks and automation to prevent recurrence
  • Mentor other SREs on best practices in infrastructure orchestration, production troubleshooting, and AI-augmented operations
  • Represent SRE in design reviews and work cross-functionally with engineering teams on operational readiness
Qualifications
  • 7+ years of experience in DevOps, Site Reliability, or infrastructure engineering
  • Expertise in multi-cloud environments - strong hands-on experience with GCP, AWS, and familiarity with OCI (Oracle Cloud Infrastructure)
  • Experience designing and operating infrastructure across multiple cloud providers, including networking, identity management, and cross-cloud connectivity
  • Expertise in Infrastructure as Code with tools such as Terraform, Ansible
  • Strong proficiency in Python and shell scripting for automation
  • Strong experience with Linux and distributed systems handling high-volume transactions
  • Familiarity with CI/CD pipelines, GitLab, and Artifactory
  • Strong fundamentals in HTTP, web servers, and networking
  • BS or MS in Computer Science, a related field, or equivalent professional experience
  • Excellent problem solving, critical thinking, communication, and teamwork skills
  • Self-disciplined, self-managed, self-motivated with a strong sense of ownership, urgency, and drive
  • Experience applying AI/ML to operational workflows (AIOps, intelligent alerting, automated remediation, or LLM-powered tooling) is a strong plus
  • Experience with cloud compliance frameworks (FedRAMP, IL5) and operating in regulated environments is a plus
  • Experience building and managing large database systems - relational (MySQL, PostgreSQL) and non-relational (Redis, BigQuery, etc.) - is a plus
Compensation Disclosure

The compensation offered for this position will depend on qualifications, experience, and work location. For candidates who receive an offer at the posted level, the starting base salary (for non-sales roles) or base salary + commission target (for sales/com-missioned roles) is expected to be the annual range listed below. The offered compensation may also include restricted stock units and a bonus. A description of our employee benefits may be found here.

Our Commitment

We’re trailblazers that dream big, take risks, and challenge cybersecurity’s status quo. It’s simple: we can’t accomplish our mission without diverse teams innovating, together.

We are committed to providing reasonable accommodations for all qualified individuals with a disability. If you require assistance or accommodation due to a disability or special need, please contact us at accommodations@paloaltonetworks.com.

Palo Alto Networks is an equal opportunity employer. We celebrate diversity in our workplace, and all qualified applicants will receive consideration for employment without regard to age, ancestry, color, family or medical care leave, gender identity or expression, genetic information, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran status, race, religion, sex (including pregnancy), sexual orientation, or other legally protected characteristics.

All your information will be kept confidential according to EEO guidelines.

Is role eligible for Immigration Sponsorship? No. Please note that we will not sponsor applicants for work visas for this position.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Staff DevOps Engineer
Senior Staff DevOps Engineer

Palo Alto Networks • United States

On-site
USD 126,000 - 204,000
Principal SRE Engineer (US Citizen)
Principal SRE Engineer (US Citizen)

Palo Alto Networks, Inc. • Santa Clara (CA)

On-site
USD 147,000 - 238,000
Staff Software Engineer(Cloudops)
Staff Software Engineer(Cloudops)

Palo Alto Networks • Santa Clara (CA)

On-site
USD 124,000 - 202,000
Staff Software Engineer(Cloudops)
Staff Software Engineer(Cloudops)

Palo Alto Networks • California (MO)

On-site
USD 124,000 - 202,000
Sr Principal Software Engineer/Architect (SRE/Platforms/Agentic AI)
Sr Principal Software Engineer/Architect (SRE/Platforms/Agentic AI)

Palo Alto Networks • Santa Clara (CA)

On-site
USD 230,000 - 340,000
Principal Software Engineer (Cloud Platform and AI Engineering)
Principal Software Engineer (Cloud Platform and AI Engineering)

Jobs Paloaltonetworks • Santa Clara (CA)

On-site
USD 147,000 - 238,000
Sr Technical Support Engineer, Cortex Cloud
Sr Technical Support Engineer, Cortex Cloud

Palo Alto Networks, Inc. • Plano (TX)

On-site
USD 106,000 - 172,000
Sr. Staff Engineer Software, Infrastructure Reliability (Chronosphere)
Sr. Staff Engineer Software, Infrastructure Reliability (Chronosphere)

Palo Alto Networks • Boston (MA)

Hybrid
USD 126,000 - 205,000
Restaurant d'entreprise
Indemnités de stage/alternance
Sr Technical Support Engineer, Cortex Cloud
Sr Technical Support Engineer, Cortex Cloud

Jobs Paloaltonetworks • Town of Texas (WI)

On-site
USD 106,000 - 172,000
Sr. Principal Software Engineer (L7 Security)
Sr. Principal Software Engineer (L7 Security)

Palo Alto Networks, Inc. • San Francisco (CA)

On-site
USD 170,000 - 277,000
Equity and stock options
Comprehensive health benefits
Work-life balance initiatives