Principal Site Reliability Engineer

Palo Alto Networks

United States

Hybrid

USD 152,000 - 245,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Palo Alto Networks seeks a Site Reliability Engineer to join a large hybrid infrastructure supporting services on GCP, AWS, and on-premises alike. You will work on automation, architecture, performance, metrics, security, and reliability within a Kubernetes-based stack.

You’ll learn and apply a broad set of tools and technologies to keep services robust, scalable, and secure. The role emphasizes collaboration with developers, researchers, data scientists, and security experts, with mentoring and

Qualifications

  • BS or MS in Computer Science or related field, or equivalent experience.
  • Experience with infrastructure automation and configuration management.
  • Proficient in Python and/or Go.
  • Experience managing apps in Kubernetes with autoscaling.
  • Background in Production Engineering, DevOps, or SRE.
  • Strong experience with GCP or AWS, especially GCP.

Responsibilities

  • Contribute to the success of SRE and DevOps.
  • Develop expertise in new technologies.
  • Collaborate with developers, researchers, data scientists, and security experts.
  • Design, build, and operate reliable, secure cloud infrastructure.
  • Ensure applications are production-ready, scalable, and reliable.
  • Develop tools and automation frameworks.
  • Automate robust deployment of services.
  • Orchestrate end-to-end monitoring and alerting.
  • Participate in on-call rotations.
  • Lead root-cause analysis of critical issues.
  • Mentor and champion SRE culture.
  • Participate in design reviews.

Skills

Python
Go
Kubernetes
CI/CD
GitLab/GitHub
Linux administration
Automation scripting
Cloud platforms (GCP/AWS)
Monitoring as code

Education

BS or MS in Computer Science

Tools

Ansible
Terraform
Helm
GCP
AWS
GitLab
GitHub

Job description

Our Mission

At Palo Alto Networks®, we’re united by a shared mission—to protect our digital way of life. We thrive at the intersection of innovation and impact, solving real-world problems with cutting-edge technology and bold thinking. Here, everyone has a voice, and every idea counts. If you’re ready to do the most meaningful work of your career alongside people who are just as passionate as you are, you’re in the right place.

Who We Are

In order to be the cybersecurity partner of choice, we must trailblaze the path and shape the future of our industry. This is something our employees work at each day and is defined by our values: Disruption, Collaboration, Execution, Integrity, and Inclusion. We weave AI into the fabric of everything we do and use it to augment the impact every individual can have. If you are passionate about solving real-world problems and ideating beside the best and the brightest, we invite you to join us!

We believe collaboration thrives in person. That’s why most of our teams work from the office full time, with flexibility when it’s needed. This model supports real-time problem-solving, stronger relationships, and the kind of precision that drives great outcomes.

Job Summary
Your Career

Palo Alto Networks runs a large hybrid infrastructure and is one of the largest GCP customers. As a Site Reliability Engineer, you will be part of a team supporting the services running on this infrastructure. This includes automation, architecture, performance, metrics, troubleshooting, security, and reliability.

Our stack includes Kubernetes, Docker, GCP, AWS, Ansible, Terraform, Vault, Gitlab, Spinnaker, Pub/sub, Bigtable, Memorystore, Bigquery, RabbitMq, Kafka, MySQL, Python, and Go. We don’t expect you to know all these, but we do expect you to learn the ones needed for this role.

Your Impact
  • Contribute to the success of SRE and DevOps

  • Develop expertise in new technologies

  • Work with developers, researchers, data scientists, and security experts

  • Design, build, and operate reliable, secure Cloud infrastructure

  • Ensure that applications are production-ready, scalable, and reliable

  • Develop tools and automation frameworks

  • Automate robust deployment of robust services

  • Orchestrate end-to-end monitoring and alerting

  • Participate with SRE and Dev teams in the on-call rotation

  • Lead root cause analysis of critical business and production issues

  • Mentor and champion SRE culture

  • Participate in design reviews

The Team

Wildfire is the industry’s largest cloud-based malware protection engine that uses machine learning and crowdsourced intelligence to instantly prevent up to 95% of unknown malware variants inline without compromising business productivity. Wildfire infrastructure team supports the scalability and high availability of Wildfire clouds.

Qualifications
Required Experience
  • BS or MS in Computer Science, a related field, or equivalent professional experience or equivalent military experience

  • Expertise in configuration management with a framework such as Ansible, Terraform, Helm, Kubernetes

  • Proficient in Python and/or Go

  • Expertise in managing applications in the Kubenetes cluster with autoscaling enabled

  • Experience in Production Engineering, DevOps, or Site Reliability

  • Expertise in the public cloud (GCP or AWS), especially in GCP

  • Strong Linux administration, internals, and network troubleshooting

  • Proficiency with programming languages like Python, Golang, and shell scripting to automate tasks

  • Experience with CI/CD pipelines, GitLab, and GitHub preferred

  • Ability to diagnose and troubleshoot complex distributed systems handling high-volume transactions

  • Excellent written and verbal communication, able to collaborate and rally support

  • Self-disciplined, self-managed, self-motivated, and strong sense of ownership, urgency, and drive

  • Passion for infrastructure and monitoring as code

  • Ready to understand and dissect new technology stacks quickly

Compensation Disclosure

The compensation offered for this position will depend on qualifications, experience, and work location. For candidates who receive an offer at the posted level, the starting base salary (for non-sales roles) or base salary + commission target (for sales/com-missioned roles) is expected to be the annual range listed below. The offered compensation may also include restricted stock units and a bonus. A description of our employee benefits may be found here.

$151,600.00 - $245,300.00/yr

Our Commitment

We’re trailblazers that dream big, take risks, and challenge cybersecurity’s status quo. It’s simple: we can’t accomplish our mission without diverse teams innovating, together.

We are committed to providing reasonable accommodations for all qualified individuals with a disability. If you require assistance or accommodation due to a disability or special need, please contact us at accommodations@paloaltonetworks.com.

Palo Alto Networks is an equal opportunity employer. We celebrate diversity in our workplace, and all qualified applicants will receive consideration for employment without regard to age, ancestry, color, family or medical care leave, gender identity or expression, genetic information, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran status, race, religion, sex (including pregnancy), sexual orientation, or other legally protected characteristics.

All your information will be kept confidential according to EEO guidelines.

Is role eligible for Immigration Sponsorship?: Yes

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Cloud Infrastructure Engineer (Advanced Threat Protection)
Principal Cloud Infrastructure Engineer (Advanced Threat Protection)

Palo Alto Networks • United States

On-site
USD 150,000 - 230,000
Principal Site Reliability Engineer ( U.S Citizenship required )
Principal Site Reliability Engineer ( U.S Citizenship required )

Palo Alto Networks, Inc. • Santa Clara (CA)

On-site
USD 151,000 - 246,000
Manager, Site Reliability Engineering
Manager, Site Reliability Engineering

Palo Alto Networks, Inc. • Santa Clara (CA)

On-site
USD 180,000 - 260,000
Manager, Site Reliability Engineering
Manager, Site Reliability Engineering

Palo Alto Networks • United States

On-site
USD 180,000 - 240,000
Sr Site Reliability Engineer (Prisma Access)
Sr Site Reliability Engineer (Prisma Access)

Palo Alto Networks • Santa Clara (CA)

On-site
USD 120,000 - 200,000
Manager, Site Reliability Engineering
Manager, Site Reliability Engineering

Socket.dev • California (MO)

On-site
USD 150,000 - 230,000
Principal Cloud Infrastructure Engineer (Advanced Threat Protection)
Principal Cloud Infrastructure Engineer (Advanced Threat Protection)

Palo Alto Networks, Inc. • Santa Clara (CA)

On-site
USD 130,000 - 170,000
Employee benefits
Diverse workplace
Principal SRE Engineer (US Citizen)
Principal SRE Engineer (US Citizen)

Palo Alto Networks, Inc. • Santa Clara (CA)

On-site
USD 147,000 - 238,000
Senior Principal Engineer Software
Senior Principal Engineer Software

Palo Alto Networks, Inc. • Santa Clara (CA)

Hybrid
USD 120,000 - 150,000
Flexible working policies
Continuous learning opportunities
Sr Staff Engineer Software
Sr Staff Engineer Software

Palo Alto Networks, Inc. • Santa Clara (CA)

On-site
USD 126,000 - 205,000