Senior Cloud SRE – Infra, CI/CD & Observability

Zerto

Puerto Rico

Hybrid

Confidential

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Hybrid work model
San Juan office

Job summary

Hewlett Packard Enterprise seeks a Principal Site Reliability Engineer to design, build, and optimize cloud infrastructure, ensuring scalability, security, and reliability across platforms. You will lead enhancements to CI/CD, monitoring, and security programs while collaborating with development and security teams.

The role requires strong Linux, cloud, and IaC skills, plus experience with Kafka/Cassandra. This is a hybrid role with San Juan office attendance twice weekly.

Qualifications

  • 10+ years hands-on experience in Infra Ops, DevOps, or Site Reliability Engineering (SRE).
  • Proficiency with Linux systems, especially Debian-based distributions.
  • Strong experience with cloud platforms such as AWS and GCP.
  • Expertise in Infrastructure as Code tools like Terraform, Packer, and Ansible.
  • Solid programming skills in Python and/or Golang.
  • Deep understanding of containerization (Docker, Container) and orchestration tools (AWS EKS, GCP GKE).
  • Experience with GitOps workflows.
  • Proven track record in implementing and maintaining CI/CD pipelines.
  • Strong background in security and familiarity with security programs.
  • Experience with monitoring and logging tools (Prometheus, Grafana, ELK).
  • Knowledge of both relational (SQL) and non-relational databases.
  • Excellent problem-solving and debugging skills with a strong sense of ownership.
  • Experience managing distributed systems like Apache Kafka and Cassandra.
  • Effective communicator and collaborative team player.
  • It is mandatory to attend to San Juan office twice a week.

Responsibilities

  • Design, build, and optimize cloud infrastructure.
  • Improve CI/CD pipelines with FluxCD and Jenkins.
  • Address container image vulnerabilities and remediation.
  • Build AMIs aligned with CIS and STIG.
  • Strengthen monitoring with Prometheus, Grafana, ELK.
  • Troubleshoot production issues for reliability.
  • Collaborate with dev, security, and ops teams.
  • Enhance IAC and enforce best practices.

Skills

Infra Ops
DevOps
SRE
Linux Debian
AWS
GCP
Terraform
Packer
Ansible
Python
Golang
Docker
Kubernetes
GitOps
CI/CD
Security
Prometheus
Grafana
ELK
SQL/NoSQL

Tools

AWS EKS
GCP GKE
Terraform
Packer
Ansible
FluxCD
Jenkins

Job description

Hewlett Packard Enterprise seeks a Principal Site Reliability Engineer to design, build, and optimize cloud infrastructure, ensuring scalability, security, and reliability across platforms. You will lead enhancements to CI/CD, monitoring, and security programs while collaborating with development and security teams.

The role requires strong Linux, cloud, and IaC skills, plus experience with Kafka/Cassandra. This is a hybrid role with San Juan office attendance twice weekly.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Cloud SRE Engineer — Infra, CI/CD & Observability
Senior Cloud SRE Engineer — Infra, CI/CD & Observability

Hewlett Packard Enterprise Development LP • San Juan (PR)

Hybrid
USD 130,000 - 190,000
Health & Wellbeing
Personal & Professional Development
Unconditional Inclusion
Hybrid Staff SRE — Cloud Infra, CI/CD & Observability
Hybrid Staff SRE — Cloud Infra, CI/CD & Observability

Zerto • Puerto Rico

Hybrid
Confidential
Staff SRE: Hybrid Cloud Infra Lead
Staff SRE: Hybrid Cloud Infra Lead

Hewlett Packard Enterprise Development LP • San Juan (PR)

Hybrid
USD 120,000 - 170,000
Senior Staff SRE - Hybrid (2 days/wk in-office)
Senior Staff SRE - Hybrid (2 days/wk in-office)

Zerto • Puerto Rico

Hybrid
Confidential
Principal SRE: Hybrid Cloud Reliability & Observability
Principal SRE: Hybrid Cloud Reliability & Observability

Hewlett Packard Enterprise • San Juan (PR)

Hybrid
USD 120,000 - 160,000
Comprehensive benefits suite
Personal & professional development opportunities
Unconditional inclusion in the workplace
Senior Site Reliability Engineer — Cloud Infra & Automation
Senior Site Reliability Engineer — Cloud Infra & Automation

Careers Inc • Puerto Rico

On-site
USD 100,000 - 140,000
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Zerto • Puerto Rico

Hybrid
Confidential
Hybrid work model
San Juan office
Site Reliability Engineer Sr. Staff
Site Reliability Engineer Sr. Staff

Zerto • Puerto Rico

Hybrid
Confidential
AI Ops Engineer — Observability & Automation
AI Ops Engineer — Observability & Automation

Hewlett Packard Enterprise • San Juan (PR)

Hybrid
USD 90,000 - 150,000
Hybrid work model (2 days in-office)
Relocation support
Senior Cloud & Distributed Systems Engineer (Hybrid)
Senior Cloud & Distributed Systems Engineer (Hybrid)

Hewlett Packard Enterprise • San Juan (PR)

Hybrid
USD 120,000 - 180,000