Senior Cloud Reliability Engineer - 24x7 SRE (FedRAMP)

Pegasystems

Norfolk (VA)

On-site

USD 102,000 - 153,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Continuous learning opportunities
Innovative, inclusive environment
Competitive global benefits + bonus/in

Job summary

Pegasystems seeks an experienced cloud operations engineer to join the Service Reliability Team (SRT). You will own and operate cloud infrastructure components across compute, storage, and platform services, collaborating with Engineering, Security, and Support to ensure reliability for mission‑critical SaaS environments.

You will work to reduce toil through automation, participate in on‑call rotations, and contribute to runbooks and SOPs, with strong focus on resilience and incident response.

Qualifications

  • 3+ years supporting enterprise cloud infrastructure or cloud operations for SaaS platforms.
  • 2+ years operating AWS infrastructure services (experience with GCP a plus).
  • 2+ years of Linux systems administration; working knowledge of AWS services and networking concepts; CCNA/CCNP a plus, but not required.

Responsibilities

  • Monitor, respond to, and resolve infrastructure alerts, incidents, service requests, and changes within SLA.
  • Own and drive customer-impacting escalations with focus on stability and service restoration.
  • Provision, operate, and upgrade cloud infrastructure components across compute, storage, and platform services.
  • Troubleshoot complex infrastructure and platform issues, perform root cause analysis, and contribute to long‑term fixes.
  • Create, maintain, and continuously improve runbooks, SOPs, and operational standards.
  • Partner with Engineering on pre‑release validation and operational readiness of new platform capabilities.
  • Identify opportunities to automate manual or repetitive operational tasks and reduce operational toil.
  • Participate in infrastructure‑focused projects and adapt to evolving business and platform requirements.
  • Participate in an after‑hours on‑call rotation, including weekend coverage.
  • Support FedRAMP‑compliant environments (U.S. citizenship and residency required)

Skills

Cloud operations
AWS
Linux administration
Networking concepts
Scripting (Bash/Python)

Tools

EC2, EBS, S3
ELB/Load Balancing
VPC, Transit Gateway, Route 53

Job description

Pegasystems seeks an experienced cloud operations engineer to join the Service Reliability Team (SRT). You will own and operate cloud infrastructure components across compute, storage, and platform services, collaborating with Engineering, Security, and Support to ensure reliability for mission‑critical SaaS environments.

You will work to reduce toil through automation, participate in on‑call rotations, and contribute to runbooks and SOPs, with strong focus on resilience and incident response.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Cloud Reliability Engineer – 24/7 Ops & FedRAMP
Cloud Reliability Engineer – 24/7 Ops & FedRAMP

Pegasystems • Salt Lake City (UT)

On-site
USD 102,000 - 153,000
Gartner leadership recognition
Continuous learning opportunities
Innovative, inclusive, agile culture
+1
Senior Cloud Reliability Engineer - 24x7 Ops
Senior Cloud Reliability Engineer - 24x7 Ops

Pegasystems • Phoenix (AZ)

On-site
USD 102,000 - 153,000
Gartner leadership recognition
Continuous learning opportunities
Flexible, inclusive work environment
+1
Senior Cloud Reliability Engineer – 24x7 SaaS
Senior Cloud Reliability Engineer – 24x7 SaaS

Pegasystems • Sterling (VA)

Hybrid
USD 102,000 - 154,000
Gartner leadership recognition
Continuous learning opportunities
Inclusive, agile, fun work environment
+1
Senior Cloud Reliability Engineer (Remote | Swing Shift)
Senior Cloud Reliability Engineer (Remote | Swing Shift)

Pegasystems • Provo (UT)

On-site
USD 102,000 - 153,000
Gartner leadership
Continuous learning
Inclusive environment
+1
Senior Cloud Operations Engineer - Global 24x7 Deployments
Senior Cloud Operations Engineer - Global 24x7 Deployments

Pegasystems • Waltham (MA), Northern (KY)

Hybrid
USD 102,000 - 153,000
Competitive global benefits
Bonus potential
Employee equity
FedRAMP SRE II — 24/7 Real-Time Operations Lead
FedRAMP SRE II — 24/7 Real-Time Operations Lead

C-Serv • United States

Hybrid
USD 96,000 - 110,000
Senior Cloud SRE & 24x7 Incident Resilience Engineer
Senior Cloud SRE & 24x7 Incident Resilience Engineer

Encora • United States

On-site
USD 90,000 - 130,000
Senior Cloud SRE – FedRAMP & Multi-Cloud Infra
Senior Cloud SRE – FedRAMP & Multi-Cloud Infra

Palo Alto Networks • Santa Clara (CA)

On-site
USD 120,000 - 200,000
Senior SRE (FedRAMP) – Cloud Reliability & Automation
Senior SRE (FedRAMP) – Cloud Reliability & Automation

Empleora • Northern (KY)

Hybrid
USD 160,000 - 210,000
Senior Platform SRE: FedRAMP & Cloud Reliability Leader
Senior Platform SRE: FedRAMP & Cloud Reliability Leader

Tata Consultancy Services • San Jose (CA)

On-site
USD 94,000 - 130,000