Cloud Reliability Engineer – 24/7 Ops & FedRAMP

Pegasystems

Salt Lake City (UT)

On-site

USD 102,000 - 153,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Gartner leadership recognition
Continuous learning opportunities
Innovative, inclusive, agile culture
Competitive global benefits including:

Job summary

Pegasystems is hiring for a Service Reliability Engineer on our SRT to ensure reliability, performance, and security of our cloud platforms. You’ll own systems, drive resiliency, and reduce toil through automation in a global, follow‑the‑sun environment.

The role requires hands‑on cloud infrastructure experience, AWS proficiency, Linux administration, and a strong platform mindset. Training and after‑hours on‑call are included as part of the shift pattern.

Qualifications

  • 3+ years supporting enterprise cloud infrastructure or cloud operations for SaaS platforms.
  • 2+ years of hands-on experience operating AWS infrastructure services (experience with GCP a plus).
  • 2+ years of Linux systems administration experience. Working knowledge of AWS services, including EC2, EBS, S3; ELB/Load Balancing; VPC, Transit Gateway, Route 53.
  • Experience supporting highly available, fault-tolerant cloud environments. Familiarity with infrastructure automation and scripting (Bash, Shell, Python, or similar).
  • Exposure to networking concepts (DNS, load balancing, firewalls) as part of broader infrastructure operations; AWS or GCP certification preferred; CCNA/CCNP a plus, but not required.

Responsibilities

  • Monitor, respond to, and resolve infrastructure alerts, incidents, service requests, and changes within SLA.
  • Own and drive customer-impacting escalations with a focus on stability and service restoration.
  • Provision, operate, and upgrade cloud infrastructure components across compute, storage, and platform services.
  • Troubleshoot complex infrastructure and platform issues, perform root cause analysis, and contribute to long‑term fixes.
  • Create, maintain, and continuously improve runbooks, SOPs, and operational standards.
  • Partner with Engineering on pre‑release validation and operational readiness of new platform capabilities.
  • Identify opportunities to automate manual or repetitive operational tasks and reduce operational toil.
  • Participate in infrastructure‑focused projects and adapt to evolving business and platform requirements.
  • Participate in an after‑hours on‑call rotation, including weekend coverage.
  • Support FedRAMP‑compliant environments (U.S. citizenship and residency required).

Skills

Cloud operations
Infrastructure engineering
AWS
Linux
Networking basics
SaaS platforms

Tools

Bash
Shell
Python

Job description

Pegasystems is hiring for a Service Reliability Engineer on our SRT to ensure reliability, performance, and security of our cloud platforms. You’ll own systems, drive resiliency, and reduce toil through automation in a global, follow‑the‑sun environment.

The role requires hands‑on cloud infrastructure experience, AWS proficiency, Linux administration, and a strong platform mindset. Training and after‑hours on‑call are included as part of the shift pattern.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Cloud Reliability Engineer - 24x7 SRE (FedRAMP)
Senior Cloud Reliability Engineer - 24x7 SRE (FedRAMP)

Pegasystems • Norfolk (VA)

On-site
USD 102,000 - 153,000
Continuous learning opportunities
Innovative, inclusive environment
Competitive global benefits + bonus/in
Senior Cloud Reliability Engineer – 24x7 SaaS
Senior Cloud Reliability Engineer – 24x7 SaaS

Pegasystems • Sterling (VA)

Hybrid
USD 102,000 - 154,000
Gartner leadership recognition
Continuous learning opportunities
Inclusive, agile, fun work environment
+1
Senior Cloud Reliability Engineer - 24x7 Ops
Senior Cloud Reliability Engineer - 24x7 Ops

Pegasystems • Phoenix (AZ)

On-site
USD 102,000 - 153,000
Gartner leadership recognition
Continuous learning opportunities
Flexible, inclusive work environment
+1
Senior Cloud Reliability Engineer (Remote | Swing Shift)
Senior Cloud Reliability Engineer (Remote | Swing Shift)

Pegasystems • Provo (UT)

On-site
USD 102,000 - 153,000
Gartner leadership
Continuous learning
Inclusive environment
+1
Senior Cloud Operations Engineer - Global 24x7 Deployments
Senior Cloud Operations Engineer - Global 24x7 Deployments

Pegasystems • Waltham (MA), Northern (KY)

Hybrid
USD 102,000 - 153,000
Competitive global benefits
Bonus potential
Employee equity
Senior Cloud Operations Engineer, Infrastruture
Senior Cloud Operations Engineer, Infrastruture

Pegasystems • Norfolk (VA)

On-site
USD 102,000 - 153,000
Continuous learning opportunities
Innovative, inclusive environment
Competitive global benefits + bonus/in
Cloud Reliability Engineer – 24/7 DoD Ops (TS/SCI)
Cloud Reliability Engineer – 24/7 DoD Ops (TS/SCI)

Peraton • Chantilly (VA)

On-site
USD 112,000 - 179,000
25 days PTO accrued annually
Eligible for bonus plan
Cloud Operations Engineer: Kubernetes & AWS (Remote, US)
Cloud Operations Engineer: Kubernetes & AWS (Remote, US)

Pegasystems • California (MO)

On-site
USD 73,000 - 110,000
Bonus potential
Employee equity
Global benefits program
Senior Cloud Operations Engineer, Infrastruture
Senior Cloud Operations Engineer, Infrastruture

Pegasystems • Salt Lake City (UT)

On-site
USD 102,000 - 153,000
Gartner leadership recognition
Continuous learning opportunities
Innovative, inclusive, agile culture
+1
Senior Cloud Operations Engineer, Infrastruture
Senior Cloud Operations Engineer, Infrastruture

Pegasystems • Phoenix (AZ)

On-site
USD 102,000 - 153,000
Gartner leadership recognition
Continuous learning opportunities
Flexible, inclusive work environment
+1