Cloud Site Reliability Engineer: Build Resilient SaaS

ADP

Town of Texas (WI)

On-site

USD 120,000 - 150,000

Full time

3 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Aspect Software is seeking a Site Reliability Engineer to build and operate reliable, scalable cloud services with a focus on observability, automation, and incident management.

You will collaborate with cross-functional teams to define SLOs/SLIs, automate toil, and strengthen incident response across SaaS platforms. The role emphasizes resilient deployment practices and continuous improvement of reliability and performance.

Qualifications

  • Bachelor's degree in Computer Science, Engineering, or related field, or equivalent practical experience.
  • 4-5+ years of experience in site reliability engineering, DevOps, or production software engineering.
  • Hands-on experience operating production workloads in AWS; Azure or GCP is valuable.
  • Experience with infrastructure as code and configuration automation (e.g., AWS CDK, CloudFormation, Terraform).
  • Experience building or supporting CI/CD pipelines and source-control workflows (GitHub).
  • Working knowledge of Linux, networking, DNS, load balancing, firewalls, certificates, and common distributed-systems failure modes.
  • Observability and monitoring experience (Datadog, Grafana, CloudWatch).
  • Understanding of incident management, root-cause analysis, and blameless post-incident practices.

Responsibilities

  • Design, build, and operate highly available, scalable production services and cloud infrastructure.
  • Define and maintain SLIs, SLOs, error budgets, and actionable alerts with engineering teams.
  • Develop automation to reduce toil and improve deployment reliability.
  • Improve observability with metrics, logs, tracing, dashboards, and health checks.
  • Participate in on-call rotations and lead incident response and post-incident reviews.
  • Enhance CI/CD pipelines with automated validation, progressive delivery, canary releases, and reliable rollbacks.
  • Review designs for reliability, capacity, security, and cost efficiency.

Tools

AWS
Terraform
GitHub Actions
Datadog
Grafana

Job description

Aspect Software is seeking a Site Reliability Engineer to build and operate reliable, scalable cloud services with a focus on observability, automation, and incident management.

You will collaborate with cross-functional teams to define SLOs/SLIs, automate toil, and strengthen incident response across SaaS platforms. The role emphasizes resilient deployment practices and continuous improvement of reliability and performance.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Cloud Reliability Engineer
Senior Cloud Reliability Engineer

Alvaria • Town of Texas (WI), Northern (KY)

Hybrid
USD 120,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

ADP • Town of Texas (WI)

On-site
USD 120,000 - 150,000
Site Reliability Engineer
Site Reliability Engineer

Alvaria • Town of Texas (WI), Northern (KY)

Hybrid
USD 120,000 - 180,000
Senior Site Reliability Engineer: Build Resilient, Scalable Systems
Senior Site Reliability Engineer: Build Resilient, Scalable Systems

Inspire-Brands • Atlanta (GA)

On-site
USD 140,000 - 190,000
Senior Site Reliability Engineer - Edge & Cloud Resilience
Senior Site Reliability Engineer - Edge & Cloud Resilience

Saasventurecapital • El Segundo (CA)

On-site
USD 170,000 - 190,000
Comprehensive health coverage
401(k) with 4% company match
Flexible paid time off
+2
Site Reliability Engineer – Cloud & On-Prem Reliability
Site Reliability Engineer – Cloud & On-Prem Reliability

Charles River Associates International • Boston (MA)

Hybrid
USD 130,000 - 150,000
Healthcare benefits
Paid time off
401(k) matching
Senior Site Reliability Engineer — Cloud, Resilience & Automation
Senior Site Reliability Engineer — Cloud, Resilience & Automation

Compunnel Inc. • Denton (TX)

Hybrid
USD 120,000 - 150,000
Senior Site Reliability Engineer — Scalable, Observability-Driven
Senior Site Reliability Engineer — Scalable, Observability-Driven

Origami Risk • Chicago (IL)

Hybrid
USD 100,000 - 120,000
Hybrid work arrangement
401(k) with company match
Medical, dental, and vision benefits
+1
Senior Site Reliability Engineer: Build Resilient Systems
Senior Site Reliability Engineer: Build Resilient Systems

IRB USA Inspire Resources • Atlanta (GA)

On-site
USD 130,000 - 180,000
AWS Cloud DevOps Engineer – SaaS Platform & CI/CD
AWS Cloud DevOps Engineer – SaaS Platform & CI/CD

ADP • Town of Texas (WI)

On-site
USD 90,000 - 140,000