AWS Cloud Ops Engineer: Automation & Reliability Lead

GuideWell

Northern (KY)

Hybrid

USD 109,000 - 178,000

Full time

3 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Medical, dental, vision insurance
Retirement plan with employer match
Paid time off and holidays
Hybrid work arrangement
Short/long-term disability coverage
Wellness program

Job summary

GuideWell is seeking a Cloud Operations Engineer to keep AWS workloads patched, backed up, monitored, and recoverable. The role focuses on building automation and operation standards to scale without headcount growth.

This is a standard business-hours role in the Eastern Time zone with a shared on-call rotation, not a NOC or shift job. You will automate recurring tasks, define runbooks, and drive reliability through tooling and incident response.

Qualifications

  • 6+ years related work experience in infrastructure operations or systems engineering with 4+ years experience operating cloud workloads.
  • 1+ years direct supervisory/management experience.
  • Related Bachelor's degree required.
  • Strong hands-on AWS experience running production workloads.
  • Proficiency in a programming or scripting language such as Python or Go.
  • Hands-on Terraform and CI/CD pipeline experience.
  • Experience defining and operating against SLOs and error budgets.
  • Observability tooling experience with CloudWatch, Prometheus, Grafana, Datadog, or equivalent.
  • Incident management and on-call experience in a production environment.
  • Kubernetes or EKS in production.

Responsibilities

  • Build and maintain monitoring and alerting for migrated workloads. Alert on conditions that require action.
  • Automate patching, backup validation, certificate renewal, and other recurring operational tasks.
  • Define and enforce operational readiness criteria a workload must meet before it goes live in cloud.
  • Build and test disaster recovery runbooks. Validate recovery time and recovery point targets through failover exercises.
  • Respond to incidents, participate in the on-call rotation, and drive root cause to closure.
  • Convert manual procedures into executable automation and self-healing where the task is repeatable.
  • Track availability and operational health against defined targets. Report on recurring issues.
  • Manage backup, retention, and restore testing for cloud workloads.
  • Troubleshoot and maintain AWS compute, including Windows Server and Linux instances.
  • Partner with infrastructure operations to transition migrated workloads into steady-state support.
  • Maintain runbooks and operational documentation a new on-call engineer can use unaided.

Skills

Python/Go
AWS
On-call/Incident management
Supervisory experience
SLOs & error budgets
Observability tools
Terraform & CI/CD
Kubernetes / EKS

Education

Bachelor's degree

Tools

Terraform
CI/CD
CloudWatch
Prometheus
Grafana
Datadog
Kubernetes / EKS

Job description

GuideWell is seeking a Cloud Operations Engineer to keep AWS workloads patched, backed up, monitored, and recoverable. The role focuses on building automation and operation standards to scale without headcount growth.

This is a standard business-hours role in the Eastern Time zone with a shared on-call rotation, not a NOC or shift job. You will automate recurring tasks, define runbooks, and drive reliability through tooling and incident response.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Cloud Automation & Reliability Engineer (AWS)
Cloud Automation & Reliability Engineer (AWS)

Socket.dev • Jacksonville (FL)

Hybrid
USD 109,000 - 178,000
Medical, dental, vision insurance
Retirement plan with employer match
Paid time off & holidays
AWS Cloud Operations & Automation Engineer
AWS Cloud Operations & Automation Engineer

Florida Blue Group • Northern (KY)

Hybrid
USD 109,000 - 178,000
Remote AWS Cloud Ops & Reliability Engineer
Remote AWS Cloud Ops & Reliability Engineer

HireArt • Austin (TX)

On-site
PTO (96 hours/year)
13 paid holidays
Pre‑tax commuter benefits
+4
Senior DevOps Engineer - AWS Automation & Reliability
Senior DevOps Engineer - AWS Automation & Reliability

Quest Global • Spring (TX)

On-site
USD 110,000 - 130,000
401(k)
401(k) matching
Dental insurance
+6
Remote Cloud Ops Engineer — Automation & Reliability
Remote Cloud Ops Engineer — Automation & Reliability

Branch App • United States

Remote
USD 135,000 - 150,000
Medical insurance
Dental insurance
Vision insurance
+4
Cloud DevOps Engineer: Automation & Reliability
Cloud DevOps Engineer: Automation & Reliability

Opsline • New York (NY)

On-site
USD 100,000 - 130,000
AWS Systems Engineer: Cloud Automation & Reliability
AWS Systems Engineer: Cloud Automation & Reliability

Elsevier • Gainesville (FL)

On-site
USD 90,000 - 120,000
Cloud Site Reliability Engineer – AWS & IaC Automation
Cloud Site Reliability Engineer – AWS & IaC Automation

Visa Hunt • United States

On-site
USD 80,000 - 133,000
Medical, Rx, Dental & Vision Insurance
Paid Holidays
Parental Leave
+13
Cloud Reliability Lead - IaC, Automation & SRE
Cloud Reliability Lead - IaC, Automation & SRE

Loftware • United States

Hybrid
USD 115,000 - 160,000
Senior AWS Cloud Engineer - DevOps & Automation
Senior AWS Cloud Engineer - DevOps & Automation

WASH • Nashville (TN)

On-site
USD 130,000 - 180,000