Senior Platform & Cloud Resilience Engineer (AWS/SRE)

Capital Group

Charlotte, Northern (NC, KY)

Hybrid

USD 137,000 - 219,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Time-away and health benefits from day
2-for-1 matching gifts for charitable
On-demand professional development

Job summary

Capital Group is seeking a Senior Infrastructure & Application Resiliency Engineer in Charlotte to architect, build, and evolve fault-tolerant infrastructure and patterns that support enterprise availability commitments. You’ll lead resilience initiatives across product, application, and platform teams, shaping reliability strategy and incident response.

You will design multi-region architectures, conduct chaos experiments, automate self-healing, and mentor engineers to raise the bar for

Qualifications

  • Bachelor’s degree in Computer Science, Engineering, or equivalent practical cloud experience.
  • At least 10 years in SRE, DevOps, or Cloud Architecture roles, with technical lead experience on enterprise-scale reliability initiatives.
  • Deep expertise in backup, recovery, disaster recovery, and cyber resilience on AWS — AWS Backup, cross-region recovery, immutable backups, recovery orchestration, application-consistent recovery, RTO/RPO planning, backup governance.
  • Expert in AWS resiliency services (e.g., Application Recovery Controller and Resiliency Hub), multi-region architectures, modern app architecture, and Kubernetes/container orchestration (e.g., EKS scaling and networking).
  • Proven experience defining SLOs/SLIs and error budgets and automating self-healing and orchestrated failover to meet strict RTO/RPO targets.
  • Hands-on experience leading chaos/failure-injection experiments with Chaos Mesh, Gremlin, or AWS Fault Injection Simulator (FIS).
  • Advanced Infrastructure‑as‑Code skills (e.g., Terraform) and observability with Prometheus/OpenTelemetry/Datadog; knowledge of AutoSys/Temporal.io/Kafka is a plus.
  • Proven ability to mentor engineers, drive operational excellence, and articulate complex technical challenges to business/IT execs.

Responsibilities

  • Working autonomously with wide latitude, you’ll be the strategic technical lead for complex, cross-functional resiliency solutions running in AWS—defining reliability strategy, including SLOs, error budgets, and disaster-recovery posture.
  • Design multi-region, multi-AZ architectures and publish reference patterns adopted across the enterprise; lead chaos experiments and game days to prove resilience; automate self-healing and orchestrated failover to hit RTO/RPO targets; coordinate incident response for high-severity outages; author blameless post‑mortems.

Skills

SRE/DevOps
Cloud Architecture
Mentoring
Operational excellence

Education

Bachelor’s degree in Computer Science/Engineering or equivalent

Tools

AWS resiliency services
Kubernetes (EKS)
Terraform
Prometheus/OpenTelemetry/Datadog
Chaos Mesh/Gremlin/FIS
AutoSys/Temporal.io/Kafka

Job description

Capital Group is seeking a Senior Infrastructure & Application Resiliency Engineer in Charlotte to architect, build, and evolve fault-tolerant infrastructure and patterns that support enterprise availability commitments. You’ll lead resilience initiatives across product, application, and platform teams, shaping reliability strategy and incident response.

You will design multi-region architectures, conduct chaos experiments, automate self-healing, and mentor engineers to raise the bar for

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Platform Resilience Engineer (AWS/SRE Lead)
Senior Platform Resilience Engineer (AWS/SRE Lead)

504 CGCG-US CG Companies Global-US • Charlotte (NC)

On-site
USD 137,000 - 219,000
Generous time-off and health benefits
2-for-1 matching gifts
Annual grants for the organizations
+1
Senior Platform Resiliency Engineer
Senior Platform Resiliency Engineer

Capital-Group-1 • Charlotte (NC)

On-site
USD 137,000 - 219,000
Flexible work options
Company-funded retirement contribution
Bonuses and comprehensive benefits
Senior Platform Engineer – Cloud & DevOps Leader
Senior Platform Engineer – Cloud & DevOps Leader

Capital One • McLean (VA)

On-site
USD 165,000 - 188,000
Platform Engineer Lead, Disaster Recovery & Resiliency
Platform Engineer Lead, Disaster Recovery & Resiliency

Capital Group • Charlotte (AR)

On-site
USD 153,000 - 247,000
Platform Engineer Lead: Cloud Platform Architect
Platform Engineer Lead: Cloud Platform Architect

504 CGCG-US CG Companies Global-US • Charlotte (NC)

Hybrid
USD 154,000 - 246,000
Annual performance bonus
Retirement plan with 15% company 9con
Hybrid schedule
+1
Senior Staff Resilience Engineer - Cloud & Reliability
Senior Staff Resilience Engineer - Cloud & Reliability

Relha LLC • Richmond (VA), Northern (KY)

Hybrid
USD 286,000 - 327,000
Senior Observability Platform Engineer
Senior Observability Platform Engineer

Capital-Group-1 • Charlotte (NC)

On-site
USD 131,000 - 219,000
Platform Engineer IV
Platform Engineer IV

504 CGCG-US CG Companies Global-US • Charlotte (NC)

On-site
USD 137,000 - 219,000
Generous time-off and health benefits
2-for-1 matching gifts
Annual grants for the organizations
+1
Platform Engineering Lead — Cloud Automation & Strategy
Platform Engineering Lead — Cloud Automation & Strategy

Capital Group • Charlotte (NC)

On-site
USD 154,000 - 246,000
Time-off & health benefits
Matching gifts program
Professional development resources
Cloud Platform Engineer - Resilience & DevOps Leader
Cloud Platform Engineer - Resilience & DevOps Leader

Capital One • Riverwoods (IL)

On-site
USD 150,000 - 171,000