Senior Platform Resiliency Engineer

Capital-Group-1

Charlotte (NC)

On-site

USD 137,000 - 219,000

Full time

12 days ago
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Flexible work options
Company-funded retirement contribution
Bonuses and comprehensive benefits

Job summary

Capital Group is seeking a Senior Infrastructure & Application Resiliency Engineer to architect and evolve fault-tolerant patterns across AWS. You will lead reliability initiatives, define SLOs/SLIs, and drive cross-functional resilience improvements with product, application, and platform teams.

You will mentor engineers, coordinate incidents, and implement self-healing and automated failover to meet strict RTO/RPO targets, while promoting best practices across the organization.

Qualifications

  • Bachelor’s degree in CS, Engineering, or equivalent practical cloud experience.
  • Minimum 10 years in SRE/DevOps/Cloud Architecture with tech lead experience.
  • Expert in AWS resiliency services (Application Recovery Controller, Resiliency Hub), multi-region architectures, Kubernetes (EKS).
  • Experience defining SLOs/SLIs and automating self-healing and orchestrated failover.

Responsibilities

  • Design multi-region, multi-AZ architectures and publish reference patterns enterprise-wide.
  • Lead chaos experiments and game days to prove resilience.
  • Automate self-healing and orchestrated failover to meet RTO/RPO targets.
  • Coordinate incident response for high-severity outages and publish blameless post-mortems.
  • Mentor senior engineers and raise reliability practices across the org.

Skills

SRE leadership
AWS expertise
Kubernetes / EKS
Terraform
Observability
Chaos engineering
Incident response
SLOs/SLIs
Automation
Communication

Education

Bachelor's degree in CS/Engineering or equivalent

Tools

Chaos Mesh
Gremlin
AWS Fault Injection Simulator
Prometheus
Datadog
OpenTelemetry
Temporal.io
Kafka
Terraform

Job description

Capital Group is seeking a Senior Infrastructure & Application Resiliency Engineer to architect and evolve fault-tolerant patterns across AWS. You will lead reliability initiatives, define SLOs/SLIs, and drive cross-functional resilience improvements with product, application, and platform teams.

You will mentor engineers, coordinate incidents, and implement self-healing and automated failover to meet strict RTO/RPO targets, while promoting best practices across the organization.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Platform Resilience Engineer (AWS/SRE Lead)
Senior Platform Resilience Engineer (AWS/SRE Lead)

504 CGCG-US CG Companies Global-US • Charlotte (NC)

On-site
USD 137,000 - 219,000
Generous time-off and health benefits
2-for-1 matching gifts
Annual grants for the organizations
+1
Senior Platform & Cloud Resilience Engineer (AWS/SRE)
Senior Platform & Cloud Resilience Engineer (AWS/SRE)

Capital Group • Charlotte (NC), Northern (KY)

Hybrid
USD 137,000 - 219,000
Time-away and health benefits from day
2-for-1 matching gifts for charitable
On-demand professional development
Platform Engineer Lead: Cloud Platform Architect
Platform Engineer Lead: Cloud Platform Architect

504 CGCG-US CG Companies Global-US • Charlotte (NC)

Hybrid
USD 154,000 - 246,000
Annual performance bonus
Retirement plan with 15% company 9con
Hybrid schedule
+1
Platform Engineer Lead – AWS Disaster Recovery & Resiliency
Platform Engineer Lead – AWS Disaster Recovery & Resiliency

Capital Group • Charlotte (AR)

On-site
USD 153,000 - 247,000
Generous time-off and health benefits
2-for-1 matching gift program
On-demand professional development resources
Cloud Platform Engineer - Resilience & DevOps Leader
Cloud Platform Engineer - Resilience & DevOps Leader

Capital One • Riverwoods (IL)

On-site
USD 150,000 - 171,000
Senior Platform Engineer — AI‑Augmented Cloud Infra
Senior Platform Engineer — AI‑Augmented Cloud Infra

Capital Group • Irvine (CA)

On-site
USD 136,000 - 255,000
Generous time off and health benefits
Flexible work options
Retirement contribution
+1
Platform Resiliency Lead: Disaster Recovery & Automation
Platform Resiliency Lead: Disaster Recovery & Automation

504 CGCG-US CG Companies Global-US • Town of Charlotte (NY)

On-site
USD 153,000 - 247,000
Individual annual performance bonus
Capital’s annual profitability bonus
15% company contribution to retirement plan
+5
Platform Reliability Architect
Platform Reliability Architect

Transformcap • San Francisco (CA)

Hybrid
USD 182,000 - 250,000
Comprehensive Health Coverage
Parental Leave & Family Support
401(k) program
+3
Senior Resilience Architect: Cloud, Incident Leadership
Senior Resilience Architect: Cloud, Incident Leadership

Capital One • Plano (TX)

On-site
USD 286,000 - 327,000
Senior Platform Engineer: Cloud, DevOps & Automation
Senior Platform Engineer: Cloud, DevOps & Automation

Relha LLC • Richmond (VA), Northern (KY)

Hybrid
USD 131,000 - 150,000