Systems Architect, Disaster Recovery

Donatech Corporation

United States

On-site

USD 180,000 - 280,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Donatech Corporation is seeking an Architectural Leader to define end-to-end disaster recovery and high-availability architectures for enterprise workloads across multi-region cloud, hybrid, and on-prem environments. You will develop blueprints, reference designs, and patterns that align with security, compliance, and cost policies, and lead DR exercise planning and execution.

In this role you will collaborate with IT Infrastructure, Cloud Engineering, Application Development, Security, and

Qualifications

  • 5+ years of experience designing and implementing DR/HA for enterprise workloads.
  • Bachelor’s degree in Computer Science, IT, Engineering, or related field (or equivalent).
  • Hands-on with cloud platforms (AWS, Azure, GCP) and DR services.
  • Strong networking, storage, virtualization, and container orchestration knowledge.
  • Excellent written and verbal communication; able to explain technical concepts clearly.
  • U.S. citizenship required.

Responsibilities

  • Define end-to-end DR and high-availability architectures for multi-region environments.
  • Develop blueprints, reference designs, and pattern libraries aligned with security and cost policies.
  • Design automated failover, replication, and failback mechanisms.
  • Evaluate emerging technologies to improve resiliency and reduce MTTR.
  • Lead DR exercises and report metrics to leadership.
  • Collaborate with IT Infrastructure, Cloud Engineering, Security, and Governance teams.
  • Drive MBSE adoption and automated documentation to keep artefacts current.

Skills

DR/HA design
Cloud platforms (AWS/Azure/GCP)
Kubernetes
Networking & storage understanding
Communication skills
Security & compliance awareness
US citizenship

Education

Bachelor’s degree in Computer Science, IT, Engineering, or related field

Tools

Site Recovery Manager (SRM)
Terraform
CloudFormation
Jenkins
GitLab CI/CD
Zerto
Veeam
IBM Resiliency Services

Job description

Position would require the candidate to be a W2 employee of Donatech. US Citizenship Required.

Architectural Leadership
  • Define end to end DR and high availability (HA) architectures for enterprise wide workloads, incorporating multi region cloud, hybrid, and on prem solutions.
  • Develop architectural blueprints, reference designs, and pattern libraries that align with client’s security, compliance, and cost optimization policies.
Solution Design & Implementation
  • Design and implement automated fail over, replication, and fail back mechanisms (e.g., Site Recovery Manager, Kubernetes based HA, database mirroring, storage level replication).
  • Evaluate and integrate emerging technologies (e.g., Immutable Infrastructure, Chaos Engineering, Serverless DR) to improve resiliency and reduce mean time to recover (MTTR).
Governance & Compliance
  • Ensure all DR solutions meet corporate policies CRX 301, CRX 302, and relevant regulatory requirements (e.g., NIST?800 34, ISO?22301, FedRAMP).
  • Create and maintain DR documentation, run books, and test plans; conduct periodic reviews and updates.
Testing & Validation
  • Lead full scale DR exercise planning, execution, and post mortem analysis for multi site, multi cloud environments.
  • Define success criteria, metrics, and KPIs; report findings to senior leadership and stakeholders.
Stakeholder Collaboration
  • Partner with IT Infrastructure, Cloud Engineering, Application Development, Security, and Governance teams to embed DR/HA considerations early in the SDLC.
  • Serve as the technical authority for DR during design reviews (SRR, PDR, CDR, TRR) and program risk assessments.
Continuous Improvement
  • Conduct risk assessments, threat modeling, and capacity planning to anticipate emerging resiliency challenges.
  • Drive adoption of Model Based Systems Engineering (MBSE) and automated documentation tools to keep architecture artefacts current.
DR Plan Modernization & Compliance
  • Review existing DR plan architectures across the enterprise, assessing their alignment with current resilience standards, best practices, and organizational Recovery Objectives.
  • Collaborate with internal teams (Application Owners, IT Service Managers, Engineering) to update and refine DR plans, ensuring that all applications and IT services meet the latest RTO/RPO targets.
  • Develop and implement remediation plans to bring legacy systems and applications up to date with modern resilience standards, ensuring compliance with corporate policies (CRX 301, CRX 302) and regulatory requirements.
  • Track progress and report status to senior leadership, providing insights into plan modernization efforts and risk mitigation strategies.
Basic Qualifications :
  • 5?+?years of experience designing and implementing DR/HA solutions for enterprise scale workloads in cloud, hybrid, and on prem environments.
  • Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related technical discipline (or equivalent experience).
  • Proven hands on experience with cloud platforms (AWS, Azure, GCP) and related services (e.g., Disaster Recovery, Site Recovery Manager, cross region replication, networking, IAM).
  • Strong understanding of networking, storage, virtualization, container orchestration (Kubernetes), and database technologies as they relate to resiliency.
  • Excellent written and verbal communication skills; ability to translate complex technical concepts for both technical and non technical audiences.
  • U.S. citizenship required
Desired skills :
  • Familiarity with automated disaster recovery (DR) solutions, including but not limited to:
  • Amazon Web Services (AWS) Disaster Recovery Service (DRS): Experience with configuring and managing replication, fail over, and fail back processes for AWS workloads.
  • Microsoft Azure Site Recovery (ASR): Knowledge of setting up and managing site recovery between on premises environments, Azure, and other clouds.
  • Zerto: Hands on experience with continuous data protection (CDP) and near zero RPO replication across VMware, Hyper V, and cloud environments.
  • Veeam Backup & Replication: Experience with agent less backup, replication, and automated fail over testing for virtual, physical, and cloud workloads.
  • IBM Resiliency Services (formerly IBM Disaster Recovery as a Service): Familiarity with managed DR services for hybrid cloud environments, including integration with IBM Cloud and on premises infrastructure.
  • Experience with DR automation, including:
  • Scripting and integration with IaC tools (Terraform, CloudFormation) and CI/CD pipelines (Jenkins, GitLab) to automate DR workflows.
  • DR exercise planning and execution, including defining success criteria, metrics, and KPIs for recovery processes.
  • Strong analytical skills: ability to perform risk assessments, impact analysis, and cost benefit modeling for DR solutions.
  • Hybrid/multi-cloud deployments, with ability to manage DR across multiple cloud providers and on premises environments.
  • Advanced certifications (e.g., AWS Certified Solutions Architect - Professional, Azure Solutions Architect Expert, VMware VCAP DCV).
Standard Job Description :

Designs, develops and oversees system architectures, including complex systems and systems design activities. Serves as focal point for the development and communication of the system architecture. Ensures conversion of mission requirements into total systems solutions that account for design and technology maturity constraints of the system. Develops systems and system element architecture, design, and interface definition. Supports internal and external design reviews. Maintains knowledge of current and developing technologies and design and analysis methodologies. Develops models and architectural guidelines for current and future system development.

Typical Minimums :

Generally has 5+ years of related experience and may have a post-secondary degree or training in a related discipline.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Enterprise DR/HA Architect – Multi-Cloud Resilience Lead
Enterprise DR/HA Architect – Multi-Cloud Resilience Lead

Donatech Corporation • United States

On-site
USD 180,000 - 280,000
Disaster Recovery Leader
Disaster Recovery Leader

Unisys • Northern (KY)

Hybrid
USD 140,000 - 210,000
Disaster Recovery Analyst
Disaster Recovery Analyst

Eliassen Group • Maryland Heights (MO)

Hybrid
USD 83,000 - 104,000
Medical, Dental, Vision benefits
401k with company matching
Life insurance
Strategic IT Disaster Recovery Lead
Strategic IT Disaster Recovery Lead

The Timberline Group • St. Louis (MO)

On-site
USD 120,000 - 180,000
Disaster Recovery Leader
Disaster Recovery Leader

Unisys Corporation • Pennsylvania

On-site
USD 150,000 - 210,000
Disaster Recovery Leader
Disaster Recovery Leader

Unisys Corporation • Norwich (CT)

On-site
USD 180,000 - 230,000
Resiliency Architect
Resiliency Architect

ALLTECH CONSULTING SVC INC • Town of Texas (WI)

On-site
USD 100,000 - 130,000
Senior Technical Architect - Cyber Resilience
Senior Technical Architect - Cyber Resilience

Calance • New York (NY)

On-site
USD 150,000 - 230,000
Senior Solution Business Continuity & Resilient Infrastructure
Senior Solution Business Continuity & Resilient Infrastructure

Buchanan Technologies • Dallas (TX), Northern (KY)

On-site
USD 140,000 - 190,000
Disaster Recovery Leader
Disaster Recovery Leader

Unisys • Pennsylvania

On-site
USD 150,000 - 190,000