Lead, Resiliency Engineer | Pune

Northern Trust

Pune District, Bengaluru

On-site

INR 3,000,000 - 4,500,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Northern Trust is seeking a Lead, Resiliency Engineer to design, implement, and manage disaster recovery programs across on-prem and cloud environments. You will leverage deep infrastructure expertise to identify risks, develop mitigations, and coordinate with cross-functional teams to keep the business secure and resilient during disruptions including cyber incidents.

You will oversee DR architecture for data centers, work with colocation providers, and drive SRM/Veeam-based recovery

Qualifications

  • 10+ years in IT infrastructure, with 6+ years focused on Disaster Recovery, Business Continuity, or Site Reliability Engineering.
  • Deep expertise designing and operating DR solutions on at least two major cloud platforms (AWS, Azure, or GCP) including cross-region replication, Route 53 / Traffic Manager / Cloud DNS failover, and managed database HA.
  • Extensive hands-on experience with on-premises DR technologies: VMware SRM, Veeam, Zerto, or equivalent.
  • Demonstrated experience building and executing ransomware recovery programs, including immutable storage and cyber recovery runbooks.
  • Strong command of compliance frameworks: ISO 22301, NIST SP 800-34 / CSF, SOC 2, and relevant sector regulations.
  • Proficiency in scripting and automation (Python, PowerShell, Bash) and IaC tools (Terraform, CloudFormation).
  • Proven track record leading enterprise-wide DR exercises and driving RTO/RPO attainment at scale.
  • Exceptional technical writing skills; able to produce executive-ready reports and detailed runbooks with equal clarity.

Responsibilities

  • Design, implement, and manage disaster recovery programs.
  • Develop DR plans, risk assessments, and response procedures.
  • Conduct regular DR testing and drills.
  • Collaborate with IT, security, and other departments.
  • Maintain DR documentation and incident response playbooks.
  • Lead cyber recovery exercises with CISO and SOC.
  • Train employees on DR and business continuity.
  • Stay updated on industry best practices and threats.

Skills

Disaster Recovery
Business Continuity
Site Reliability
Cloud Platforms
Scripting & IaC

Tools

VMware SRM
Veeam
Zerto
Terraform
CloudFormation
Route 53

Job description

Role & responsibilities

The Lead, Resiliency Engineer will be responsible for designing, implementing, and managing our disaster recovery programs. Your strong technical background and infrastructure expertise will be critical in identifying potential risks and developing strategies to mitigate the impact of disasters. You will work collaboratively with cross-functional teams to ensure our business remains secure and resilient in the event of any unforeseen disruptions including cyber recovery.

Key Responsibilities:
  • Disaster Recovery Planning: Develop and maintain comprehensive disaster recovery plans, including risk assessments, continuity strategies, and response procedures.
  • Risk Assessment: Identify potential threats and vulnerabilities, conducting risk assessments to evaluate their impact on business operations.
  • Disaster Recovery Testing: Plan, execute, and evaluate regular disaster recovery exercises to validate the effectiveness of recovery plans and make necessary adjustments.
  • Coordination: Collaborate with IT, security, and other relevant departments to ensure alignment between disaster recovery and security strategies.
  • Documentation: Maintain accurate documentation of disaster recovery plans, procedures, and incident response protocols.
  • Incident Response: Lead disaster recovery efforts in the event of a disruption, coordinating the response and recovery activities.
  • Training and Awareness: Develop and provide training to employees on disaster recovery and business continuity procedures to enhance preparedness.
  • Continuous Improvement: Stay updated on industry best practices, emerging technologies, and evolving threats to continually improve disaster recovery capabilities.
On-Premises & Data Center Resilience
  • Oversee DR architecture for on-premises data centers including storage replication (NetApp SnapMirror, Pure Storage ActiveCluster), SAN/NAS failover, and bare-metal recovery.
  • Manage relationships with colocation and secondary data center providers to ensure contractual alignment with DR objectives.
  • Drive server virtualization recovery strategies using VMware Site Recovery Manager (SRM) and Veeam.
Cybersecurity & Ransomware Recovery
  • Design and maintain immutable backup architectures and air-gapped environments to protect against ransomware and destructive cyberattacks.
  • Lead cyber recovery exercises simulating ransomware scenarios; document and refine clean-room recovery playbooks.
  • Collaborate with the CISO and SOC to integrate DR procedures into the Incident Response (IR) lifecycle, ensuring seamless handoffs from containment to recovery.
  • Champion zero-trust recovery principles, including identity verification during failover and integrity validation of recovered workloads.
Compliance, Audit & Governance
  • Maintain DR program alignment with ISO 22301, NIST SP 800-34, SOC 2 Type II, and applicable industry regulations (HIPAA, PCI-DSS, GDPR as relevant).
  • Own all DR-related evidence gathering, documentation, and remediation activities for internal and external audits.
  • Report program health, test outcomes, and risk metrics to executive leadership and the Board on a regular cadence.
  • Establish governance frameworks including DR policy, standards, and exception management processes.
Testing, Drills & Continuous Improvement
  • Plan and execute full-scale DR tests (tabletop, functional, and full failover) across production-equivalent environments at least twice per year.
  • Track and drive closure of test findings; maintain a risk register for unresolved gaps.
  • Implement chaos engineering principles and game-day exercises to proactively uncover resilience weaknesses.
Qualifications:
  • 10+ years in IT infrastructure, with at least 6 years focused on Disaster Recovery, Business Continuity, or Site Reliability Engineering.
  • Deep expertise designing and operating DR solutions on at least two major cloud platforms (AWS, Azure, or GCP) including cross-region replication, Route 53 / Traffic Manager / Cloud DNS failover, and managed database HA.
  • Extensive hands-on experience with on-premises DR technologies: VMware SRM, Veeam, Zerto, or equivalent.
  • Demonstrated experience building and executing ransomware recovery programs, including immutable storage and cyber recovery runbooks.
  • Strong command of compliance frameworks: ISO 22301, NIST SP 800-34 / CSF, SOC 2, and relevant sector regulations.
  • Proficiency in scripting and automation (Python, PowerShell, Bash) and IaC tools (Terraform, CloudFormation).
  • Proven track record leading enterprise-wide DR exercises and driving RTO/RPO attainment at scale.
  • Exceptional technical writing skills; able to produce executive-ready reports and detailed runbooks with equal clarity.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Disaster Recovery Manager
Disaster Recovery Manager

Insight Global • Mumbai

Hybrid
INR 1,800,000 - 2,400,000
CLOUD ARCHITECT - Ansible
CLOUD ARCHITECT - Ansible

Happiest Minds Technologies • Bengaluru

On-site
INR 1,400,000 - 2,100,000
IT Resiliency Engineer
IT Resiliency Engineer

AppSierra • Pune City

On-site
INR 1,500,000 - 2,500,000
Manager (IT Infrastructure)
Manager (IT Infrastructure)

Metaphor Infotech Mumbai • Mumbai

On-site
INR 1,500,000 - 2,300,000
Lead, Technology Resiliency Governance
Lead, Technology Resiliency Governance

Northern Trust Corp • Pune District

Hybrid
INR 2,800,000 - 4,200,000
Senior Disaster Recovery (DR) Engineer
Senior Disaster Recovery (DR) Engineer

Alter Domus • Hyderabad

On-site
INR 2,500,000 - 4,000,000
ACC A study leave
Birthday day off
Employee Share Plan
Senior Manager - Disaster Recovery Governance
Senior Manager - Disaster Recovery Governance

Cognizant • Bengaluru Urban

Hybrid
INR 3,000,000 - 4,500,000
Hybrid work model
Wellbeing programs
Manager - Disaster Recovery
Manager - Disaster Recovery

Golden Opportunities • Chennai District

On-site
INR 1,200,000 - 2,400,000
Senior Backup Engineer
Senior Backup Engineer

Forward Eye Technologies • Dadri

On-site
INR 1,200,000 - 1,800,000
Senior Backup Engineer
Senior Backup Engineer

Forward Eye Technologies • Pune District

On-site
INR 1,800,000 - 2,400,000