Principal Platform Engineer (High Availability & Disaster Recovery)

Anaplan Inc

Gurugram District

On-site

INR 3,500,000 - 7,000,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Anaplan Inc. in Gurugram, India, is seeking a Principal HA & DR Engineer to own the architectural strategy and execution for the platform's resilience, uptime, and disaster recovery.

You will design and deliver automated failover, cross-region coordination, and self-healing mechanisms to protect critical workloads. We value a deep cloud and on‑prem hybrid skill set, experience with Kubernetes, IaC (Terraform/Ansible), and incident management to drive uptime and meet strict SLAs across global

Qualifications

  • 8+ years of HA engineering experience across cloud and on-prem environments.
  • Proven ability to evangelize resiliency standards and influence cross-functional engineering teams.
  • Hands-on expertise with real-time data replication and automated, zero-downtime traffic failover.
  • Infrastructure as Code (IaC) using Terraform or Ansible to manage highly available environments.
  • Strong Kubernetes and Linux internals knowledge; experience with container orchestration.

Responsibilities

  • Drive the HA Strategy & Roadmap: end-to-end design and execution for global High Availability across cloud and on-prem environments.
  • Build & Implement HA Architectures: active-active clustering, global load balancing, real-time DB replication to eliminate single points of failure.
  • Establish the Resiliency Practice: define standards for fault tolerance and self-healing across the platform.
  • Influence Engineering Teams: ensure self-healing mechanisms and automated recovery are embedded in core deliverables.
  • Bootstrap Chaos Engineering: define strategy and guardrails for first automated fault-injection and live-fire drills.
  • Optimize Uptime & Manage Risks: monitor KPIs, mitigate bottlenecks, and ensure SLAs.

Skills

HA Engineering
Technical Influence
Failover & Replication
Infrastructure as Code
Kubernetes
Hybrid Networking
Strategic Execution
SaaS Experience
Incident Management
Observability

Tools

Terraform
Ansible
Kubernetes

Job description

At Anaplan, we are a team of innovators focused on optimizing business decision-making through our leading AI-infused scenario planning and analysis platform so our customers can outpace their competition and the market.

What unites Anaplanners across teams and geographies is our collective commitment to our customers' success and to our Winning Culture.

Our customers rank among the who's who in the Fortune 50. Coca-Cola, LinkedIn, Adobe, LVMH and Bayer are just a few of the 2,400+ global companies who rely on our best-in-class platform.

Our Winning Culture is the engine that drives our teams of innovators. We champion diversity of thought and ideas, we behave like leaders regardless of title, we are committed to achieving ambitious goals, and we love celebratingour wins - big and small.

Supported by operating principles of being strategy-led, values -based and disciplined in execution, you'll be inspired, connected, developed and rewarded here. Everything that makes you unique is welcome; join us and let's build what's next - together!

Principle Platform DR & Capacity Engineer

At Anaplan, we are a team of innovators focused on optimizing business decision-making through our leading AI-infused scenario planning and analysis platform so our customers can outpace their competition and the market.

What unites Anaplanners across teams and geographies is our collective commitment to our customers' success and to our Winning Culture.

Our customers rank among the who's who in the Fortune 50. Coca-Cola, LinkedIn, Adobe, LVMH and Bayer are just a few of the 2,400+ global companies who rely on our best-in-class platform.

Our Winning Culture is the engine that drives our teams of innovators. We champion diversity of thought and ideas, we behave like leaders regardless of title, we are committed to achieving ambitious goals, and we love celebratingour wins - big and small.

Supported by operating principles of being strategy-led, values -based and disciplined in execution, you'll be inspired, connected, developed and rewarded here. Everything that makes you unique is welcome; join us and let's build what's next - together!

Your impact

As the Principal HA & DR Engineer, you will own the architectural strategy and technical execution for the platform's resilience, high availability, and disaster recovery lifecycle. You will design, build, and deliver technical solutions that support Platform availability, resiliency, and disaster response.

You will act as the principal authority for business continuity across our global infrastructure, defining the guardrails for system fault-tolerance and leading cross-functional coordination for multi-region disaster preparedness between both cloud and on prem. You will also design automated failover mechanisms, champion chaos engineering practices, and develop robust mitigation strategies to eliminate single points of failure across both cloud and on-premises infrastructure.

You will :
  • Drive the HA Strategy & Roadmap: Own the end-to-end design and technical execution of our global High Availability roadmap, ensuring continuous platform uptime across both cloud and on-premises environments.
  • Build & Implement HA Architectures: Hand-on code, configure, and engineer active-active clustering, global load balancing, and real-time database replication to completely eliminate single points of failure.
  • Establish the Resiliency Practice: Build and champion our foundational Resiliency Engineering framework from scratch, defining platform-wide standards for fault tolerance and self-healing.
  • Influence Engineering Teams: Partner with and influence cross-functional engineering squads to ensure self-healing mechanisms and automated recovery protocols are embedded directly into their core deliverables.
  • Bootstrap Chaos Engineering: Introduce and champion our first-ever Chaos Engineering program; you will define the strategy, establish the safe guardrails, and prepare the organization to execute its initial automated fault-injection and live-fire failover drills.
  • Build & Implement HA Architectures: Hand-on code, configure, and engineer active-active clustering, global load balancing, and real-time database replication to completely eliminate single points of failure.
  • Optimize Uptime & Manage Risks: Monitor infrastructure performance KPIs to identify potential bottlenecks, mitigate shortage or overload risks before they cause downtime, and ensure compliance with strict Customer SLA requirements.

Your skills and experience

We work with a broad range of technologies, and wedon'texpect you to know everything on day one.You’llhave time to learn our tools and grow into the role.We'relooking for diverse experiences to help strengthen our team.

  • 8+ Years of HA Engineering: Extensive HA experience across cloud (AWS, GCP, or Azure) and on-premises environments.
  • Technical Influence: Proven ability to evangelize resiliency standards and influence cross-functional engineering teams to build self-healing features.
  • Failover & Replication: Hands-on expertise with real-time data replication, database clustering, and automated, zero-downtime traffic failover.
  • Infrastructure as Code (IaC): Advanced proficiency with Terraform or Ansible to build and replicate highly available environments programmatically.
  • Systems & Container Engineering: Deep technical knowledge of Linux internals and container orchestration using Kubernetes (K8s).
  • Hybrid Network Engineering: Hands-on experience managing complex hybrid networking topologies, including BGP routing, DNS management, Anycast, and CDNs.
  • Strategic Execution: Ability to translate high-level uptime requirements into a technical roadmap and personally execute the engineering work.
  • Experience supporting SaaS products.
  • Experience with Incident Management, Post Mortems and related practices.
  • Knowledge of observability and monitoring best practices.
  • Experience operating within one or more public clouds (AWS, GCP, Azure).
  • Experience with configuration management, and infrastructure as code
  • Knowledge of observability andmonitoringbest practices
Our Commitment to Diversity, Equity, Inclusionand Belonging (DEIB)

We believe attracting and retaining the best talent and fostering an inclusive culture strengthens our business. DEIB improves our workforce, enhances trust with our partners and customers, and drives business success. Build your career in a place where diversity, equity, inclusion and belonging aren’t just words on paper - this is what drives our innovation, it’s how we connect, and it contributes to what makes us a market leader. We believe in a hiring and working environment where all people are respected and valued, regardless of gender identity or expression, sexual orientation, religion, ethnicity, age, neurodiversity, disability status, citizenship, or any other aspect which makes people unique. We hire you for who you are, and we want you to bring your authentic self to work every day!

We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, perform essential job functions, and receive equitable benefits and all privileges of employment. Please contact us to request accommodation.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Platform Engineer (High Availability & Disaster Recovery)
Principal Platform Engineer (High Availability & Disaster Recovery)

Anaplan • Gurugram District

On-site
INR 5,000,000 - 9,000,000
Platform DR & Capacity Engineer
Platform DR & Capacity Engineer

Anaplan • Gurugram District

On-site
INR 1,800,000 - 2,800,000
Platform DR & Capacity Engineer
Platform DR & Capacity Engineer

Anaplan Inc • Gurugram District

On-site
INR 1,800,000 - 3,000,000
Infrastructure Engineer
Infrastructure Engineer

Anaplan • Gurugram District

Hybrid
INR 1,200,000 - 1,800,000
Senior Data Engineer
Senior Data Engineer

Anaplan Inc • India

On-site
INR 1,500,000 - 2,300,000
Senior QA Engineer - Platform Experience(SDET)
Senior QA Engineer - Platform Experience(SDET)

Anaplan • Gurugram District

On-site
INR 1,200,000 - 1,600,000
Sr. Application Analyst-Anaplan Model Builder- Supply Chain
Sr. Application Analyst-Anaplan Model Builder- Supply Chain

Anaplan • Gurugram District

On-site
INR 1,800,000 - 2,600,000
Senior Software Engineer
Senior Software Engineer

Anaplan • Gurugram District

On-site
INR 4,000,000 - 7,500,000
Flexible working
Catered lunches
Fully stocked kitchen
+3
Data & Analytics Systems Engineer
Data & Analytics Systems Engineer

Anaplan Inc • India

On-site
INR 1,000,000 - 1,400,000
Program Manager
Program Manager

Anaplan Inc • India

On-site
INR 1,800,000 - 2,400,000