Staff Replication Engineer – High-Availability Data Platform

Ddn

Raleigh (NC)

On-site

USD 150,000 - 190,000

Full time

4 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

DDN is seeking a Staff Replication Development Engineer to lead the design and development of the replication engine for the Infinia AI Data Platform. This role focuses on building enterprise-grade asynchronous replication capabilities for reliable disaster recovery in large-scale data systems.

You will develop high-performance replication pipelines, secure data transfer systems, and robust data integrity mechanisms, partnering with backend, security, and platform teams to deliver end-to-end

Qualifications

  • 8+ years of experience in distributed systems, storage systems, or backend software engineering.
  • Strong programming skills in C++, Go, Java, or Rust.
  • Experience designing and building data replication systems, data pipelines, or distributed data services.
  • Deep understanding of distributed systems concepts (consistency, availability, scalability, fault tolerance).
  • Strong expertise in multi-threading, concurrency, and parallel processing.
  • Knowledge of networking protocols and secure communication (TCP/IP, HTTP/HTTPS, TLS).
  • Experience implementing data integrity mechanisms (checksums, validation, consistency checks).
  • Experience designing and building REST APIs and service-based architectures.
  • Familiarity with checkpointing, failure recovery, and retry mechanisms in distributed systems.
  • Basic understanding of observability concepts (metrics, logging, alerting).
  • Strong debugging, problem-solving, and system design skills.

Responsibilities

  • Design and develop multi-threaded asynchronous replication systems with parallel streaming capabilities.
  • Build object-level delta replication with checkpointing and resume functionality.
  • Develop replication engines supporting bucket/share-level replication controls.
  • Implement secure data transfer mechanisms using TLS 1.3 with mutual authentication.
  • Ensure end-to-end data integrity through checksum validation and verification pipelines.
  • Design and implement manual failover workflows for disaster recovery scenarios.
  • Build and maintain REST APIs for replication configuration, control, and automation.
  • Develop metadata tracking and change detection systems to enable efficient replication.
  • Implement RPO visibility, alerting, and operational insights for replication status.
  • Contribute to monitoring dashboards focused on replication health and performance.
  • Ensure systems are designed for high availability, fault tolerance, and scalability.
  • Partner with QA teams to drive performance, resiliency, and scale validation.
  • Collaborate with backend, security, and platform teams to deliver end-to-end replication workflows.
  • Participate in debugging, production issue resolution, and continuous improvement of replication reliability.
  • Provide technical leadership, architectural guidance, and mentorship to the engineering team.

Skills

Distributed systems
C++
Go
Java
Rust
Multi-threading
Concurrency
Networking
Security
Observability

Job description

DDN is seeking a Staff Replication Development Engineer to lead the design and development of the replication engine for the Infinia AI Data Platform. This role focuses on building enterprise-grade asynchronous replication capabilities for reliable disaster recovery in large-scale data systems.

You will develop high-performance replication pipelines, secure data transfer systems, and robust data integrity mechanisms, partnering with backend, security, and platform teams to deliver end-to-end

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff Engineer, Replication
Staff Engineer, Replication

Ddn • Raleigh (NC)

On-site
USD 150,000 - 190,000
Staff Replication Development Engineer
Staff Replication Development Engineer

DDN • San Francisco (CA)

On-site
USD 185,000 - 230,000
Lead Replication Systems Engineer
Lead Replication Systems Engineer

DDN • San Francisco (CA)

On-site
USD 185,000 - 230,000
Staff Engineer
Staff Engineer

Ddn • Santa Clara (CA)

On-site
USD 180,000 - 240,000
Data Replication Engineer: IMS/SQL (On-Site)
Data Replication Engineer: IMS/SQL (On-Site)

DXC Technology • Indiana (PA)

On-site
USD 110,000 - 150,000
Senior Database Replication Engineer — On-site in Indiana
Senior Database Replication Engineer — On-site in Indiana

DXC Technology • Burns Harbor (IN)

On-site
USD 120,000 - 150,000
Staff Engineer, Lakeflow Disaster Recovery & Replication
Staff Engineer, Lakeflow Disaster Recovery & Replication

Cacheflow • San Francisco (CA)

On-site
USD 192,000 - 260,000
Technical Lead, AI Data Platform & Distributed Storage
Technical Lead, AI Data Platform & Distributed Storage

Ddn • Santa Clara (CA)

On-site
USD 180,000 - 280,000
Staff Software Engineer, AiDP
Staff Software Engineer, AiDP

Ddn • Santa Clara (CA)

On-site
USD 180,000 - 280,000
Senior AI Data Path & Storage Architect
Senior AI Data Path & Storage Architect

DDN • California (MO)

Hybrid
USD 180,000 - 250,000