Senior Replication Engineer - Scalable Data Platform

Data Direct Networks

Santa Clara (CA)

On-site

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job description

We are on our way to being the first company to power 1 MILLION GPUs and want world-class talent to join our amazing team!

The world is moving faster than ever, and yet, it will never move this slowly again. We are at the forefront of an incredible technological revolution but, at its core, it is fueled by incredible people. People like you.

1,000 Global Employees

11k Happy Customers

16 International Offices

We’re Looking for the Best and Brightest

We are the world’s leading data intelligence platform that reliably accelerates massive datasets for actionable real-time insights. Join our team to help the best and brightest minds tackle the world’s biggest challenges in business, science, medicine, academia and government.

Do What Can’t be Done

For the past 20 years, our team has kept us at the forefront of storage technology and has provided the foundation for enabling researchers to push the limits of “what can be done.”
These innovations take research and discovery to the next level, enabling them to discover cures to disease, observe global warming patterns, model innovative automotive and aerospace designs, discover new sources of energy, make communities safer, and accelerate business results across a wide variety of industries.

DDN Helps Build Your Future, Too

At DDN, we understand our customers’ diverse needs. Whether you’re a data scientist, IT professional, executive, or researcher, our solutions empower you with cutting‑edge technology and unparalleled support.

Highly Competitive Vacation Plans

Paid Holidays

Bonus Programs

Tuition Reimbursement

Employee Referral Program

Excellent Medical, Dental and Vision Benefits

Paid Leave Programs

Anniversary and Recognition Awards

Location
Employment Type

Full time

Location Type

Hybrid

DDN is seeking a Staff Replication Development Engineer to lead the design and development of the replication engine for the Infinia AI Data Platform. This role focuses on building enterprise‑grade asynchronous replication capabilities that enable reliable and secure disaster recovery for large‑scale data systems.

You will work on developing high‑performance replication pipelines, efficient data synchronization mechanisms, and secure data transfer systems. This role requires deep expertise in distributed systems and strong technical leadership to deliver a scalable and resilient replication foundation.

Key Responsibilities

Design and develop multi‑threaded asynchronous replication systems with parallel streaming capabilities

Build object‑level delta replication with checkpointing and resume functionality

Implement secure data transfer mechanisms using TLS 1.3 with mutual authentication

Ensure end‑to‑end data integrity through checksum validation and verification pipelines

Design and implement manual failover workflows for disaster recovery scenarios

Build and maintain REST APIs for replication configuration, control, and automation

Develop metadata tracking and change detection systems to enable efficient replication

Implement RPO visibility, alerting, and operational insights for replication status

Contribute to monitoring dashboards focused on replication health and performance

Ensure systems are designed for high availability, fault tolerance, and scalability

Partner with QA teams to drive performance, resiliency, and scale validation

Collaborate with backend, security, and platform teams to deliver end‑to‑end replication workflows

Participate in debugging, production issue resolution, and continuous improvement of replication reliability

Provide technical leadership, architectural guidance, and mentorship to the engineering team

Required Qualifications

8+ years of experience in distributed systems, storage systems, or backend software engineering

Strong programming skills in one or more languages: C++, Go, Java, or Rust

Experience designing and building data replication systems, data pipelines, or distributed data services

Deep understanding of distributed systems concepts (consistency, availability, scalability, fault tolerance)

Strong expertise in multi‑threading, concurrency, and parallel processing

Knowledge of networking protocols and secure communication (TCP/IP, HTTP/HTTPS, TLS)

Experience implementing data integrity mechanisms (checksums, validation, consistency checks)

Experience designing and building REST APIs and service‑based architectures

Familiarity with checkpointing, failure recovery, and retry mechanisms in distributed systems

Basic understanding of observability concepts (metrics, logging, alerting)

Strong debugging, problem‑solving, and system design skills

Preferred Qualifications

Experience with asynchronous replication, disaster recovery (DR), or backup systems

Familiarity with object storage or large‑scale data storage systems

Knowledge of delta encoding, change data capture, or incremental data synchronization techniques

Experience building high‑throughput, low‑latency data movement systems

Exposure to security practices including mutual TLS, encryption, and authentication

Experience working on enterprise‑scale data platforms or storage products

Familiarity with performance optimization and large‑scale system tuning

We pride ourselves on our commitment to delivering tangible and consistent results.

Regional Director of Sales, Middle East and Africa

Let’s Forge a Better Future, Together

Explore our current job openings and find the perfect opportunity to advance your career with a company that values expertise, creativity and growth.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Replication Development Engineer
Staff Replication Development Engineer

Data Direct Networks • Santa Clara (CA)

On-site
Staff Replication Development Engineer
Staff Replication Development Engineer

DDN • San Francisco (CA)

On-site
USD 185,000 - 230,000
Staff Replication Development Engineer
Staff Replication Development Engineer

Ddn • Sacramento (CA)

On-site
USD 180,000 - 240,000
Staff Replication Development Engineer
Staff Replication Development Engineer

DDN • Santa Clara (CA)

On-site
USD 150,000 - 190,000
Staff Replication Development Engineer
Staff Replication Development Engineer

DDN • North Carolina

Hybrid
USD 120,000 - 170,000
Lead Replication Engineer, Distributed Data Systems
Lead Replication Engineer, Distributed Data Systems

DDN • Santa Clara (CA)

On-site
USD 150,000 - 190,000
Staff Replication Systems Architect
Staff Replication Systems Architect

Doist • North Carolina

Hybrid
USD 120,000 - 170,000
Sr Staff Engineer
Sr Staff Engineer

Data Direct Networks • California (MO)

On-site
USD 180,000 - 260,000
Highly Competitive Vacation Plans
Paid Holidays
Bonus Programs
+5
Senior Staff Engineer
Senior Staff Engineer

Data Direct Networks • North Carolina

Hybrid
USD 225,000 - 275,000
Highly Competitive Vacation Plans
Paid Holidays
Bonus Programs
+5
Director, Engineering – Release Engineering, DevOps & SRE
Director, Engineering – Release Engineering, DevOps & SRE

Data Direct Networks • North Carolina

Hybrid
USD 250,000 - 300,000
Vacation plans
Paid holidays
Bonus programs
+5