Senior Replication Engineer - Scalable Data Platform
Listed on 2026-08-04
-
Software Development
Data Engineering
We are on our way to being the first company to power 1 MILLION GPUs and want world-class talent to join our amazing team!
The world is moving faster than ever, and yet, it will never move this slowly again. We are at the forefront of an incredible technological revolution but, at its core, it is fueled by incredible people. People like you.
1,000
Global Employees
11k
Happy Customers
16
International Offices
We are the world’s leading data intelligence platform that reliably accelerates massive datasets for actionable real-time insights. Join our team to help the best and brightest minds tackle the world’s biggest challenges in business, science, medicine, academia and government.
Do What Can’t be DoneFor the past 20 years, our team has kept us at the forefront of storage technology and has provided the foundation for enabling researchers to push the limits of “what can be done.”
These innovations take research and discovery to the next level, enabling them to discover cures to disease, observe global warming patterns, model innovative automotive and aerospace designs, discover new sources of energy, make communities safer, and accelerate business results across a wide variety of industries.
At DDN, we understand our customers’ diverse needs. Whether you’re a data scientist, IT professional, executive, or researcher, our solutions empower you with cutting‑edge technology and unparalleled support.
Highly Competitive Vacation Plans
Paid Holidays
Bonus Programs
Tuition Reimbursement
Employee Referral Program
Excellent Medical, Dental and Vision Benefits
Paid Leave Programs
Anniversary and Recognition Awards
LocationEmployment Type
Full time
Location TypeHybrid
DDN is seeking a Staff Replication Development Engineer to lead the design and development of the replication engine for the Infinia AI Data Platform. This role focuses on building enterprise‑grade asynchronous replication capabilities that enable reliable and secure disaster recovery for large‑scale data systems.
You will work on developing high‑performance replication pipelines, efficient data synchronization mechanisms, and secure data transfer systems. This role requires deep expertise in distributed systems and strong technical leadership to deliver a scalable and resilient replication foundation.
Key ResponsibilitiesDesign and develop multi‑threaded asynchronous replication systems with parallel streaming capabilities
Build object‑level delta replication with checkpointing and resume functionality
Implement secure data transfer mechanisms using TLS 1.3 with mutual authentication
Ensure end‑to‑end data integrity through checksum validation and verification pipelines
Design and implement manual failover workflows for disaster recovery scenarios
Build and maintain REST APIs for replication configuration, control, and automation
Develop metadata tracking and change detection systems to enable efficient replication
Implement RPO visibility, alerting, and operational insights for replication status
Contribute to monitoring dashboards focused on replication health and performance
Ensure systems are designed for high availability, fault tolerance, and scalability
Partner with QA teams to drive performance, resiliency, and scale validation
Collaborate with backend, security, and platform teams to deliver end‑to‑end replication workflows
Participate in debugging, production issue resolution, and continuous improvement of replication reliability
Provide technical leadership, architectural guidance, and mentorship to the engineering team
Required Qualifications8+ years of experience in distributed systems, storage systems, or backend software engineering
Strong programming skills in one or more languages: C++, Go, Java, or Rust
Experience designing and building data replication systems, data pipelines, or distributed data services
Deep understanding of distributed systems concepts (consistency, availability, scalability, fault tolerance)
Strong expertise in multi‑threading, concurrency, and parallel processing
Knowledge of networking protocols and secure communication (TCP/IP, HTTP/HTTPS, TLS)
Experience…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).