DDN Storage Logo

DDN Storage

Staff Replication Development Engineer

Posted Yesterday
Be an Early Applicant
Remote or Hybrid
Hiring Remotely in North Carolina, USA
Senior level
Remote or Hybrid
Hiring Remotely in North Carolina, USA
Senior level
Leads the design and development of high-performance asynchronous replication systems for a large-scale AI data platform. Responsibilities include distributed replication pipelines, delta synchronization, secure data transfer, integrity validation, REST APIs, failover workflows, observability, and disaster recovery capabilities. Provides architectural guidance and mentorship while partnering with QA, security, backend, and platform teams to validate scalability, resilience, and performance.
The summary above was generated by AI

DDN is seeking a Staff Replication Development Engineer to lead the design and development of the replication engine for the Infinia AI Data Platform. This role focuses on building enterprise-grade asynchronous replication capabilities that enable reliable and secure disaster recovery for large-scale data systems.

You will work on developing high-performance replication pipelines, efficient data synchronization mechanisms, and secure data transfer systems. This role requires deep expertise in distributed systems and strong technical leadership to deliver a scalable and resilient replication foundation.

 
Key Responsibilities
  • Design and develop multi-threaded asynchronous replication systems with parallel streaming capabilities

  • Build object-level delta replication with checkpointing and resume functionality

  • Develop replication engines supporting bucket/share-level replication controls

  • Implement secure data transfer mechanisms using TLS 1.3 with mutual authentication

  • Ensure end-to-end data integrity through checksum validation and verification pipelines

  • Design and implement manual failover workflows for disaster recovery scenarios

  • Build and maintain REST APIs for replication configuration, control, and automation

  • Develop metadata tracking and change detection systems to enable efficient replication

  • Implement RPO visibility, alerting, and operational insights for replication status

  • Contribute to monitoring dashboards focused on replication health and performance

  • Ensure systems are designed for high availability, fault tolerance, and scalability

  • Partner with QA teams to drive performance, resiliency, and scale validation

  • Collaborate with backend, security, and platform teams to deliver end-to-end replication workflows

  • Participate in debugging, production issue resolution, and continuous improvement of replication reliability

  • Provide technical leadership, architectural guidance, and mentorship to the engineering team

 
Required Qualifications
  • 8+ years of experience in distributed systems, storage systems, or backend software engineering

  • Strong programming skills in one or more languages: C++, Go, Java, or Rust

  • Experience designing and building data replication systems, data pipelines, or distributed data services

  • Deep understanding of distributed systems concepts (consistency, availability, scalability, fault tolerance)

  • Strong expertise in multi-threading, concurrency, and parallel processing

  • Knowledge of networking protocols and secure communication (TCP/IP, HTTP/HTTPS, TLS)

  • Experience implementing data integrity mechanisms (checksums, validation, consistency checks)

  • Experience designing and building REST APIs and service-based architectures

  • Familiarity with checkpointing, failure recovery, and retry mechanisms in distributed systems

  • Basic understanding of observability concepts (metrics, logging, alerting)

  • Strong debugging, problem-solving, and system design skills

 
Preferred Qualifications
  • Experience with asynchronous replication, disaster recovery (DR), or backup systems

  • Familiarity with object storage or large-scale data storage systems

  • Knowledge of delta encoding, change data capture, or incremental data synchronization techniques

  • Experience building high-throughput, low-latency data movement systems

  • Exposure to security practices including mutual TLS, encryption, and authentication

  • Experience working on enterprise-scale data platforms or storage products

  • Familiarity with performance optimization and large-scale system tuning

Similar Jobs

Yesterday
Remote or Hybrid
North Carolina, USA
Senior level
Senior level
Artificial Intelligence • Machine Learning • Software • Analytics
Leads the design and development of a highly available asynchronous replication engine for the Infinia AI Data Platform. Responsibilities include multi-threaded replication pipelines, delta synchronization, secure TLS data transfer, integrity validation, failover workflows, REST APIs, metadata tracking, observability, and performance validation. Provides architectural guidance, mentors engineers, collaborates across teams, and resolves production issues while ensuring scalability, resiliency, and disaster recovery capabilities.
Top Skills: C++GoHttp/HttpsJavaMutual TlsRest ApisRustTcp/IpTls 1.3
32 Minutes Ago
Remote or Hybrid
45K-85K Annually
Junior
45K-85K Annually
Junior
Artificial Intelligence • Fintech • Insurance • Marketing Tech • Software • Analytics
Handle inbound calls and warm leads, assess customers’ insurance needs, recommend appropriate property and casualty coverages, and convert prospects into policyholders. Complete paid training and obtain a Property and Casualty license if needed. Maintain strong customer communication, persuasion, organization, and computer skills while working remotely on assigned schedules, including one weekend day.
Top Skills: Cable InternetDsl InternetFiber InternetPcWired High-Speed Internet
An Hour Ago
Remote or Hybrid
United States
111K-180K Annually
Senior level
111K-180K Annually
Senior level
Fintech • Information Technology • Insurance • Financial Services • Big Data Analytics
Leads architecture, modernization, optimization, and reliability initiatives for mainframe CICS, MQ, and z/OS Connect environments. Provides technical direction across development and operations teams, establishes governance and change processes, tunes performance using telemetry, resolves incidents, and develops modernization roadmaps. Collaborates with stakeholders and enterprise architects to deliver secure, scalable, high-availability solutions while evaluating automation, cloud integration, and AI technologies.
Top Skills: AnsibleCicsCobolDevOpsIbm MqIbm Z/OsOpenshiftPythonRed Hat Ansible Automation PlatformZ/Os ConnectZlinux

What you need to know about the Charlotte Tech Scene

Ranked among the hottest tech cities in 2024 by CompTIA, Charlotte is quickly cementing its place as a major U.S. tech hub. Home to more than 90,000 tech workers, the city’s ecosystem is primed for continued growth, fueled by billions in annual funding from heavyweights like Microsoft and RevTech Labs, which has created thousands of fintech jobs and made the city a go-to for tech pros looking for their next big opportunity.

Key Facts About Charlotte Tech

  • Number of Tech Workers: 90,859; 6.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lowe’s, Bank of America, TIAA, Microsoft, Honeywell
  • Key Industries: Fintech, artificial intelligence, cybersecurity, cloud computing, e-commerce
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (CED)
  • Notable Investors: Microsoft, Google, Falfurrias Management Partners, RevTech Labs Foundation
  • Research Centers and Universities: University of North Carolina at Charlotte, Northeastern University, North Carolina Research Campus

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account