Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Replication Development Engineer

DataDirect Networks

DDN is seeking a Staff Replication Development Engineer to lead the design and development of the replication engine for the Infinia AI Data Platform. This role focuses on building enterprise-grade asynchronous replication capabilities that enable reliable and secure disaster recovery for large-scale data systems.You will work on developing high-performance replication pipelines, efficient data synchronization mechanisms, and secure data transfer systems. This role requires deep expertise in distributed systems and strong technical leadership to deliver a scalable and resilient replication foundation.Key ResponsibilitiesDesign and develop multi-threaded asynchronous replication systems with parallel streaming capabilitiesBuild object-level delta replication with checkpointing and resume functionalityDevelop replication engines supporting bucket/share-level replication controlsImplement secure data transfer mechanisms using TLS 1.3 with mutual authenticationEnsure end-to-end data integrity through checksum validation and verification pipelinesDesign and implement manual failover workflows for disaster recovery scenariosBuild and maintain REST APIs for replication configuration, control, and automationDevelop metadata tracking and change detection systems to enable efficient replicationImplement RPO visibility, alerting, and operational insights for replication statusContribute to monitoring dashboards focused on replication health and performanceEnsure systems are designed for high availability, fault tolerance, and scalabilityPartner with QA teams to drive performance, resiliency, and scale validationCollaborate with backend, security, and platform teams to deliver end-to-end replication workflowsParticipate in debugging, production issue resolution, and continuous improvement of replication reliabilityProvide technical leadership, architectural guidance, and mentorship to the engineering teamRequired Qualifications8+ years of experience in distributed systems, storage systems, or backend software engineeringStrong programming skills in one or more languages: C++, Go, Java, or RustExperience designing and building data replication systems, data pipelines, or distributed data servicesDeep understanding of distributed systems concepts (consistency, availability, scalability, fault tolerance)Strong expertise in multi-threading, concurrency, and parallel processingKnowledge of networking protocols and secure communication (TCP/IP, TLS)Experience implementing data integrity mechanisms (checksums, validation, consistency checks)Experience designing and building REST APIs and service-based architecturesFamiliarity with checkpointing, failure recovery, and retry mechanisms in distributed systemsBasic understanding of observability concepts (metrics, logging, alerting)Strong debugging, problem-solving, and system design skillsPreferred QualificationsExperience with asynchronous replication, disaster recovery (DR), or backup systemsFamiliarity with object storage or large-scale data storage systemsKnowledge of delta encoding, change data capture, or incremental data synchronization techniquesExperience building high-throughput, low-latency data movement systemsExposure to security practices including mutual TLS, encryption, and authenticationExperience working on enterprise-scale data platforms or storage productsFamiliarity with performance optimization and large-scale system tuningLocationSanta ClaraEmployment TypeFull timeLocation TypeHybridDepartmentDepartmentEngineeringEngineeringInfinia Engineering

Vacancy posted more than 2 months ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Replication Development Engineer. Be the first to apply!