Staff Replication Development Engineer
Overview
As Staff Replication Development Engineer, you will lead the design and development of the replication engine for the Infinia AI Data Platform. You will build enterprise-grade asynchronous replication capabilities to enable reliable disaster recovery for large-scale data systems. You will shape high-performance replication pipelines, secure data transfer, and end-to-end integrity, partnering with cross-functional teams to deliver a scalable, resilient foundation. This role offers the chance to influence DR strategy and data availability at scale.
ResponsibilitiesDesign and develop multi-threaded asynchronous replication systems with parallel streaming
Key requirements8+ years of experience in distributed systems, storage systems, or backend software engineeringStrong programming skills in C++, Go, Java, or RustExperience designing and building data replication systems, data pipelines, or distributed data servicesDeep understanding of distributed systems concepts (consistency, availability, scalability, fault tolerance)Strong expertise in multi-threading, concurrency, and parallel processingKnowledge of networking protocols and secure communication (TCP/IP, HTTP/HTTPS, TLS)Experience implementing data integrity mechanisms (checksums, validation, consistency checks)Experience designing and building REST APIs and service-based architecturesFamiliarity with checkpointing, failure recovery, and retry mechanisms in distributed systemsBasic understanding of observability concepts (metrics, logging, alerting)Strong debugging, problem-solving, and system design skillsstrong communicationleadership and mentorshipproblem-solvingasynchronous replicationoffset/checkpointingdelta encoding / change data capture