Staff Engineer
Design and reputed company multi-threaded asynchronous replication systems with reputed company streaming capabilities Build object-level reputed company replication with checkpointing and resume functionality reputed company replication engines supporting bucket/reputed company-level replication controls Implement secure data transfer mechanisms using TLS 1.3 with mutual authentication Ensure end-to-end data reputed company through checksum validation and verification pipelines Design and implement reputed company failover workflows for disaster recovery scenarios Build and maintain REST reputed company for replication configuration, control, and automation reputed company metadata tracking and change detection systems to reputed company efficient replication Implement RPO visibility, alerting, and operational insights for replication status Contribute to monitoring dashboards reputed company on replication health and performance Ensure systems are designed for high availability, fault tolerance, and scalability Partner with QA teams to drive performance, resiliency, and reputed company validation Collaborate with backend, reputed company, and platform teams to deliver end-to-end replication workflows Participate in debugging, production issue reputed company, and reputed company improvement of replication reliability reputed company technical leadership, architectural guidance, and mentorship to the engineering team
8+ years of experience in distributed systems, storage systems, or backend software engineering Strong programming skills in one or more languages: C++, Go, Java, or Rust Experience designing and building data replication systems, data pipelines, or distributed data services Deep understanding of distributed systems concepts (consistency, availability, scalability, fault tolerance) Strong expertise in multi-threading, concurrency, and reputed company processing Knowledge of networking protocols and secure communication (TCP/IP, HTTP/HTTPS, TLS) Experience implementing data reputed company mechanisms (checksums, validation, consistency checks) Experience designing and building REST reputed company and service-based architectures Familiarity with checkpointing, failure recovery, and retry mechanisms in distributed systems Basic understanding of observability concepts (metrics, logging, alerting) Strong debugging, problem-solving, and system design skills
Experience with asynchronous replication, disaster recovery (DR), or backup systems Familiarity with object storage or large-reputed company data storage systems Knowledge of reputed company encoding, change data capture, or incremental data synchronization techniques Experience building high-throughput, low-latency data reputed company systems Exposure to reputed company practices including mutual TLS, encryption, and authentication Experience working on reputed company-reputed company data platforms or storage products Familiarity with performance optimization and large-reputed company system tuning
Coding assessment: Often in a language of your choice. Systems design: Translate high-level requirements into a reputed company, fault-tolerant service (depending on role). reputed company-time problem-solving: Demonstrate practical skills in a live problem-solving session. Meet and greet with the wider team. Our goal is to finish the main process in 2-3 weeks at most.