We are looking for a Lead Engineer/Manager to join our newly formed Backup & Disaster Recovery group. The role needs strong design, development, delivery and leadership skills for our Backup & Disaster Recovery capabilities across our platforms. This is a hands-on technical leadership role for someone who has built or worked deeply on backup, restore, replication/continuous data protection (CDP) technologies, and understands what it takes to help customers meet aggressive RPO (Recovery Point Objective) and RTO (Recovery Time Objective) targets.
You will own a critical part of our product ecosystem, he systems that ensure customer data and workloads can be reliably protected, replicated, and recovered ,while also mentoring and growing a small team of engineers.
What You'll Do- Design, develop, and own core Back up & DR features spanning backup, restore, replication/CDP, and management features for PCD environments
- Drive engineering decisions to meet the business objectives including performance, reliability, cost, platform compatibility.
- Architect and build storage-driver interface for various storage vendors, ensuring consistency, integrity, and recoverability of data across failure scenarios.
- Architect and build storage independent replication technology for consistency, integrity, performance recoverability across data centers.
- Lead and mentor a team of engineers, QA, UI and UX, providing technical direction, SDLC, code/design reviews, and career guidance.
- Collaborate with Product and Support teams to define requirements, validate scenarios, troubleshoot production issues and train the involved parties
- Drive root-cause analysis for failures and lead the resolution of complex, distributed-systems issues.
- Contribute to architecture and roadmap discussions for the storage and data protection stack.
- Establish and enforce engineering best practices testing, automation, documentation, and operational readiness.
- 8+ years of software/systems engineering experience, with a significant portion focused on Disaster Recovery, Backup, Restore, or Replication technologies.
- Prior experience leading or mentoring a team, even informally (tech lead, module owner, etc.).
- Strong understanding of storage stack including FC, iSCSI, NVMe, RBD, DRBD, RPO/RTO concepts and trade-offs in real-world architectures.
- Hands-on experience with Continuous Data Protection (CDP), snapshot-based or log-based replication, and storage systems (block, file, or object).
- Solid background in systems-level software development — strong coding skills in one or more of: Rust, C/C++, Go, Java, or Python.
- Experience working with distributed systems, storage stacks, or virtualization/cloud infrastructure.
- Strong debugging and problem-solving skills for complex, data-critical, low-tolerance-for-error systems.
- Excellent communication skills and comfort working cross-functionally with Product and Support teams.
- Direct experience with enterprise DR/backup platforms.
- Exposure to CDP-based DR-as-a-service offerings (e.g., VMware Site Recovery Manager , HPE Zerto Software) with an understanding of how RPO/RTO targets are engineered into orchestrated failover pipelines.
- Familiarity with backup-target/appliance architectures (e.g.VMware SRM, Site REcovery Manager) and how they integrate into a broader data-protection stack.
- Exposure to hypervisor or hyperconverged infrastructure environments.
- Experience with public cloud DR/replication services (AWS, Azure, GCP).
- Familiarity with storage protocols (iSCSI, NFS, SMB, FC, iSCSI, NVMe, RBD, DRBD) or software-defined storage.
Benefits and Perks:
Employees today are looking for companies that truly care and recognise their whole person. Platform 9's benefits and perks have been carefully designed to ensure that we take care of an employee's emotions and physical well-being. Many of our benefits extend to families, who form a significant part of our well-being at work
Please note that benefits change by country.
- Competitive Compensation and Equity
- Medical Healthcare for you and your family
- Hybrid Work Model
- Wellness Benefits
- Professional Development/Global certifications
- Reward and Recognition Programs
- Team Building Activities
Our benefits have been carefully selected, keeping in mind employees requirements and personal situations now and for the future
Skills Required
- 8+ years of software or systems engineering experience
- Significant experience focused on disaster recovery, backup, restore, or replication technologies
- Prior experience leading or mentoring an engineering team
- Strong understanding of storage technologies and protocols, including FC, iSCSI, NVMe, RBD, and DRBD
- Understanding of RPO and RTO concepts and trade-offs
- Hands-on experience with Continuous Data Protection, snapshot-based replication, or log-based replication
- Experience with block, file, or object storage systems
- Strong coding skills in Rust, C, C++, Go, Java, or Python
- Experience with distributed systems, storage stacks, or virtualization/cloud infrastructure
- Strong debugging and problem-solving skills for complex data-critical systems
- Excellent communication skills and ability to collaborate cross-functionally with Product and Support teams
- Direct experience with enterprise disaster recovery or backup platforms
- Exposure to CDP-based disaster-recovery-as-a-service offerings such as VMware Site Recovery Manager or HPE Zerto
- Familiarity with backup-target or appliance architectures
- Exposure to hypervisor or hyperconverged infrastructure environments
- Experience with public-cloud disaster recovery or replication services, including AWS, Azure, or GCP
- Familiarity with NFS, SMB, software-defined storage, or related storage protocols
What We Do
Platform9 is the open distributed cloud company, offering the power of the public cloud on infrastructure of customers’ choice—powered by Kubernetes and cloud-native technologies. Public clouds are walled gardens, and DIY is difficult and time-consuming. Platform9 offers a third option—an open and faster option—enabling a better way to go cloud-native. Platform9’s service powers 40K+ nodes across private, public and edge clouds. Innovative enterprises like Juniper, Kingfisher Plc, Mavenir, Redfin and Cloudera achieve 4x faster time-to-market, up to 90% reduction in operational costs, and 99.9% uptime. Platform9 is an inclusive, globally distributed company, backed by leading investors.








