Backup and Disaster Recovery Questions

Keeping data durable and recoverable when systems fail: backup design (full, incremental, differential, and snapshot strategies; point-in-time recovery), backup verification and restore testing, retention and archival policy (including compliance retention and legal holds), encryption and key management for backups, and disaster-recovery planning measured against recovery-time and recovery-point objectives (RTO/RPO). Tests whether a candidate can design a backup strategy that actually restores, choose the right retention tier for a business's downtime and data-loss tolerance, and operate backup systems safely under failure, compliance, and ransomware threats. Distinct from code-level fault-tolerance patterns (circuit breakers, retries, bulkheads) and multi-region failover architecture, which belong to high-availability-and-disaster-recovery.

MediumSystem Design
117 practiced

Design a monitoring and alerting plan for backup systems: what would tell you a backup silently failed, and what would you put in front of an on-call engineer versus what a manager needs to see on a dashboard?

MediumTechnical
78 practiced

Explain the differences between full, incremental, and differential backups. For each type, describe how the restore operation actually works, typical storage and I/O characteristics, how complex the recovery chain gets, and a realistic scenario where you'd prefer that type over the others.

EasyTechnical
101 practiced

Explain the difference between filesystem or volume snapshots and traditional backups. Discuss scenarios where snapshots are sufficient (e.g., rapid rollback) and cases where snapshots are not a replacement for backups (e.g., offsite, long-term retention, or provider failures).

EasyTechnical
67 practiced

A small company needs a backup retention policy: daily backups kept 30 days, weekly backups kept 6 months, yearly backups kept 7 years. Walk through how you'd actually structure and enforce a policy like this so it holds up over time, including what should happen when a legal request requires keeping something past its normal deletion date.

HardSystem Design
71 practiced

Design a backup and disaster recovery system for 200 TB of production block storage spread across multiple data centers. Requirements: daily incremental backups, weekly full backups, RTO < 4 hours for critical datasets, RPO < 1 hour for highest-priority data, and retention/compliance policies. Detail architecture (snapshot vs block-level copy vs agent), cataloging, verification, restore runbooks, and how SREs should operate and test the system.

Unlock Full Question Bank

Get access to all Backup and Disaster Recovery interview questions and detailed answers.

Sign in to Continue

Join thousands of developers preparing for their dream job.