A company needs a Disaster Recovery (DR) strategy with an RTO of 30 minutes and an RPO of 15 minutes. They want to minimize costs while having a scaled-down version of their core environment always running in a second region. Which DR strategy should they use?
Trap 1: Backup and Restore.
Backup and Restore is the most cost-effective but slowest DR strategy. It involves restoring backups after a disaster occurs, which can take hours or days. This would fail to meet the 30-minute RTO requirement, as provisioning new infrastructure and restoring large datasets from backups cannot typically be completed within that aggressive time window.
Trap 2: Pilot Light.
In a Pilot Light strategy, only the most critical core elements (like the database) are kept running and up-to-date in the DR region. Other resources, like application servers, are provisioned only during a disaster. While faster than Backup and Restore, it often takes longer than 30 minutes to scale and configure the full environment.
Trap 3: Multi-site Active-Active.
Multi-site Active-Active involves running the full production environment in two or more regions simultaneously and splitting traffic between them. This provides the lowest RTO and RPO (near zero) but is the most expensive strategy because you are paying for full duplicate infrastructure at all times, which does not align with the goal of minimizing costs.
- A
Backup and Restore.
Why it fails: Backup and Restore is the most cost-effective but slowest DR strategy. It involves restoring backups after a disaster occurs, which can take hours or days. This would fail to meet the 30-minute RTO requirement, as provisioning new infrastructure and restoring large datasets from backups cannot typically be completed within that aggressive time window.
- B
Pilot Light.
Why it fails: In a Pilot Light strategy, only the most critical core elements (like the database) are kept running and up-to-date in the DR region. Other resources, like application servers, are provisioned only during a disaster. While faster than Backup and Restore, it often takes longer than 30 minutes to scale and configure the full environment.
- C
Warm Standby.
Warm Standby keeps a minimized version of the full environment running in the second region. This ensures that the application is always ready to handle traffic, and only needs to be scaled up to meet production loads during a failover. This strategy easily fits the 30-minute RTO and 15-minute RPO requirements while remaining cost-conscious.
- D
Multi-site Active-Active.
Why it fails: Multi-site Active-Active involves running the full production environment in two or more regions simultaneously and splitting traffic between them. This provides the lowest RTO and RPO (near zero) but is the most expensive strategy because you are paying for full duplicate infrastructure at all times, which does not align with the goal of minimizing costs.