The sequence
- Confirm the failure and its scope: hardware, storage, OS or application
- Decide whether to repair or fail over — repairing first often costs the most time
- Boot the most recent verified image or begin a restore
- Validate data currency against the RPO
- Bring dependent systems up in the documented order
- Communicate status and the interim process to staff
- Rebuild the primary and plan the failback window
- Review what the failure revealed about the design