Real-world incidents have showcased the importance of DR and BC, demonstrating both successes where companies adeptly handled crises, and failures where lack of preparation led to significant repercussions.
Successful DR and BC Initiatives:
- Delta Airlines (2016):
- IncidentA power outage at Delta’s operations center led to the grounding of 2,000 flights over a period of three days.
- ResponseDespite the initial disruption, Delta’s transparent communication strategy shone through. They kept passengers informed via social media, offered compensation for affected travelers, and worked tirelessly to resume operations.
- OutcomeThough the incident was a setback, Delta’s handling of the situation, especially their communication, minimized reputational damage and showcased the importance of having a robust BC communication plan.
- Maersk (2017):
- IncidentThe shipping giant was hit by the NotPetya ransomware, affecting its global operations and shutting down port terminals.
- ResponseMaersk swiftly initiated its BC plan. Instead of paying the ransom, they decided to reinstall their entire infrastructure. Over ten days, 4,000 servers and 45,000 PCs were reinstalled.
- OutcomeMaersk’s efficient response and commitment to not ceding to the ransomware’s demands earned praise. They were transparent about the situation, which bolstered their reputation despite the incident.
Failed DR and BC Initiatives:
- BlackBerry (2011):
- IncidentA core switch failure led to a massive service disruption for BlackBerry users worldwide, lasting several days.
- ResponseBlackBerry’s reaction was slow, and initial communication to users was lacking in clarity. It took them a significant amount of time to restore services fully.
- OutcomeThe already struggling brand suffered immensely due to this prolonged outage, hastening its decline in the smartphone market. The incident showcased the importance of not only having a DR plan but also testing and ensuring it works efficiently.
- Knight Capital (2012):
- IncidentA software glitch in Knight Capital’s trading system triggered millions of unintended trades, leading to a loss of $440 million within 45 minutes.
- ResponseThe company was unable to halt the malfunctioning system immediately. While they eventually stopped the erroneous trades, the financial damage was done.
- OutcomeThe firm’s stock plummeted, and they were forced to secure emergency financing to stay afloat. Later, they merged with another company. The incident emphasized the need for efficient DR in IT systems, especially where high-frequency operations are involved.
Both successful and unsuccessful cases underline the importance of preparation, swift response, clear communication, and the continuous refinement of DR and BC plans. They also emphasize the potential financial and reputational risks organizations face when they are unprepared for disruptions.
Key terms in plain language
Open a term for a concise explanation of language used on this page.
Colocation
Placing customer-owned servers and network equipment in a professionally operated data center that provides power, cooling, physical security, and connectivity.
Content Delivery Network (CDN)
A distributed system that serves website or application content from locations closer to users, improving speed, resilience, and capacity.
Cloud Computing
Computing resources—such as applications, servers, storage, or databases—delivered from remote infrastructure and scaled as requirements change.
Infrastructure as a Service (IaaS)
Cloud-based servers, storage, and networking that customers configure and manage without owning the underlying data-center hardware.
Bandwidth
The amount of data a connection can carry in a given time, usually measured in Mbps or Gbps. More bandwidth supports more users, devices, and simultaneous applications.
Disaster Recovery (DRaaS)
A plan and service for restoring applications, data, and operations after an outage or disruption. DRaaS provides recovery infrastructure through a managed cloud service.