1. Disaster Recovery for PaaS 2.0

Disaster Recovery 2.0 roles and responsibilities

Version:

This topic describes the roles and responsibilities for the different stages in the disaster recovery process for PaaS 2.0. Roles and responsibilities are described using the RACI model, where R = Responsible, A = Accountable, C = Consultant, and I = Informed.

Customer-owned Business continuity plan

The Business continuity plan is the responsibility of the customer or partner, with Sitecore providing expertise.

Task IDTask definition/descriptionSitecore rolesCustomer/Partner roles
1Selection of Sitecore DR servicesI/CR/A
2Creation of an end-to-end business continuity planI/CR/A
3End-to-end testing of business continuity planI/CR/A
4Execution of business continuity planI/CR/A

Disaster recovery architecture and testing process

Sitecore will test the following items as part of our Disaster Recovery 2.0 service.

Task IDTask definition/descriptionSitecore rolesCustomer/Partner roles
1Sitecore DR 2.0 pattern creationR/AI/C
2Sitecore DR 2.0 pattern maintenanceR/AI/C
3Sitecore DR 2.0 pattern supportR/AI/C
4Sitecore DR 2.0 initial pattern testing and validationR/AI/C
5Sitecore DR 2.0 initial pattern testing and validation with Sitecore XM v10.3.1 or later, and XP v10.3.1 or laterR/AI/C

Disaster Recovery end-to-end process

The activities described in this section are intended to guide the Disaster Recovery implementation and failover scenarios.

DR Deployment

Sitecore Managed Cloud provides the following steps as part of the DR 2.0 offering.

Task IDTask definition/descriptionSitecore rolesCustomer/Partner roles
1Deploy a new Sitecore environment in the secondary region and scale down the Web Apps and Redis.R/AI/C
2Sync the data between the primary and secondary Azure SQL using Failover Groups.R/AI/C
3Sync the sizes/tiers of all Azure resourcesR/AI/C
4Sync the file contents of all Web AppsR/AI/C
5Set up Front Door to switch between primary CD and outage pageR/AI/C
6Set up email alerts to notify the Sitecore Managed Cloud operations team when the availability tests failR/AI/C

DR invocation

The DR failover Initiation process starts when you've notified Sitecore of your intent to perform a complete DR failover to the secondary Azure region.

Task IDTask definition/descriptionSitecore rolesCustomer/Partner roles
1Switch Front Door traffic to the maintenance pageR/AC/I
2Disable environments synchronizationR/AC/I
3Perform SQL failoverR/AC/I
4Restore App Service backupsR/AC/I
5Scale up Disaster Recovery environmentR/AC/I
6Verify Disaster Recovery healthR/AC/I
7Switch Front Door traffic to the Disaster Recovery siteR/AC/I
8Perform post-DR recovery updates to the Sitecore applicationC/IR/A
9Perform post-DR recovery updates to non-standard Azure services or componentsC/IR/A

DR failback

The failback steps for DR 2.0 are similar to the DR invocation steps noted previously. However, we assume there were no changes to the App services files or configuration during the failover state - that is, there were no code updates or deployments. Therefore, there are no steps to disable environment synchronization, restore app service backups, or to scale up the DR environment.

Task IDTask definition/descriptionSitecore rolesCustomer/Partner roles
1Switch Front Door traffic to the maintenance pageR/AC/I
2Perform SQL failoverR/AC/I
3Verify Primary healthR/AC/I
4Switch Front Door traffic to the Production page.R/AC/I
5Perform post-DR recovery updates to the Sitecore applicationC/IR/A
6Perform post-DR recovery updates to non-standard Azure services or componentsC/IR/A
Note

If any changes have been applied to the App services in the disaster recovery region (while failed over), these changes must be reapplied to the App services when the failback process has been completed.

If you have suggestions for improving this article, let us know!