US12481545B2ActiveUtilityA1

Detection and recovery of platform specific replication failures in SDNAS solution on a storage system

Assignee: DELL PRODUCTS LPPriority: Sep 14, 2023Filed: Sep 14, 2023Granted: Nov 25, 2025
Est. expirySep 14, 2043(~17.1 yrs left)· nominal 20-yr term from priority
G06F 11/0706G06F 11/1402G06F 2201/805G06F 11/0727G06F 11/2094
56
PatentIndex Score
0
Cited by
2
References
18
Claims

Abstract

An example methodology includes monitoring a replication workflow execution on a storage system and, responsive to a detection of a failure of the replication workflow, determining whether the failure is in a Software Defined Network Attached Storage (SDNAS) or in an underlying storage array platform. The method also includes, responsive to a determination that the failure is in the SDNAS, validating health of the SDNAS for performing the replication workflow and, responsive to validating the health of the SDNAS, performing recovery of the failed replication workflow. The method further includes, responsive to a determination that the failure is in the underlying storage array platform, determining that the underlying storage array platform can be rolled back to a previously known good state for performing the replication workflow, performing a rollback of the underlying storage array platform to the previously known good state, and performing the recovery of the failed replication workflow.

Claims

exact text as granted — not AI-modified
What is claimed is: 
     
         1 . A method comprising:
 monitoring, by a computing device, a replication workflow execution on a storage system; and   responsive to a detection of a failure of the replication workflow, by the computing device:
 determining whether the failure is in a Software Defined Network Attached Storage (SDNAS) or in an underlying storage array platform; and 
 responsive to a determination that the failure is in the SDNAS:
 validating a health of the SDNAS for performing the replication workflow; and 
 responsive to validating the health of the SDNAS, performing a recovery of the failed replication workflow that includes attempting forward execution of the replication workflow from a failed replication operation in the replication workflow. 
 
   
     
     
         2 . The method of  claim 1 , wherein the validating a health of the SDNAS for performing the replication workflow includes checking one or more components of the SDNAS which are needed to perform the replication workflow. 
     
     
         3 . The method of  claim 1 , further comprising, responsive to failing to validate the health of the SDNAS for performing the replication workflow, reporting an original error associated with the failure of the replication workflow in the SDNAS. 
     
     
         4 . The method of  claim 1 , further comprising reporting a result of the recovery of the failed replication workflow. 
     
     
         5 . The method of  claim 1 , wherein determining whether the failure is in the underlying storage array platform includes determining whether the failure is in a data replication facility of the underlying storage array platform. 
     
     
         6 . The method of  claim 1 , further comprising:
 responsive to a determination that the failure is in the underlying storage array platform:
 determining whether the underlying storage array platform can be rolled back to a previously known good state for performing the replication workflow; and 
 responsive to a determination that the underlying storage array platform can be rolled back to the previously known good state:
 performing a rollback of the underlying storage array platform to the previously known good state; and 
 performing the recovery of the failed replication workflow. 
 
   
     
     
         7 . The method of  claim 6 , wherein the determining whether the underlying storage array platform can be rolled back to the previously known good state includes determining whether a data replication facility of the underlying storage array platform is available. 
     
     
         8 . The method of  claim 6 , further comprising, responsive to the determination that the failure is in the underlying storage array platform, responsive to a determination that the replication workflow is an unplanned failover, performing the recovery of the failed replication workflow without the rollback of the underlying storage array platform to the previously known good state. 
     
     
         9 . The method of  claim 7 , further comprising, responsive to a determination that the underlying storage array platform cannot be rolled back to the previously known good state, reporting an original error associated with the failure of the replication workflow in the underlying storage array platform. 
     
     
         10 . The method of  claim 6 , wherein the performing the recovery of the failed replication workflow responsive to the determination that the underlying storage array platform can be rolled back to the previously known good state includes retrying the failed replication workflow. 
     
     
         11 . The method of  claim 10 , further comprising reporting a result of the retrying of the failed replication workflow. 
     
     
         12 . A computing device comprising:
 one or more non-transitory machine-readable mediums configured to store instructions; and   one or more processors configured to execute the instructions stored on the one or more non-transitory machine-readable mediums, wherein execution of the instructions causes the one or more processors to carry out a process comprising:
 monitoring a replication workflow execution on a storage system; and 
 responsive to a detection of a failure of the replication workflow:
 determining whether the failure is in a Software Defined Network Attached Storage (SDNAS) or in an underlying storage array platform; and 
 responsive to a determination that the failure is in the SDNAS:
 validating a health of the SDNAS for performing the replication workflow; and 
 responsive to validating the health of the SDNAS, performing a recovery of the failed replication workflow that includes attempting forward execution of the replication workflow from a failed replication operation in the replication workflow. 
 
 
   
     
     
         13 . The computing device of  claim 12 , wherein the validating a health of the SDNAS for performing the replication workflow includes checking one or more components of the SDNAS which are needed to perform the replication workflow. 
     
     
         14 . The computing device of  claim 12 , wherein determining whether the failure is in the underlying storage array platform includes determining whether the failure is in a data replication facility of the underlying storage array platform. 
     
     
         15 . The computing device of  claim 12 , wherein the process further comprises:
 responsive to a determination that the failure is in the underlying storage array platform:
 determining whether the underlying storage array platform can be rolled back to a previously known good state for performing the replication workflow; and 
 responsive to a determination that the underlying storage array platform can be rolled back to the previously known good state:
 performing a rollback of the underlying storage array platform to the previously known good state; and 
 performing the recovery of the failed replication workflow. 
 
   
     
     
         16 . The computing device of  claim 15 , wherein the determining whether the underlying storage array platform can be rolled back to the previously known good state includes determining whether a data replication facility of the underlying storage array platform is available. 
     
     
         17 . The computing device of  claim 15 , wherein the process further comprises, responsive to the determination that the failure is in the underlying storage array platform, responsive to a determination that the replication workflow is an unplanned failover, performing the recovery of the failed replication workflow without the rollback of the underlying storage array platform to the previously known good state. 
     
     
         18 . A non-transitory machine-readable medium encoding instructions that when executed by one or more processors cause a process to be carried out, the process including:
 monitoring a replication workflow execution on a storage system; and   responsive to a detection of a failure of the replication workflow:
 determining whether the failure is in a Software Defined Network Attached Storage (SDNAS) or in an underlying storage array platform; 
 responsive to a determination that the failure is in the SDNAS:
 validating a health of the SDNAS for performing the replication workflow; and 
 responsive to validating the health of the SDNAS, performing a recovery of the failed replication workflow; and 
 
 responsive to a determination that the failure is in the underlying storage array platform:
 determining whether the underlying storage array platform can be rolled back to a previously known good state for performing the replication workflow; and 
 responsive to a determination that the underlying storage array platform can be rolled back to the previously known good state:
 performing a rollback of the underlying storage array platform to the previously known good state; and 
 performing the recovery of the failed replication workflow.

Join the waitlist — get patent alerts

Track US12481545B2 — get alerts on status changes and closely related new filings.

We store only your email — no account needed. See our privacy policy.