RAID Disk Failure Control via Protection Mode

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

As the number of disks in a RAID increases, the probability of simultaneous disk failures or disconnections rises, leading to a higher risk of user data loss due to the uncertainty of manual operations and the life cycle of disks.

Innovation Solution

A method is introduced to determine the number of failed disks in a RAID, compare it to a predetermined threshold, and set non-failing disks into a protection mode to prevent disconnection, using mechanical locking mechanisms or indication marks to ensure data security.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Quantity of substance

If the number of disks in RAID is increased to improve storage capacity, then storage capacity is improved, but the probability of simultaneous disk failures or disconnections increases leading to higher risk of data loss

Engineering Contradiction:
Improvestorage capacityVSAvoiddata security
Core Design Contradiction:
Quantity of substanceVSReliability

Solution Approach 1:

The system proactively identifies disks that are at risk of failure based on health indicators and preemptively locks them in protection mode before actual failure occurs. This preliminary action prevents the disk from being disconnected, thereby maintaining RAID data security even as the number of disks increases.

Inventive Principle:
Principle #10Preliminary action

2Ease of operation

If manual operations are allowed on disks to improve ease of operation, then ease of operation is improved, but the uncertainty of manual operations increases the risk of data loss

Engineering Contradiction:
Improvedisk accessibilityVSAvoiddata security
Core Design Contradiction:
Ease of operationVSReliability

Solution Approach 1:

The system continuously monitors disk health indicators and provides feedback by automatically locking disks in protection mode when risk thresholds are exceeded. This feedback mechanism restricts manual operations on at-risk disks, preventing human error from causing data loss while maintaining ease of operation for healthy disks.

Inventive Principle:
Principle #23Feedback

3Reliability

If protection mode is activated to prevent disk disconnection and improve data security, then data security is improved, but the number of restricted operations on disks increases

Engineering Contradiction:
Improvedata securityVSAvoiddisk operation flexibility
Core Design Contradiction:
ReliabilityVSAdaptability or versatility

Solution Approach 1:

The protection mode is applied locally and selectively only to specific disks that exhibit risk indicators, rather than restricting all disks in the RAID array. This localized approach maintains data security for at-risk disks while preserving operational flexibility for healthy disks that do not require protection.

Inventive Principle:
Principle #3Local quality

Data Source

PatentUS11314581B2Control method of disk failure, electronic device and computer readable storage medium
Publication Date: 2022.04.26 EMC IP HLDG CO LLC
  • US11314581B2 patent drawing
  • US11314581B2 patent drawing
  • US11314581B2 patent drawing

AI summary

Techniques for disk failure control involve determining the number of failed disks in a Redundant Array of Independent Disks (RAID). The techniques further involve comparing the number of failed disks with a predetermined threshold; and in accordance with a determination that the number of failed disks exceeds the predetermined threshold, setting at least one non-failing disk in the RAID into a protection mode to prevent the at least one non-failing disk from being disconnected. Such techniques facilitate prevention of the user data loss in the RAID.