Software Agent for Suspected Storage Drive Failure Identification

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional storage management approaches fail to detect suspected drive failures in software-defined storage systems, leading to data unavailability and loss, as they cannot identify drives that are not currently failing but are likely to do so in the future.

Innovation Solution

Implementing a software agent in the operating system to monitor and process predefined storage drive attributes, using algorithmic logic to identify suspected failures and perform automated actions, such as notifying the system to create a redundant copy of data and initiate a replacement workflow.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If conventional storage management approaches are used, then the system operates with simple monitoring, but suspected drive failures cannot be detected leading to data loss

Engineering Contradiction:
Improvedrive failure detection capabilityVSAvoidmonitoring system complexity
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

A software agent is introduced as an intermediary component that runs within the operating system to monitor storage drive attributes. This agent collects attribute data from drives, processes it through algorithmic logic, and communicates failure suspicions to the storage management system, thereby enabling sophisticated detection without requiring complex changes to the core storage management architecture

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The patent replaces traditional hardware-based or firmware-based monitoring mechanisms with a software-based monitoring system. The software agent uses algorithmic logic to analyze drive attributes and detect suspected failures, substituting mechanical or firmware monitoring approaches with flexible software-based detection that can adapt to different drive types and failure modes

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Reliability

If proactive suspected failure detection is implemented, then data loss is prevented, but additional monitoring and processing resources are consumed

Engineering Contradiction:
Improvedata availabilityVSAvoidmonitoring resource consumption
Core Design Contradiction:
ReliabilityVSUse of energy by moving object

Solution Approach 1:

The software agent monitors only predefined storage drive attributes that are most indicative of potential failures, rather than continuously analyzing all possible drive parameters. The algorithmic logic processes attribute values selectively to identify suspected failures, performing partial monitoring actions that are sufficient for detection without consuming excessive computational resources

Inventive Principle:
Principle #16Partial or excessive action

Solution Approach 2:

The monitoring system leverages existing drive self-diagnostic capabilities and attribute reporting mechanisms. The software agent utilizes attributes that drives already report through existing interfaces (such as SMART attributes), allowing the system to perform proactive detection without adding significant overhead for data collection and processing

Inventive Principle:
Principle #25Self-service

Data Source

PatentUS11474904B2Software-defined suspected storage drive failure identification
Publication Date: 2022.10.18 EMC IP HLDG CO LLC
  • US11474904B2 patent drawing
  • US11474904B2 patent drawing
  • US11474904B2 patent drawing

AI summary

Methods, apparatus, and processor-readable storage media for software-defined suspected storage drive failure identification are provided herein. An example computer-implemented method includes implementing at least one software agent in an operating system associated with at least one storage system, wherein the at least one software agent is configured to monitor and process one or more predefined storage drive attributes; obtaining, using the at least one software agent, attribute values for the one or more predefined storage drive attributes from one or more storage drives within the at least one storage system; identifying, using the at least one software agent, at least one suspected failure among the one or more storage drives by processing the obtained attribute values using algorithmic logic; and performing at least one automated action based on the at least one identified suspected failure among the one or more storage drives.