Service Processor Forensics via NC-SI During Storage Controller Failure

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

In a cluster environment, when a storage controller fails, administrators face challenges in obtaining forensics to diagnose the failure since communication with the service processor is routed through the storage controller, making remote access unavailable.

Innovation Solution

Configuring service processors to collect and expose forensics directly to a cluster health monitor using alternative communication paths, such as the Network Controller Sideband Interface (NC-SI), even during storage controller failures, ensuring continuous access for diagnostic purposes.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If communication with the service processor is routed through the storage controller, then the service processor can be managed and monitored, but remote access becomes unavailable when the storage controller fails

Engineering Contradiction:
Improveservice processor availabilityVSAvoidremote access capability
Core Design Contradiction:
ReliabilityVSEase of operation

Solution Approach 1:

The patent introduces a cluster health monitor as an intermediary component that can directly access the service processor through alternative communication paths (such as the NC-SI interface) when the storage controller fails. This mediator enables continuous monitoring and forensics collection without relying on the failed storage controller, resolving the contradiction between maintaining service processor availability and preserving remote access capability.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Stability of the object's composition

If the service processor operates independent of the storage controller, then it can continue functioning during storage controller failure, but communication paths become unavailable

Engineering Contradiction:
Improveservice processor operational independenceVSAvoidforensics accessibility
Core Design Contradiction:
Stability of the object's compositionVSLoss of information

Solution Approach 1:

The patent adds another communication dimension by implementing alternative communication paths (such as the Network Controller Sideband Interface or NC-SI) that are independent of the storage controller. This allows the service processor to maintain operational independence while preserving forensics accessibility through a different communication channel, effectively adding a dimensional alternative to the failed communication path.

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

3Difficulty of detecting and measuring

If administrators want to obtain forensics during storage controller failure, then diagnostic capability is needed, but communication through the failed controller is unavailable

Engineering Contradiction:
Improvefailure diagnosis capabilityVSAvoidcommunication availability
Core Design Contradiction:
Difficulty of detecting and measuringVSReliability

Solution Approach 1:

The cluster health monitor serves as a mediator that can directly communicate with the service processor through alternative paths to collect forensics during storage controller failures. This intermediary approach enables failure diagnosis capability while maintaining communication availability, as the health monitor can access the service processor without relying on the failed storage controller.

Inventive Principle:
Principle #24Intermediary (Mediator)

Data Source

PatentUS9836345B2Forensics collection for failed storage controllers
Publication Date: 2017.12.05 NETAPP INC
  • US9836345B2 patent drawing
  • US9836345B2 patent drawing
  • US9836345B2 patent drawing

AI summary

One or more techniques and/or systems are provided for collecting forensics associated with a failure of a storage controller. For example, a storage node, of a cluster environment, may comprise a service processor and a storage controller. The storage controller may manage a storage device accessible, through the storage controller, to one or more client devices. The service processor may manage the storage controller (e.g., collect operational statistics of the storage controller, perform software and/or firmware updates for the storage controller, etc.). The service processor may obtain forensics associated with a failure of the storage controller, and may provide the forensics to a cluster health monitor notwithstanding the storage controller being in an inoperable state (e.g., the service processor may send the forensics through a network interface controller of the storage node, over a non-client storage management network, to the cluster health monitor).