Cloud Operation Failure Prediction Using Resource Utilization Analysis

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Predicting and preventing operation failures in cloud systems due to insufficient resources is challenging, as current technologies struggle to anticipate the impact of resource utilization on cloud operations, leading to potential failures without timely alerts.

Innovation Solution

A system that determines historical and current resource utilization in cloud systems, using a determination engine to assess the likelihood of future operation failures and generates alerts when thresholds are exceeded, allowing for proactive management.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If cloud systems operate without predictive monitoring, then system complexity and resource allocation flexibility are maintained, but operation failures occur due to insufficient resource anticipation

Engineering Contradiction:
Improveoperation reliabilityVSAvoidmonitoring system complexity
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

The system performs preliminary actions by determining historical utilization of resources for performing operations and current utilization of resources before operations execute. This advance assessment allows the system to predict potential failures before they occur, enabling proactive alert generation rather than reactive response.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system implements feedback mechanisms by continuously monitoring resource utilization and comparing it against historical data and thresholds. The determination engine uses this feedback loop to assess likelihood of future operation failures and generate alerts when predictive thresholds are exceeded, creating a closed-loop monitoring system.

Inventive Principle:
Principle #23Feedback

2Reliability

If historical and current resource utilization analysis is implemented, then future operation failures can be predicted, but computational overhead and processing time increase

Engineering Contradiction:
Improvefailure prediction accuracyVSAvoidprocessing time
Core Design Contradiction:
ReliabilityVSLoss of time

Solution Approach 1:

The system extracts only the essential data elements needed for failure prediction by determining specific historical utilization metrics and current utilization states. Rather than analyzing all possible system parameters, the system focuses on extracting and analyzing only the resource utilization data that directly correlates with operation failure likelihood.

Inventive Principle:
Principle #2Taking out (Extraction)

3Reliability

If proactive alert generation is implemented, then operation failures can be prevented, but false alerts may increase system noise and reduce operational efficiency

Engineering Contradiction:
Improveoperation reliabilityVSAvoidoperational efficiency
Core Design Contradiction:
ReliabilityVSProductivity

Solution Approach 1:

The system uses parameter changes by establishing predictive thresholds for resource utilization that trigger alerts. The determination engine dynamically assesses whether current and historical utilization patterns indicate a high likelihood of failure based on these parameters, generating alerts only when predictive criteria are met rather than on fixed schedules.

Inventive Principle:
Principle #35Parameter changes

Data Source

PatentUS11561878B2Determining a future operation failure in a cloud system
Publication Date: 2023.01.24 HEWLETT PACKARD ENTERPRISE DEV LP
  • US11561878B2 patent drawing
  • US11561878B2 patent drawing
  • US11561878B2 patent drawing

AI summary

Examples described relate to determining a future operation failure in a cloud system. In an example, a historical utilization of resources for performing an operation in a cloud system may be determined. A current utilization of resources in the cloud system may be determined. Based on the historical utilization of resources for performing the operation in the cloud system and the current utilization of resources in the cloud system, a determination may be made whether a future performance of the operation in the cloud system is likely to be a failure. In response to a determination that the future performance of the operation in the cloud system is likely to be a failure, an alert may be generated.