Workload Distribution Controller for Compute Nodes

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Modern computing systems face challenges in efficiently distributing workload assignments among compute nodes due to errors that can lead to the 'Storm Drain Problem, where an error-generating node rapidly consumes tasks without explicit feedback, causing inefficiency and performance issues.

Innovation Solution

Implementing a distribution controller that monitors consumption rates of workload assignments across compute nodes and adjusts distribution based on consumption patterns to identify and mitigate errors, thereby preventing the 'Storm Drain Problem by redistributing tasks away from error-generating nodes.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Productivity

If the distribution controller distributes workload assignments to compute nodes based on simple availability, then the distribution process is fast and simple, but error-generating nodes rapidly consume tasks without explicit feedback causing the Storm Drain Problem and system inefficiency

Engineering Contradiction:
Improveworkload distribution speedVSAvoidworkload distribution accuracy
Core Design Contradiction:
ProductivityVSReliability

Solution Approach 1:

The patent implements a feedback mechanism where the distribution controller monitors consumption of workload assignments by each compute node and uses this feedback to dynamically adjust future distribution decisions. This allows the system to identify error-generating nodes through their consumption patterns and redistribute tasks away from them, resolving the Storm Drain Problem while maintaining efficient workload distribution

Inventive Principle:
Principle #23Feedback

Solution Approach 2:

The patent applies preliminary action by having the distribution controller proactively monitor consumption patterns and identify potential error-generating nodes before they cause significant system degradation. By detecting abnormal consumption behaviors early and preemptively redistributing workload, the system prevents the Storm Drain Problem from fully developing, maintaining both productivity and reliability

Inventive Principle:
Principle #10Preliminary action

2Reliability

If the distribution controller monitors consumption of workload assignments by each compute node, then workload distribution accuracy improves, but the complexity of the distribution controller increases

Engineering Contradiction:
Improveworkload distribution accuracyVSAvoiddistribution controller complexity
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

The patent implements self-service by having compute nodes report their own workload assignment consumption status to the distribution controller. This self-reporting mechanism eliminates the need for complex centralized monitoring infrastructure, as each node autonomously provides the necessary consumption data, thereby improving workload distribution accuracy without proportionally increasing controller complexity

Inventive Principle:
Principle #25Self-service

Data Source

PatentUS8688831B2Managing workload distribution among a plurality of compute nodes
Publication Date: 2014.04.01 INTERNATIONAL BUSINESS MACHINE CORPORATION
  • US8688831B2 patent drawing
  • US8688831B2 patent drawing
  • US8688831B2 patent drawing

AI summary

Methods, apparatuses, and computer program products for managing workload distribution among a plurality of compute nodes are provided. Embodiments include monitoring, by the distribution controller, consumption of workload assignments by each compute node of the plurality of compute nodes; and distributing, by the distribution controller, unconsumed workload assignments to one or more compute nodes of the plurality of compute nodes based on the consumption of the workload assignments of each compute node of the plurality of compute nodes.