Mapped Cluster Stretching for Data Storage Workload Scaling

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional data storage systems face inefficiencies in handling increased workloads, leading to either prolonged processing times or the risk of destabilizing the storage system when trying to accelerate data analytics, especially with large amounts of archived data.

Innovation Solution

The technology involves 'cluster stretching' by increasing the number of mapped nodes in a mapped Redundant Array of Independent Nodes (RAIN) system while maintaining the same number of storage devices, allowing for improved computing resources and flexible redistribution of storage devices among nodes, which can be further stretched or un-stretched as needed.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Speed

If computing resources are heavily loaded to process large amounts of archived data faster, then processing speed is improved, but the storage system becomes unstable

Engineering Contradiction:
Improvedata processing speedVSAvoidstorage system stability
Core Design Contradiction:
SpeedVSReliability

Solution Approach 1:

The patent segments the storage system into multiple mapped clusters, each handling a portion of the data workload. By dividing the large-scale data processing task across multiple independent mapped clusters rather than overloading a single cluster, the system achieves faster aggregate processing speed while maintaining stability through distributed load management.

Inventive Principle:
Principle #1Segmentation

2Productivity

If the number of mapped nodes is increased to handle increased workload, then computing power is improved, but system complexity increases

Engineering Contradiction:
Improveworkload handling capacityVSAvoidmapped cluster configuration complexity
Core Design Contradiction:
ProductivityVSDevice complexity

Solution Approach 1:

The patent creates universal mapped cluster templates that can be replicated and scaled. Instead of manually configuring each mapped cluster individually, the system uses a standardized template approach where a single mapped cluster configuration can serve multiple purposes and be copied to create additional clusters, thereby increasing productivity while minimizing the complexity increase through reuse of proven configurations.

Inventive Principle:
Principle #6Universality (Multi-functionality)

3Reliability

If data is processed at moderate speed to maintain system stability, then system reliability is preserved, but processing time becomes unacceptably long

Engineering Contradiction:
Improvestorage system stabilityVSAvoiddata processing time
Core Design Contradiction:
ReliabilityVSLoss of time

Solution Approach 1:

The patent transitions from a single-dimension approach (one mapped cluster processing data sequentially) to a multi-dimensional approach by creating multiple mapped clusters that process different portions of data simultaneously in parallel. This dimensional expansion from sequential to parallel processing dramatically reduces total processing time while each individual cluster operates at stable, moderate speeds, thus preserving reliability.

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

Data Source

PatentUS11209996B2Mapped cluster stretching for increasing workload in a data storage system
Publication Date: 2021.12.28 EMC IP HLDG CO LLC
  • US11209996B2 patent drawing
  • US11209996B2 patent drawing
  • US11209996B2 patent drawing

AI summary

The described technology is generally directed towards stretching a mapped storage clusters by adding nodes to a mapped cluster of mapped nodes and storage devices mapped to a real cluster of nodes and storage devices. Stretching the mapped cluster can provide additional computing resources to a set of storage devices. In one implementation, one or more newly mapped nodes are added to increase the node count of an existing mapped cluster to form a stretched cluster, with the storage devices distributed among the increased number of nodes; a mapping table is updated to relate the stretched cluster nodes and storage devices to the real cluster nodes and storage devices. Also described is un-stretching a stretched cluster, or further stretching a stretched cluster.