Virtualized Storage De-duplication via Pointer Mapping
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
In virtualized storage environments, redundant data leads to inefficient use of storage capacity, as identical data across multiple host systems is stored separately, occupying unnecessary space and requiring solutions for de-duplication.
Innovation Solution
A data de-duplication application is implemented within the virtualized storage environment, which identifies redundant data and replaces it with pointers to a single instance, using a virtualization layer to manage storage capacity and map I/O requests, allowing for efficient reduction of redundant data across pooled storage.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Quantity of substance
If storage capacity is allocated to multiple host systems in a virtualized environment, then each host can access its required data, but redundant data occupies unnecessary storage space
Solution Approach 1:
The patent implements de-duplication by identifying identical data blocks across multiple host systems and replacing redundant copies with references or pointers to a single master copy. This eliminates redundant data storage while maintaining data accessibility for all hosts, directly resolving the contradiction between providing data access to multiple hosts and eliminating redundant storage.
2Ease of operation
If identical data is stored separately for each host system, then data access is simplified and fast, but storage efficiency deteriorates due to redundant copies
Solution Approach 1:
The system maintains data access efficiency by implementing a reference mechanism where multiple hosts can quickly access data through pointers or references to the master copy, eliminating the need to physically replicate identical data blocks across multiple storage locations while preserving fast access capabilities.
Solution Approach 2:
The patent merges identical data blocks from multiple host systems into a single master copy in the pooled storage capacity, consolidating redundant data while maintaining the ability for all hosts to access the consolidated data efficiently, thereby improving storage efficiency without significantly compromising data access ease.
3Adaptability or versatility
If pooled storage capacity is used to consolidate storage resources, then storage management is centralized and flexible, but redundant data accumulation increases
Solution Approach 1:
The de-duplication application operates within the pooled storage environment to identify and eliminate redundant data blocks across different host systems, replacing them with references to master copies. This maintains the centralized and flexible pooled storage management structure while actively preventing redundant data accumulation.
Solution Approach 2:
The system implements a feedback mechanism where the de-duplication application continuously monitors the pooled storage capacity for redundant data patterns, identifies duplicate blocks, and automatically replaces them with references to master copies, creating a self-regulating system that prevents redundant data accumulation while maintaining pooled storage flexibility.
Data Source
AI summary
In one example, a method for de-duplicating redundant data in a virtualized storage environment includes operating a data de-duplication application on a host system that is one of a plurality of host systems in a computer architecture, where the data de-duplication application is operable to globally de-duplicate data in a pooled storage capacity that comprises a virtualization of a plurality of storage devices.


