Hierarchical Data Storage System Content Indexing
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current data protection systems face challenges in efficiently searching and retrieving data across multiple copies stored on various devices, as they are time-consuming, put a significant load on systems, and lack integrated data rights security control and the ability to apply search criteria throughout the data management system.
Innovation Solution
A hierarchical data storage system with a global cell and storage operations cells that perform data storage operations, including content indexing and searching, allowing for the creation of indices across all data types and platforms, enabling users to search and retrieve data based on integrated data security policies and apply search criteria through a unified interface.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If typical search systems search over multiple copies on multiple devices, then data retrieval capability is improved, but system load increases significantly and search time becomes excessive
Solution Approach 1:
The system creates and maintains content indices in advance for all secondary copies across multiple devices and formats. This preliminary indexing allows search queries to be executed quickly by simply searching the pre-built indices rather than scanning actual data content, thereby improving search efficiency while maintaining comprehensive data retrieval capability
Solution Approach 2:
The patent introduces content indices as an intermediary layer between the search function and the actual data copies. The indices serve as mediators that contain metadata and keywords extracted from all secondary copies, allowing the search system to query this intermediate structure instead of directly searching through multiple data copies on multiple devices, thus reducing system load and search time
2Ease of operation
If typical systems perform restore operations to retrieve data, then data accessibility is improved, but production data is overwritten and security control is lost
Solution Approach 1:
The system extracts and separates the data retrieval function from the restore operation. Instead of requiring a restore operation that copies data back to production servers, the system allows users to search and access data directly from secondary copies through the content indices, extracting only the necessary retrieval capability while leaving production data intact and secure
Solution Approach 2:
The patent uses the existing secondary copies as the basis for direct access without creating additional restoration copies. The content indices reference these secondary copies, allowing users to access data from the secondary copies themselves rather than restoring them to production environments, thus maintaining production data integrity while providing data accessibility
Data Source
AI summary
A complete document management system is disclosed. Accordingly, systems and methods for managing data associated with a data storage component coupled to multiple computers over a network are disclosed. Systems and methods for managing data associated with a data storage component coupled to multiple computers over a network are further disclosed. Additionally, systems and methods for accessing documents available through a network, wherein the documents are stored on one or more data storage devices coupled to the network, are disclosed.


