Container File Archive Deduplication Legal Validity
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional electronic archives face challenges with large storage space requirements and slow data retrieval times, especially when dealing with numerous small files, and lack mechanisms to ensure data immutability and bit-precision restoration while meeting legal standards for long-term archiving.
Innovation Solution
A method involving the creation of a container file that combines archive data and administration indices, utilizing deduplication to link duplicate files and ensure immutability, along with an efficient indexing system for quick search functionality, while maintaining qualified electronic signatures for legal validity.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If archive data is stored in separate files with individual qualified electronic signatures, then legal validity is maintained, but storage space requirements increase significantly
Solution Approach 1:
The patent combines multiple archived files with their individual qualified electronic signatures into a single container file. This merging approach maintains the legal validity of each signed file while significantly reducing the total storage space required, as the signature verification infrastructure can efficiently handle multiple files within one container rather than requiring separate storage for each signed file.
Solution Approach 2:
The container file serves multiple functions: it stores the archived data, contains all necessary qualified electronic signatures for legal validity, and provides a unified structure for storage and retrieval. This multi-functionality eliminates the need for separate storage mechanisms for each signed file, thereby reducing overall storage requirements while maintaining legal compliance.
2Quantity of substance
If conventional archiving systems store large numbers of small archived data records, then comprehensive archiving is achieved, but searching and filtering becomes very time-consuming
Solution Approach 1:
The patent implements pre-computed indexes within the container file that organize and categorize the archived data before any search operation occurs. These indexes are created during the archiving process, allowing for extremely fast retrieval and filtering operations later, thus eliminating the time-consuming search problem associated with large numbers of small records.
3Adaptability or versatility
If archived data is stored in external repositories with reference databases, then storage scalability is improved, but data retrieval and verification complexity increases
Solution Approach 1:
The patent merges the archived data and the verification infrastructure (including signature verification capabilities) into a single container file. This integration maintains storage scalability while reducing system complexity, as the container file is a self-contained unit that can be stored, transferred, and verified without requiring complex external reference database systems.
Data Source
AI summary
The present invention relates to a method for producing and managing a large-volume long-term archive, which comprises an archive data memory and a management file, and to an appropriate long-term archive. The method according to the invention involves archive data being relocated in a container file, so that the legal validity of the data is maintained by virtue of qualified signing. The method for producing and managing a large-volume long-term archive, which archive comprises an archive data memory and a management database, involves the steps of - selecting archive data using prescribed features; - relocating the selected archive data from the archive data memory to an independent archive data file; - erasing the relocated archive data from the archive data memory; - relocating the indices for the selected archive data from the management database to an independent database file; - erasing the relocated indices from the management database; - combining the independent archive data file with the independent database file in a container file; and - deduplicating entries in the container file by subjecting the entries to bit pattern comparison and replacing identical patterns with references to the relevant entries in the database.