Electronic Document Content Location via Stream Descriptors
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing electronic document processing techniques rely on actual character counts for identifying characters, which leads to inconsistencies and failures when different processes or versions use different filtering mechanisms, causing issues in accurately locating and redacting content across multiple processes or routines.
Innovation Solution
The method involves generating a location identification rule specifying parameters for identifying a location identifier (LID) based on stream descriptors in the file format of an electronic document, allowing for the assignment of LIDs to content items and enabling accurate location of content across different processes or routines without relying on actual character counts.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If actual character counts are used for identifying characters, then the filtering process can identify unique characters based on their position, but the redaction process may fail to locate the same characters if it uses a different filtering mechanism or version
Solution Approach 1:
The patent introduces an intermediary mapping mechanism that decouples the filtering process from the redaction process. Instead of directly using actual character counts, the system creates an intermediate layer of location identification that translates between different filtering mechanisms. This intermediary mapping ensures that the filtering process can identify characters using one method while the redaction process can locate them using a different method, resolving the inconsistency issue between processes.
Solution Approach 2:
The patent segments the character identification process into distinct components: the filtering process that identifies characters to be redacted, and the redaction process that actually performs redaction. By separating these functions and introducing independent location tracking for each, the system allows each process to use its own filtering mechanism without interfering with the other, thus maintaining both measurement precision and process reliability.
2Adaptability or versatility
If different versions of the filtering mechanism are used, then each version can incorporate upgraded features or defect fixes, but actual character counts may differ causing redaction errors
Solution Approach 1:
The patent applies preliminary action by pre-establishing a mapping between different versions of the filtering mechanism and their corresponding actual character counts. Before redaction occurs, the system creates a translation layer that accounts for differences between filtering mechanism versions. This preliminary mapping ensures that even when different versions are used, the character identification remains accurate because the system has already compensated for version-specific variations in advance.
3Ease of operation
If the redaction process starts counting characters from a different position (e.g., from body instead of header), then it may redact incorrect characters, but requiring the same exact filtering mechanism reduces flexibility
Solution Approach 1:
The patent introduces an intermediary location tracking system that acts as a mediator between the filtering process and redaction process. This intermediary layer records the actual positions of characters independently of how each process counts them. The filtering process can start counting from the header while the redaction process can start from the body, and the intermediary mapping ensures they still refer to the same actual characters, thus maintaining both process independence and measurement precision.
Data Source
AI summary
Techniques are described relating to the identification of location of content within an electronic document. Techniques may include generating a location identification rule specifying one or more parameters for identifying a location identifier (LID) for each of the one or more streams associated with the content of the electronic document. Further, the LID may be generated in accordance with the location identification rule. The LID may be assigned to at least a portion of the content, such that the portion of the content within the electronic document may be located in accordance with the LID.


