Electronic Document Indexing via Anchor and Sheet Name Detection
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current machine-readable text analysis and database indexing systems fail to effectively identify and index internal cross-references within electronic documents, leading to inefficient navigation and linkage between document pages.
Innovation Solution
A system and method that analyze machine-encoded text to identify and categorize sheet names, anchors, and anchor references, generating an index that maps associations between these notations, allowing for the creation of linked pages with hyperlinks and integrated reports, using optical character recognition (OCR) and regular expressions to improve accuracy and pattern recognition.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If traditional text analysis methods are used to index electronic documents, then the indexing process is simple, but internal cross-references and anchors cannot be effectively identified
Solution Approach 1:
The patent segments the text analysis process into multiple specialized modules: OCR processing module, notation identification module (for sheet names, anchors, anchor references), and indexing module. Each module handles a specific aspect of the complex task, improving identification accuracy while managing system complexity through functional decomposition
Solution Approach 2:
The patent introduces an intermediary indexing structure that stores identified notations and their relationships. This intermediary layer between text analysis and document navigation enables accurate cross-reference identification without requiring direct complex analysis between all document elements
2Productivity
If manual navigation methods are used in electronic documents, then the system complexity is low, but navigation efficiency and user experience deteriorate
Solution Approach 1:
The patent performs preliminary actions by automatically identifying and indexing all anchors, anchor references, and sheet names during document processing. This pre-established index enables efficient navigation without requiring complex real-time analysis during user interaction, improving productivity while managing complexity through advance preparation
Solution Approach 2:
The patent creates a copied structural representation of document relationships in the index, separate from the original document content. This copied structure allows efficient navigation operations to be performed on the index rather than analyzing the full document structure during navigation, improving speed while isolating complexity
3Adaptability or versatility
If comprehensive indexing of all notations is performed, then navigation capability is improved, but processing time and computational resources increase
Solution Approach 1:
The patent extracts only the essential notations (sheet names, anchors, anchor references) that are necessary for navigation, rather than indexing all text elements. This selective extraction provides adequate navigation capability while reducing processing time and computational resources by focusing on critical elements only
Data Source
AI summary
The present disclosure provides various systems and methods for indexing digital (electronic) documents. The systems and methods may utilize various software, hardware, and firmware modules to identify notations, such as sheet names, anchors, and anchor references on construction documents. The identified notations are indexed and used to create hyperlinked pages that are easily navigable. In some embodiments, the hyperlinked pages may include previous- and next-sheet hyperlinks that allow for direct navigation within a set of pages, according to an order provided in an index.


