Electronic Bookmarking via Heading-Image Association
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing methods for converting documents with tables of contents into electronic format are time-consuming and require users to repeatedly navigate between table of contents and main body pages, lacking efficient bookmarking functionality.
Innovation Solution
An information processing apparatus comprising a reading unit, recognition unit, table-of-contents analysis unit, main-body analysis unit, and creation unit that reads and analyzes images of table of contents and main body pages, performs character recognition, and creates electronic bookmarked information associating heading items with corresponding main body page images.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Loss of time
If users manually navigate through table of contents and main body pages to find desired information, then complete document coverage is achieved, but navigation time increases significantly
Solution Approach 1:
The system performs preliminary analysis during the document conversion process to identify and extract heading items from the table of contents, associating them with corresponding main body page images. This preliminary action creates electronic bookmarked information that stores the relationship between heading items and page images, eliminating the need for users to manually navigate through the table of contents during actual document usage.
Solution Approach 2:
The system creates a simplified copy of the navigation structure by extracting only the essential heading items and their associated page images, storing them as electronic bookmarked information. This copy provides direct access points to main body pages without requiring users to interact with the complete table of contents, thereby reducing navigation time while maintaining access to all necessary information.
2Productivity
If electronic documents are created from scanned images without automated bookmarking, then document conversion is simple, but user efficiency decreases due to lack of direct navigation
Solution Approach 1:
The system performs self-service by automatically analyzing the document structure during conversion, extracting heading items, and generating electronic bookmarked information without requiring manual intervention. The analysis units automatically identify table of contents pages versus main body pages, extract heading information through character recognition, and create the bookmark structure, enabling efficient document access while keeping the conversion process automated and user-friendly.
Solution Approach 2:
The system replaces manual bookmarking operations with automated mechanical processes. Instead of requiring users to manually create bookmarks or navigate through table of contents, the system uses character recognition technology and automated analysis to extract heading items and generate electronic bookmarked information, substituting manual mechanical operations with automated digital processing.
3Ease of operation
If complete table of contents navigation is maintained in electronic format, then all content is accessible, but user frustration increases due to repeated returns to table of contents
Solution Approach 1:
The system extracts the essential navigation elements (heading items and their associated page images) from the complete table of contents structure, storing them as separate electronic bookmarked information. This extraction allows users to access main body pages directly through bookmarks without needing to return to the table of contents, reducing repeated navigation time while maintaining access to all content through the bookmark system.
Data Source
AI summary
An information processing apparatus includes a reading unit, a recognition unit, a table-of-contents analysis unit, a main-body analysis unit, and a creation unit. The reading unit reads a table of contents page and a main body page as images. The recognition unit performs character recognition on the images of the table of contents and main body pages. The table-of-contents analysis unit analyzes the image of the table of contents page, and acquires at least a heading item in accordance with a result of character recognition. The main-body analysis unit analyzes the image of the main body page, and associates an image including the heading item with the heading item in accordance with a result of character recognition. The creation unit creates electronic bookmarked information in which bookmark information for associating the heading item with the image of the main body page is added to electronic information of the read images.


