Image Text Tag Synchronization via Alignment Data
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing data management systems face challenges in efficiently identifying and aligning matching portions between image-based and text documents, due to recognition errors and formatting issues, which hinder effective tagging and data review.
Innovation Solution
A method and system that access both image-based and text documents, allowing users to select and tag content in the image-based document and automatically align these tags with corresponding text documents, using OCR and alignment data to synchronize tags across formats, and normalize text data for better matching.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Loss of information
If text is extracted from image-based documents via OCR, then text data becomes available for searching and tagging, but recognition errors and formatting artifacts are introduced
Solution Approach 1:
The patent introduces alignment data as an intermediary element that connects text portions in the text document to their corresponding portions in the image-based document. This alignment data acts as a mediator that allows tags applied to text to be automatically propagated to the corresponding image-based document portions, resolving the issue of text accuracy while maintaining text availability for searching and tagging.
2Productivity
If tags are applied to text portions in text documents, then tagging efficiency improves, but tags do not automatically appear in image-based documents
Solution Approach 1:
The system implements a feedback mechanism where alignment data is used to automatically propagate tags from text documents to image-based documents. When a user applies a tag to text, the system uses the alignment data to identify the corresponding portion in the image-based document and automatically applies the same tag, ensuring tag synchronization across both document formats without manual intervention.
3Adaptability or versatility
If both image-based and text versions of documents are maintained, then user flexibility increases, but determining alignment between corresponding portions becomes difficult
Solution Approach 1:
The patent applies preliminary action by pre-establishing alignment data that maps text portions to their corresponding portions in image-based documents before the tagging process begins. This alignment data is created in advance and stored in the system, allowing the tagging system to quickly and automatically determine corresponding portions without complex real-time analysis, thus reducing the complexity of alignment determination while maintaining user flexibility.
Data Source
AI summary
A computing system accesses an image-based document and a text document having text extracted from the image-based document and provides a user interface displaying at least a portion of the image-based document. In response to selection of a text portion of the image-based document, the system determines an occurrence of the text portion within at least a portion of the image-based document and then applies a search model on the text document to identify the same occurrence of the text portion. Once matched, alignment data indicating a relationship between a selected tag and both the text portion of the image-based document and the text portion of the text document is stored.


