Metadata Candidate Display for Scanned Documents
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing technologies fail to appropriately display candidates of selection-type values in metadata associated with scanned image data, leading to incomplete or incorrect metadata settings.
Innovation Solution
An information processing apparatus that extracts character strings from scanned image data and assigns metadata in a key-value format, featuring a display control unit that prioritizes and displays candidate values matching the extracted character strings for selection-type keys, ensuring accurate and efficient metadata input.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If candidates of selection-type values are obtained only from past registration data, then the system can maintain data consistency, but candidates cannot be displayed when no past data exists
Solution Approach 1:
The system performs preliminary OCR extraction on the scanned document to obtain character strings before the user needs to input metadata. This extracted information is then used to pre-populate or suggest values in the metadata input form, so that when the user needs to input data, relevant candidates are already prepared and displayed based on the document content itself, not just past registration data
Solution Approach 2:
The patent introduces OCR-extracted character strings as an intermediary source between the scanned document and the metadata input form. This intermediary provides additional candidate values that complement the database-derived candidates, allowing the system to display relevant options even when no matching past data exists, while maintaining data consistency through validation against the predefined value lists
2Adaptability or versatility
If all candidates of selection-type values are displayed in the metadata input form, then users have complete options, but it becomes difficult for users to identify appropriate candidates
Solution Approach 1:
Instead of uniformly displaying all candidates or using a single display method, the patent applies different display strategies to different candidates based on their relevance. Candidates extracted from the document content are given prominence (e.g., displayed at the top or highlighted), while other candidates from the database are displayed in the standard list. This local differentiation helps users quickly identify the most relevant options without hiding less relevant ones
Solution Approach 2:
The system pre-processes and prioritizes candidates before display by extracting character strings from the document and matching them against the value lists. This preliminary sorting ensures that the most relevant candidates appear first in the display, reducing the cognitive load on users and making it easier to select appropriate values without having to scan through the entire list
Data Source
AI summary
According to the technology of the present disclosure, candidates of a selection-type value included in metadata in a key-value format, which is registered in association with scanned image data, can be appropriately displayed. An information processing apparatus, which assigns metadata in a key-value format to scanned image data obtained by scanning a document, includes: a first obtaining unit configured to obtain a character string extracted from the scanned image data; a second obtaining unit configured to obtain a template of the metadata; a display control unit configured to display a screen for inputting the metadata; and a setting unit configured to set a value in association with the scanned image data, based on operation by a user via the screen, the value corresponding to a key included in the template of the metadata obtained by the second obtaining unit.


