Image Processing Device Synonym Embedding for Search Reliability
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional image processing systems face difficulties in searching for relevant items within image data due to inconsistent spellings and require manual selection of similar words, leading to poor operability and potential failure in finding all relevant items, especially when the search environment changes.
Innovation Solution
An image processing device and method that extracts words from image data, obtains synonyms, and embeds them in an accompanying layer at the same display position as the original word, allowing for easy text search without environmental influence, using a thesaurus and customized dictionaries to provide relevant synonyms.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If a user performs an OR search with multiple keywords to find inconsistent spelled words, then the search coverage is improved, but the operability deteriorates due to the difficulty of sorting out all inconsistent spelled words
Solution Approach 1:
The system performs preliminary extraction of inconsistent spelled words and generates a candidate list before the user executes the search. This pre-processing step prepares the search candidates in advance, so when the user performs the search, the results are already organized and ready for selection, eliminating the need for manual sorting during the search process.
Solution Approach 2:
The system introduces an intermediary component that automatically extracts and presents candidate inconsistent spelled words to the user. This intermediary layer handles the complex task of identifying and organizing potential search terms, serving as a bridge between the raw image data and the user's search query, thereby simplifying the user's interaction.
2Measurement precision
If a user manually selects a string for searching among extracted similar strings, then the search precision is improved, but the time consumption increases due to the need for repeated selection
Solution Approach 1:
The system performs automatic extraction and presentation of candidate inconsistent spelled words without requiring manual intervention for each selection. The system serves itself by automatically preparing the search candidates and presenting them in an organized manner, reducing the need for repeated manual selections while maintaining search precision.
3Reliability
If data for search is registered in advance with standard descriptions, then the search reliability is improved, but the adaptability deteriorates when searching in different environments or with different users
Solution Approach 1:
The system extracts inconsistent spelled words directly from the image data itself, making the extraction process independent of any pre-registered search data. This universal approach allows the same extraction mechanism to work across different environments, users, and document types without requiring environment-specific pre-registration, thereby achieving both reliability and adaptability.
4Reliability
If fuzzy search is performed to extract similar strings, then the search coverage is improved, but the difficulty of detecting and measuring increases due to the need to select from multiple candidate strings
Solution Approach 1:
The system extracts and isolates candidate inconsistent spelled words from the image data, presenting them as a separate, organized list of candidates. This extraction process separates the candidate identification from the selection process, making it easier for the user to review and select from a manageable set of pre-processed candidates rather than searching through all possible variations.
Data Source
AI summary
An image processing device, comprises: an input part for inputting image data; a word extracting part for extracting a word from texts contained in the image data; a synonym obtaining part for obtaining a synonym corresponds to the word, and for associating the obtained synonym with the word; a position identifying part for identifying a display position on the image data of the word with which the synonym is associated; a layer creating part for creating an accompanying layer to add to an original layer, which is the image data containing the word, and for embedding the synonym associated with the word within a position on the accompanying layer the same as the display position identified by the position identifying part; and an output image generating part for generating output image data including the original layer containing the word and the accompanying layer within which the synonym is embedded.


