Character Recognition Device for Partially-Hidden Text Estimation
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional character recognition technologies fail to account for hidden characters in images, leading to incomplete text recognition and reduced translation accuracy and search recall rates due to obstruction, over-exposure, or defocusing issues.
Innovation Solution
A character recognition device that detects visible text areas, estimates partially-hidden text areas by integrating visible and hidden text, and uses linguistic evaluation to supplement missing characters, ensuring comprehensive character recognition and improved text retrieval.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If conventional character recognition technology is used to recognize only visible characters, then the recognition process is simple and fast, but the text completeness is poor and translation accuracy declines
Solution Approach 1:
The system performs preliminary detection of visible text areas and preliminary recognition of visible characters before estimating hidden text areas. This preliminary action allows the system to establish a baseline and then supplement hidden characters based on contextual information, improving text completeness without completely redesigning the recognition process
Solution Approach 2:
The system introduces an intermediary estimation process that uses linguistic information and contextual clues to infer hidden characters. This intermediary step bridges the gap between visible and hidden text, allowing the system to recover incomplete text while maintaining a manageable recognition process
2Measurement precision
If hidden characters are supplemented through estimation and integration, then translation accuracy and search recall rate improve, but the processing time and computational load increase
Solution Approach 1:
The system applies partial action by focusing estimation efforts on specific hidden text areas rather than processing the entire image uniformly. By identifying and estimating only the necessary hidden portions based on visible text context, the system improves recognition accuracy while minimizing additional processing time
Solution Approach 2:
The system applies different processing qualities to different regions of the text. Visible text areas undergo standard recognition processing, while hidden text areas receive enhanced estimation processing using linguistic models. This local differentiation improves overall accuracy without uniformly increasing processing time across the entire image
3Loss of information
If the system integrates visible and hidden text areas with linguistic evaluation, then the recall rate of search results improves, but the device complexity and computational resources required increase
Solution Approach 1:
The system merges visible text recognition results with hidden text estimation results into a unified text output. By combining these two sources of information through integration processes, the system recovers complete text information while managing system complexity through coordinated processing of both visible and hidden components
Data Source
AI summary
According to an embodiment, a device includes a detector, first and second recognizers, an estimator, a second recognizer, and an output unit. The detector is configured to detect a visible text area including a visible character from an image. The first recognizer is configured to perform character pattern recognition on the visible text area, and calculate a recognition cost according to a likelihood of a character pattern. The estimator is configured to estimate a partially-hidden text area into which a hidden text area estimated to have a hidden character and the visible text area are integrated. The second recognizer is configured to calculate an integrated cost into which the calculated cost and a linguistic cost corresponding to a linguistic likelihood of a text that fits in the entire partially-hidden text area are integrated. The output unit is configured to output a text selected or ranked based on the integrated cost.


