OCR Proofreading System with Line Movement Detection
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current proofreading software for digitized documents generated by OCR processing lacks efficiency and accuracy, particularly in maintaining the original text position and order, leading to poor operational efficiency and low correction accuracy due to the lack of functions like spell checking and grammar checking, and the inability to accurately reflect user corrections in the OCR output.
Innovation Solution
An information processor and method that reads OCR-processed text, generates a document file based on the original text and position information, detects line movements during user proofreading, and merges the corrected text back into the output information, using techniques like Levenshtein distance calculation to determine the degree of matching between line operations and reflect the changes accurately.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If OCR processing is performed on source manuscript images, then text extraction is achieved, but text position accuracy and reading order maintenance deteriorate
Solution Approach 1:
The patent segments the OCR output into individual character elements with associated positional information, then reconstructs them into lines and paragraphs while maintaining the original reading order. This segmentation allows precise control over text assembly while preserving spatial relationships from the original document.
Solution Approach 2:
The patent performs preliminary processing to establish the reading order and positional relationships before the actual proofreading operation. By pre-organizing the OCR output according to the original document structure, the system ensures that corrections maintain both accuracy and reading order without requiring complex real-time adjustments.
2Manufacturing precision
If proofreading operations are performed on OCR output, then correction accuracy is improved, but operational efficiency deteriorates
Solution Approach 1:
The patent implements a feedback mechanism where the proofreading operations are automatically reflected back into the OCR output information. This allows the system to learn from user corrections and maintain consistent formatting, reducing the need for manual intervention and improving both accuracy and efficiency.
Solution Approach 2:
The system performs self-service by automatically maintaining the document structure and reading order during proofreading operations. The merge unit automatically integrates corrections while preserving the original layout, eliminating the need for users to manually reformat or reorganize text after corrections.
3Manufacturing precision
If line movement detection is implemented, then proofreading accuracy is improved, but device complexity increases
Solution Approach 1:
The patent introduces an intermediary layer (the merge unit) that mediates between the proofreading operations and the OCR output. This intermediary component handles the complexity of line movement detection and integration automatically, shielding the user from system complexity while maintaining high proofreading accuracy.
4Measurement precision
If OCR processing is performed on source manuscript images, then text extraction is achieved, but spell checking and grammar checking functions are lost
Solution Approach 1:
The patent creates a universal proofreading system that handles multiple functions including spell checking, grammar checking, and structural verification within a single integrated platform. The merge unit serves multiple purposes: integrating corrections, maintaining reading order, preserving layout, and enabling various types of proofreading operations.
Data Source
AI summary
A method, apparatus and program for proofreading a document. The information processor includes a first storage unit for storing output information which includes information text and positional information obtained by performing Optical Character Recognition (OCR) on a source manuscript image. A second storage unit for storing a document file that is proofread by a user, wherein the document file is generated by reading the OCR-processed text according to the order of reading the output information A line movement detection unit for detected movement of a line which includes text in the document file based on the proofreading performed by the user on the document file. A merge unit for reflecting result of the proofreading of the document file in the output information.


