Image Reading Apparatus Document Division via Page Number Layout Title Recognition
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing image reading apparatuses struggle to efficiently divide a document bundle into individual documents without relying on workflow execution history, leading to manual effort in document classification and reading.
Innovation Solution
An image reading apparatus equipped with a control device that includes a processor functioning as a page number recognizer, layout recognizer, title recognizer, controller, and divider. This apparatus acquires document images and uses these recognizers to determine the first page of a document based on page numbers, layout, and titles, then divides the images into documents accordingly.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If multiple recognition methods (page number, layout, title) are used to determine the first page, then document division accuracy is improved, but device complexity increases
Solution Approach 1:
The document analysis function is segmented into three independent recognizers: page number recognizer, layout recognizer, and title recognizer. Each recognizer independently analyzes different aspects of the document and outputs determination results. This segmentation allows the system to achieve high accuracy through multiple recognition methods while managing complexity by modularizing each recognizer as a separate functional unit.
Solution Approach 2:
The control device integrates multiple recognition functions (page number recognition, layout recognition, title recognition) into a single unified system that determines the first page of documents. This multi-functional approach allows one device to perform diverse document analysis tasks, improving accuracy without requiring separate dedicated devices for each recognition type.
2Measurement precision
If manual document classification is performed, then document division accuracy can be ensured, but productivity decreases
Solution Approach 1:
The system enables automatic document division by having the control device autonomously determine the first page of each document through integrated recognition methods. The apparatus processes document images automatically without requiring manual classification, thereby maintaining high accuracy while significantly improving productivity by eliminating manual intervention in the document division process.
3Measurement precision
If workflow execution history is required for document division, then document accuracy can be maintained, but adaptability decreases
Solution Approach 1:
The system performs preliminary analysis of document characteristics (page numbers, layout patterns, titles) directly from the document images themselves to determine the first page. This preliminary action eliminates the need for external workflow execution history, allowing the system to adapt to any document bundle without requiring pre-existing workflow information, thereby improving adaptability while maintaining accuracy through direct document analysis.
Data Source
AI summary
In an image reading apparatus, the image reading apparatus reads a document bundle to acquire a document image, and determines a first page of the document image according to the selected division method by a user. A page number recognizer extracts a page number from the document image by executing page number recognition processing, and determines the document image indicating a first page to be a first page of the document. A layout recognizer detects a marginal area or background color from the document image by executing layout recognition processing, and determines the first page of the document. A title recognizer extracts a title by executing title recognition processing and determines the first page of the document. A divider divides the document image into documents on the basis of the determined first page, converts the divided document images into files, and stores the files in a storage device.


