Document Scanning Control Device for Adaptive File Format Selection
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing OCR software can only generate electronic files in a single format, failing to adapt to the characteristics of scanned documents, such as table images, which limits flexibility in file generation based on image content.
Innovation Solution
A system comprising a reading device and a control device with computer-readable instructions that determine whether a scanned image is a table image, generating files in different formats (spreadsheet, word processor, or PDF) based on the presence and size of table images within the scanned documents.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If existing OCR software generates electronic files in a single format, then the system complexity is reduced and ease of manufacture is improved, but the adaptability to different document characteristics (table vs. non-table images) deteriorates
Solution Approach 1:
The control device performs preliminary analysis of the scanned image to detect whether it contains a table structure before generating the final electronic file. This preliminary detection enables the system to select the appropriate file format (spreadsheet for tables, word processor or PDF for non-tables) in advance, achieving adaptability without requiring complex real-time format switching mechanisms
Solution Approach 2:
The patent segments the document processing task into distinct pathways based on image characteristics: one pathway for table images that generates spreadsheet files, and another pathway for non-table images that generates word processor or PDF files. This segmentation allows the system to maintain simplicity within each pathway while achieving overall versatility through conditional routing
2Adaptability or versatility
If the control device determines and generates different file formats based on image content, then the versatility and user satisfaction are improved, but the processing time and operational complexity increase
Solution Approach 1:
The system performs table structure detection as a preliminary step immediately after scanning, before any file generation occurs. This early detection prevents unnecessary processing of images that don't require specialized handling, reducing overall processing time while maintaining format versatility
Solution Approach 2:
The patent applies different processing qualities and methods to different parts of the input stream: table images receive structured analysis and are converted to spreadsheet format with optimized processing, while non-table images receive standard OCR processing and are converted to word processor or PDF format. This local quality approach ensures each image type receives appropriate processing intensity, minimizing total processing time
Data Source
AI summary
A system including a reading device that reads an image from a document and a control device that controls the reading device is disclosed herein. The control device includes a processor and a memory storing computer-readable instructions. The computer-readable instructions instruct the processor to determine whether a read image read by the reading device comprises a table image. The computer-readable instructions instruct the processor to generate a first file in a first file format when the processor determines that the read image comprises the table image. The computer-readable instructions instruct the processor to generate a second file in a second file format when the processor determines that the read image does not comprise the table image. The second file is different from the first file. The second file format is different from the first file format. Computer-readable media storing the computer-readable instructions and corresponding methods also are disclosed herein.


