Image Processing Apparatus Document Segmentation Metadata Integration
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing information processing systems lack the capability to efficiently divide and process images of multiple documents, such as business cards, and integrate additional information common to these images into compatible data formats for further use.
Innovation Solution
An information processing apparatus with a dividing unit to split read images into individual documents, an acquiring unit to extract common information, and an addition processing unit to incorporate this information into compatible data files, enabling seamless processing and formatting for use across various platforms.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Productivity
If a read image containing multiple documents is processed as a single image, then the processing is simple, but the individual document information cannot be extracted and utilized separately
Solution Approach 1:
The patent applies segmentation by dividing a read image containing multiple documents into separate divided images through automatic document separation. The system detects document boundaries and splits the composite image into individual document images, enabling separate processing and utilization of each document while maintaining automated high-efficiency workflow.
2Loss of information
If additional information is manually added to each document image, then the information accuracy is high, but the processing time and labor cost increase
Solution Approach 1:
The patent implements self-service by enabling the system to automatically extract and add metadata (such as document type, date, and other common information) to each divided image without manual intervention. The system autonomously performs information extraction and attachment processes, ensuring information completeness while eliminating manual labor and reducing processing time.
Solution Approach 2:
The patent applies preliminary action by pre-extracting common information from the original read image or from recognized patterns before dividing the images. This extracted information is then automatically attached to each divided image in advance, so that when individual documents are processed later, the metadata is already available, eliminating the need for repeated manual information entry.
3Adaptability or versatility
If different document formats are processed separately, then the format-specific processing is optimized, but the integration and统一管理 of multiple formats becomes complex
Solution Approach 1:
The patent applies universality by creating a unified processing framework that handles multiple document formats (PDF, XDW, TIFF, etc.) through a common architecture. The system provides format-agnostic document separation and metadata attachment capabilities that work across different file types, while allowing format-specific optimizations where needed, thereby achieving both format versatility and simplified unified management.
Data Source
AI summary
An information processing apparatus includes a dividing unit configured to divide a read image into plural divided images, an acquiring unit configured to acquire additional information including a content common to the plural divided images, and an addition processing unit configured to perform addition processing for adding the acquired additional information to plural data files respectively including the divided images.


