Tag Information Extraction and Document Image Segmentation
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing image processing technologies fail to effectively separate document images from tags attached to them, leading to difficulties in maintaining tag information, requiring manual efforts for classification, and risking document loss or tag detachment during scanning.
Innovation Solution
An image processing device and method that extracts tag information, removes tags from document images, and generates separate images without tags, while associating tag information with the document images, allowing for grouping and outputting of tag information and document images independently.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Productivity
If a document with a tag is read by a scanner, then the document can be converted into data, but the tag portion appears dirty and may not be desired on the document image
Solution Approach 1:
The patent segments the document image processing into two distinct outputs: a clean document image with the tag portion removed, and a separate tag information image containing the extracted tag data. This segmentation allows the document image to be free from tag contamination while preserving tag information separately.
Solution Approach 2:
The patent extracts the tag information from the document image and removes the tag portion from the document image. The tag information is then placed in a separate tag information image, effectively taking out the harmful element (tag appearance) while preserving its informational value.
2Ease of operation
If manual efforts are used for classification based on tag type, then documents can be sorted, but there is a high risk of losing documents and peeling off tags
Solution Approach 1:
The patent replaces the manual mechanical classification process with an automated image processing system. The system automatically detects tags in the document images, extracts tag information, and performs classification without human intervention, thereby eliminating the risks of document loss and tag detachment associated with manual handling.
Solution Approach 2:
The patent creates a digital copy of the tag information in the tag information image, preserving the classification data without requiring physical tags to remain attached to documents during handling and sorting processes.
3Object-affected harmful factors
If tag information is removed from the document image, then the document image is cleaned, but it becomes difficult to trace the history of the tag
Solution Approach 1:
The patent moves tag information from the spatial dimension (visual appearance in document image) to a separate informational dimension (tag information image). This dimensional separation allows the document image to be clean while the tag history is preserved in a different format and location.
Solution Approach 2:
The patent introduces a tag information image as an intermediary that carries the tag history information. This intermediary preserves the tag information that would otherwise be lost when the tag portion is removed from the document image.
4Productivity
If documents are grouped using delimitation by tag, then classification is achieved, but it cannot be employed if documents are not arranged in order
Solution Approach 1:
The patent replaces the mechanical arrangement requirement (documents must be physically ordered) with an automated image processing system that can detect and analyze tags regardless of document sequence. The system processes documents in any order and performs grouping based on tag information extraction and analysis.
Data Source
AI summary
An image processing device reads a plurality of documents selectively attached with a tag, and processes the read document images with a tag. The image processing device includes a tag information extractor that extracts tag information about a tag from the document images with a tag, and an image processor that separately generates, based on the tag information, document images without a tag in which a tag is removed from the document images with a tag, and a tag information image in which the tag information about the document images without a tag is written.


