OCR-Based File Naming for Image Content Identification
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing image processing systems, such as digital color copy machines, automatically assign file names to image files based on date and time of scanning, which do not reflect the contents of the images, making it time-consuming for users to identify and search for specific files.
Innovation Solution
An information processing apparatus with a character identifying unit that performs optical character recognition (OCR) on image data to extract character strings, which are then used to set a file name that accurately reflects the image contents, allowing for automatic and informative naming of image files.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Extent of automation
If file names are created based on date and time of scanning, then file naming is automatic and simple, but the file name does not reflect the contents of the image
Solution Approach 1:
The patent extracts text information from the image content using OCR technology. The character identifying unit reads and extracts text strings from the scanned image, which are then used to create meaningful file names that reflect the actual content of the document rather than just temporal metadata.
Solution Approach 2:
The patent introduces an intermediary process between scanning and file naming: the character identifying unit acts as a mediator that bridges the image data and the file naming system. It extracts text from the image and transforms it into a format suitable for file naming, creating a meaningful connection between image content and file identifier.
2Productivity
If file names are created based on date and time, then the naming process is quick and simple, but users must open each file to see contents which is time consuming
Solution Approach 1:
The patent performs preliminary action by extracting and analyzing text content from the image during the scanning process itself. The character identifying unit reads the text and the file name setting unit creates a content-based file name before the user needs to access the file, so when users later search for files, they can immediately identify contents from the file name without opening each file.
Solution Approach 2:
The patent replaces the mechanical approach of manually opening files to check contents with an automated information processing system. The character identifying unit and file_name setting unit automatically extract and process text information, substituting manual inspection with automated optical character recognition and string processing.
3Ease of manufacture
If traditional automatic naming is used, then file naming is straightforward, but searching for required files is inconvenient
Solution Approach 1:
The patent implements self-service by enabling the system to automatically generate meaningful file names without user intervention. The character identifying unit extracts text from the image and the file_name setting unit creates an appropriate file name based on the extracted content, allowing the system to serve itself in the naming task while producing results that facilitate easy user search and identification.
Applied Scientific Principles
This section explains which scientific principles are used to turn an abstract innovation direction into a practical engineering solution.
Function Achieved in This Case
Enables users to automatically assign file names that accurately represent the image contents, improving file organization and retrieval efficiency by providing a name that reflects the image data, reducing the time and effort required to find specific files.
Implementation Method 1
a character identifying unit that performs character identification with respect to an image in an image data
Data Source
AI summary
An information processing apparatus includes a character identifying unit that performs character identification with respect to an image in an image data. A file name setting unit sets a file name of the image data based on a character or a character string obtained as a result of the character identification.


