Camera Zoom Adjustment for Character Recognition in OCR Systems

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional camera-based Optical Character Recognition (OCR) systems face inefficiencies and recognition failures due to excessively large or small character sizes in images, which are not within a predetermined range, leading to reduced recognition efficiency.

Innovation Solution

A method and apparatus that automatically adjust the zoom ratio of a camera to resize characters within a preset range for precise recognition, utilizing a camera module, recognizer module, OCR recognition engine, and dictionary module to detect character sizes and apply zoom functions as needed.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If Snapshot OCR is used with high-resolution images for accurate character recognition, then character recognition capability is improved, but recognition time increases

Engineering Contradiction:
Improvecharacter recognition capabilityVSAvoidrecognition time
Core Design Contradiction:
ReliabilityVSLoss of time

Solution Approach 1:

The patent applies partial action by performing OCR on only the selected region containing characters rather than processing the entire high-resolution image. The system allows users to select a specific region of interest, and the OCR engine processes only that portion, reducing computation time while maintaining recognition accuracy for the target characters.

Inventive Principle:
Principle #16Partial or excessive action

2Adaptability or versatility

If characters in the image are excessively large or small, then image capture flexibility is improved, but recognition efficiency deteriorates

Engineering Contradiction:
Improveimage capture flexibilityVSAvoidrecognition efficiency
Core Design Contradiction:
Adaptability or versatilityVSProductivity

Solution Approach 1:

The patent implements dynamic region selection where the system automatically identifies and adjusts the region of interest based on character detection. The OCR engine dynamically adapts to varying character sizes by selecting appropriate regions, allowing the system to handle both large and small characters effectively while maintaining recognition efficiency.

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The system changes the parameter of region selection based on character size detection. When characters are detected to be outside the optimal recognition range, the system adjusts the selected region parameters to focus on areas with appropriately sized characters, thereby maintaining recognition efficiency across varying character scales.

Inventive Principle:
Principle #35Parameter changes

3Reliability

If full recognition is performed on all characters in the image, then completeness of recognition is improved, but processing time increases

Engineering Contradiction:
Improvecompleteness of recognitionVSAvoidprocessing time
Core Design Contradiction:
ReliabilityVSLoss of time

Solution Approach 1:

The patent applies segmentation by dividing the image into a selected region of interest and the rest of the image. The OCR processing is segmented to operate only on the selected region containing characters, while other regions are excluded from processing. This segmentation maintains completeness for the target characters while significantly reducing processing time.

Inventive Principle:
Principle #1Segmentation

Data Source

PatentUS10079978B2Apparatus and method for automatically adjusting size of characters using camera
Publication Date: 2018.09.18 SAMSUNG ELECTRONICS CO LTD
  • US10079978B2 patent drawing
  • US10079978B2 patent drawing
  • US10079978B2 patent drawing

AI summary

A method is provided for automatically adjusting a size of characters using a camera. The method includes receiving an image with characters; adjusting a focus of the image with characters and detecting a region and a size of characters in the image; determining whether the size of the characters in the image falls within a preset range; recognizing the characters in the image and displaying the recognition results, if the size of the characters falls within the preset range; and automatically adjusting a zoom ratio of the image and recognizing the characters in the resized image, if the size of the characters does not fall within the preset range.