Document Image Capture with Spatial Positioning for OCR Accuracy

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

The ability to extract useful information from images, especially for software applications, is often restricted by image quality, leading to errors and time-consuming post-acquisition operations such as editing and re-capture, which limits user willingness to use image-based software applications.

Innovation Solution

An electronic device captures multiple images of a document with integrated sensors, including accelerometers and gyroscopes, before the user activates the image-activation mechanism, storing images with timestamps and spatial-position information, and analyzes them using optical character recognition to improve accuracy and reduce user effort.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If multiple images are captured before user activation, then information extraction accuracy is improved, but device complexity increases

Engineering Contradiction:
Improveinformation extraction accuracyVSAvoiddevice complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The system captures multiple images of the document before the user activates the image-activation mechanism. This preliminary action ensures that high-quality images are already available for analysis, improving information extraction accuracy without requiring complex post-capture operations.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The electronic device automatically captures and processes multiple images without requiring user intervention or knowledge. The sensor integrates spatial-position information and timestamps automatically, and the system自行 selects the best images for analysis, making the process self-service and reducing operational complexity.

Inventive Principle:
Principle #25Self-service

2Measurement precision

If multiple images are captured and analyzed, then information extraction accuracy is improved, but time consumption increases

Engineering Contradiction:
Improveinformation extraction accuracyVSAvoidtime consumption
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

Multiple images are captured in advance during the time interval before user activation. This preliminary capture eliminates the need for post-acquisition re-capture operations, and the automated selection process quickly identifies the best images for analysis, reducing overall time consumption.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system skips unnecessary post-acquisition operations by having already captured multiple images with varying quality. The automated image selection process quickly identifies suitable images for analysis, rushing through the selection phase and eliminating time-wasting manual editing and re-capture operations.

Inventive Principle:
Principle #21Skipping (Rushing through)

3Ease of operation

If images are captured without user knowledge, then ease of operation is improved, but reliability may worsen due to potential image quality issues

Engineering Contradiction:
Improveease of operationVSAvoidimage quality reliability
Core Design Contradiction:
Ease of operationVSReliability

Solution Approach 1:

The system performs preliminary image capture without user knowledge during the preparation phase. Multiple images are captured in advance, ensuring that at least some high-quality images are available for analysis, thereby maintaining reliability while improving ease of operation.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system captures more images than the single image a user would typically take. By capturing multiple images during the time interval before activation, the system ensures that even if some images are of poor quality, there will be sufficient high-quality images available for accurate analysis, compensating for potential quality issues.

Inventive Principle:
Principle #16Partial or excessive action

Data Source

PatentUS9471833B1Character recognition using images at different angles
Publication Date: 2016.10.18 INTUIT INC
  • US9471833B1 patent drawing
  • US9471833B1 patent drawing
  • US9471833B1 patent drawing

AI summary

During this information-extraction technique, a user of the electronic device may be instructed by an application executed by the electronic device (such as a software application) to point an imaging sensor, which is integrated into the electronic device, at a location on a document. For example, the user may be instructed to point a cellular-telephone camera at a field on an invoice. After providing the instruction and before the user activates an image-activation mechanism associated with the imaging device, the electronic device captures multiple images of the document by communicating a signal to the imaging device to acquire the images. Then, the electronic device stores the images with associated timestamps and spatial-position information, which is provided by a sensor which is integrated into the electronic device. After the user activates the image-activation mechanism, the electronic device analyzes the images to extract the information proximate to the location on the document.