Image Text Region Color Modification for Recognition Speed

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing technologies indirectly draw attention to text in images, slowing user recognition compared to direct attention methods.

Innovation Solution

An information processing device that acquires an image, divides it into text and background regions, and modifies color attributes to enhance visual recognition differences between character, character background, and background regions, using a visual recognition distance table to maximize differences in hue, brightness, or saturation.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If a cursor is displayed on an image including text containing region to draw user's attention, then the text containing region is highlighted, but the user's recognition of the text is slower compared to direct attention methods

Engineering Contradiction:
Improvetext recognition speedVSAvoidtime for text recognition
Core Design Contradiction:
Ease of operationVSLoss of time

Solution Approach 1:

The patent modifies the color attributes (hue, saturation, brightness) of the text containing region to enhance visual recognition. By calculating representative color values for the text region, character background region, and image background region, and then adjusting these colors to maximize visual recognition distance, the system directly draws attention to the text without requiring cursor-based interaction, thereby speeding up text recognition.

Inventive Principle:
Principle #32Color changes

Solution Approach 2:

The patent changes the visual parameters of the text containing region by modifying color attributes (hue, saturation, brightness) based on calculated representative values. The modification unit adjusts these parameters to maximize the visual recognition distance between the text region and surrounding areas, enabling direct attention drawing and improving text recognition speed without the delay associated with cursor-based methods.

Inventive Principle:
Principle #35Parameter changes

2Productivity

If color attributes of text containing region are modified to enhance visual recognition, then text recognition is quickened, but the original image impression may be altered

Engineering Contradiction:
Improvetext recognition speedVSAvoidoriginal image impression
Core Design Contradiction:
ProductivityVSStability of the object's composition

Solution Approach 1:

The patent applies color modification only to the text containing region while preserving the rest of the image. By calculating representative color values specifically for the text region and its surrounding areas, and modifying only those localized areas to maximize visual recognition distance, the system enhances text recognition speed without significantly altering the overall image impression, as the modifications are confined to specific regions rather than the entire image.

Inventive Principle:
Principle #3Local quality

Data Source

PatentUS9384557B2Information processing device, image modification method, and computer program product
Publication Date: 2016.07.05 TOSHIBA DIGITAL SOLUTIONS CORP
  • US9384557B2 patent drawing
  • US9384557B2 patent drawing
  • US9384557B2 patent drawing

AI summary

According to an embodiment, an information processing device includes: a first division unit divides an image into a text containing region and a background region other than the text containing region; a second division unit divides a text containing region into a character region constituted by lines forming characters and a character background region other than the character region; a calculator calculates a first representative value of an attribute of the character region, a second representative value of the attribute of the character background region, and a third representative value of the attribute of the background region; a modification unit makes modification so that a first difference based on the first and third representative values, a second difference based on the first and second representative values, and a third difference based on the second and third representative values become larger; and an output unit outputs a modified image.