Image Area Sound Linkage for Multi-Voice Correlation

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing electronic devices face difficulties in linking and managing sound data with images, particularly in recognizing correlations between images and sound data, and in conveniently editing or playing back multiple voices linked to images.

Innovation Solution

An electronic device method and apparatus that allows users to select specific areas of an image to link sound data, selectively play back sound data, and convert sound data to text for display, enabling better image-sound data correlation and management.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Device complexity

If a single piece of sound data is linked to a single image, then the linkage is simple, but it becomes difficult to recognize the correlation between the image and the linked sound data when multiple persons are involved

Engineering Contradiction:
Improvelinkage complexityVSAvoidcorrelation recognition
Core Design Contradiction:
Device complexityVSLoss of information

Solution Approach 1:

The patent divides the image into multiple selectable areas (first area and second area) and links different sound data (first sound data and second sound data) to each area separately. This segmentation allows users to clearly identify which sound corresponds to which person or region in the image, resolving the correlation recognition problem while maintaining simple individual linkages.

Inventive Principle:
Principle #1Segmentation

2Adaptability or versatility

If a user wants to link multiple persons' voices to an image, the user should record voices sequentially or edit multiple sound data files, which is inconvenient

Engineering Contradiction:
Improvemulti-voice linkage capabilityVSAvoidoperation convenience
Core Design Contradiction:
Adaptability or versatilityVSEase of operation

Solution Approach 1:

The patent enables independent selection and linkage of sound data to different areas of the image. Users can select the first area and link first sound data, then select the second area and link second sound data, without sequential recording or complex editing. This area-based segmentation simplifies the operation while supporting multiple voice linkages.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent introduces image areas as intermediary elements between the user and sound data. Instead of directly managing multiple sound data files, users interact with visual areas in the image, which serve as mediators to associate sounds with specific persons or objects, making the process more intuitive and convenient.

Inventive Principle:
Principle #24Intermediary (Mediator)

3Ease of manufacture

If sound data is linked to the entire image, then the linkage is straightforward, but the user cannot selectively play back sound data from specific persons or areas

Engineering Contradiction:
Improvelinkage simplicityVSAvoidselective playback capability
Core Design Contradiction:
Ease of manufactureVSAdaptability or versatility

Solution Approach 1:

The patent segments the image into multiple selectable areas, each capable of having sound data linked independently. This allows the system to maintain the simplicity of direct linkage while enabling selective playback, as users can choose to play back sound data associated with specific areas rather than the entire image.

Inventive Principle:
Principle #1Segmentation

4Quantity of substance

If the entire sound data is played back, then all sound information is provided, but the user cannot selectively listen to specific persons' voices

Engineering Contradiction:
Improvesound information completenessVSAvoidselective listening
Core Design Contradiction:
Quantity of substanceVSEase of operation

Solution Approach 1:

The patent associates different sound data with different areas of the image, enabling selective playback. Users can choose to play back sound data from specific areas (e.g., first person's voice or second person's voice) while still having access to all sound information if needed, thus providing both completeness and selectivity.

Inventive Principle:
Principle #1Segmentation

Data Source

PatentUS10684754B2Method of providing visual sound image and electronic device implementing the same
Publication Date: 2020.06.16 SAMSUNG ELECTRONICS CO LTD
  • US10684754B2 patent drawing
  • US10684754B2 patent drawing
  • US10684754B2 patent drawing

AI summary

A method of providing a visual sound image, which may generate, edit, and play back a visual sound image in which sound data is linked to an image, and an electronic device implementing the same are provided. The method includes an electronic device including a display, an image including at least one object on the display, receiving, by the electronic device, a selection of at least a certain area of the object in the image displayed on the display or a selection of a certain area of the image, and linking, by the electronic device, sound data to the at least the certain area of the object or the certain area of the image. In addition, various embodiments are possible.