Image Area Sound Linkage for Multi-Voice Correlation
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing electronic devices face difficulties in linking and managing sound data with images, particularly in recognizing correlations between images and sound data, and in conveniently editing or playing back multiple voices linked to images.
Innovation Solution
An electronic device method and apparatus that allows users to select specific areas of an image to link sound data, selectively play back sound data, and convert sound data to text for display, enabling better image-sound data correlation and management.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Device complexity
If a single piece of sound data is linked to a single image, then the linkage is simple, but it becomes difficult to recognize the correlation between the image and the linked sound data when multiple persons are involved
Solution Approach 1:
The patent divides the image into multiple selectable areas (first area and second area) and links different sound data (first sound data and second sound data) to each area separately. This segmentation allows users to clearly identify which sound corresponds to which person or region in the image, resolving the correlation recognition problem while maintaining simple individual linkages.
2Adaptability or versatility
If a user wants to link multiple persons' voices to an image, the user should record voices sequentially or edit multiple sound data files, which is inconvenient
Solution Approach 1:
The patent enables independent selection and linkage of sound data to different areas of the image. Users can select the first area and link first sound data, then select the second area and link second sound data, without sequential recording or complex editing. This area-based segmentation simplifies the operation while supporting multiple voice linkages.
Solution Approach 2:
The patent introduces image areas as intermediary elements between the user and sound data. Instead of directly managing multiple sound data files, users interact with visual areas in the image, which serve as mediators to associate sounds with specific persons or objects, making the process more intuitive and convenient.
3Ease of manufacture
If sound data is linked to the entire image, then the linkage is straightforward, but the user cannot selectively play back sound data from specific persons or areas
Solution Approach 1:
The patent segments the image into multiple selectable areas, each capable of having sound data linked independently. This allows the system to maintain the simplicity of direct linkage while enabling selective playback, as users can choose to play back sound data associated with specific areas rather than the entire image.
4Quantity of substance
If the entire sound data is played back, then all sound information is provided, but the user cannot selectively listen to specific persons' voices
Solution Approach 1:
The patent associates different sound data with different areas of the image, enabling selective playback. Users can choose to play back sound data from specific areas (e.g., first person's voice or second person's voice) while still having access to all sound information if needed, thus providing both completeness and selectivity.
Data Source
AI summary
A method of providing a visual sound image, which may generate, edit, and play back a visual sound image in which sound data is linked to an image, and an electronic device implementing the same are provided. The method includes an electronic device including a display, an image including at least one object on the display, receiving, by the electronic device, a selection of at least a certain area of the object in the image displayed on the display or a selection of a certain area of the image, and linking, by the electronic device, sound data to the at least the certain area of the object or the certain area of the image. In addition, various embodiments are possible.


