Image Audio Object Mapping for Selective Multimedia Output
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing electronic devices lack the capability to effectively combine and selectively output audio corresponding to specific subjects within an image, limiting the multimedia experience by not seamlessly integrating visual and auditory elements.
Innovation Solution
An electronic device method that extracts and maps audio objects to image objects based on features, storing a combination data set including image and audio data, and corresponding relationship information, allowing for the display of images with associated audio output when specific image objects are selected.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If audio data is combined with image data and selectively output based on image object selection, then the multimedia experience and integration of visual and auditory elements is enhanced, but the device complexity and data processing requirements increase
Solution Approach 1:
The patent segments audio data into multiple audio objects that correspond to different image objects within the image. Each audio object is independently associated with specific image objects through mapping data, allowing selective output based on user selection. This segmentation enables the system to handle complex multimedia content by breaking it down into manageable, independently controllable units.
Solution Approach 2:
The patent introduces mapping data as an intermediary component that establishes correspondence relationships between image objects and audio objects. This mapping data serves as a bridge, allowing the system to selectively associate and output specific audio objects based on which image objects are selected, without requiring direct complex integration of all audio and image data.
2Measurement precision
If audio objects are extracted and mapped to image objects based on features, then the precision of audio-image correspondence is improved, but the data processing time and computational requirements increase
Solution Approach 1:
The patent performs preliminary extraction of audio objects and creation of mapping data during the data collection and preparation phase, before actual playback or user interaction. By pre-processing the audio data and establishing correspondence relationships in advance, the system reduces real-time processing requirements when users interact with the multimedia content.
Solution Approach 2:
The patent creates simplified representations or copies of audio objects that are mapped to image objects based on extracted features. Instead of processing the entire audio data stream in real-time, the system works with pre-extracted audio object copies that contain the essential characteristics needed for correspondence with image objects, reducing computational burden.
Data Source
AI summary
A method for generating an image combined with audio, and image display and audio output includes displaying an image, when a first image object within the image is selected, outputting a first audio object corresponding to the first image object and, when a second image object within the image is selected, outputting a second audio object corresponding to the second image object.


