Object Audio Playback Apparatus With Scene Description Separation
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing audio technologies cannot individually control or manipulate sound sources from multiple objects in a synthesized audio signal, limiting user interaction and customization in audio content production.
Innovation Solution
An apparatus that separates and decodes scene description and object audio compression data from input files, allowing for individual audio effect addition to each object's signal based on corresponding scene description information, enabling realistic object audio playback and production.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If audio signals from multiple objects are synthesized into one sound source, then the audio service can be provided through conventional channels, but the user cannot individually control or manipulate sound sources from specific objects
Solution Approach 1:
The patent separates the audio signal into multiple independent object audio signals, each corresponding to a specific object in the audio content. This segmentation allows users to individually control and manipulate each sound source while maintaining the overall audio service framework.
Solution Approach 2:
The patent introduces a new dimension of object-based audio processing by adding scene description information that identifies and characterizes individual objects within the audio content, enabling granular control beyond traditional channel-based audio.
2Ease of operation
If object-based audio service technology is implemented to allow individual object control, then user interaction and customization are improved, but the requirement for scene description information and integrated analysis technology increases system complexity
Solution Approach 1:
The patent performs integrated analysis of audio signals and scene description information in advance, extracting and storing object characteristics and relationships before playback. This preliminary processing reduces the complexity during actual user interaction while maintaining advanced functionality.
Solution Approach 2:
The patent introduces an object audio effect unit as an intermediary component that bridges the gap between raw audio signals and user control interfaces, handling the complex analysis and processing tasks while presenting simplified controls to users.
3Reliability
If conventional audio codecs are used for compression, then compatibility is maintained, but the ability to preserve and process individual object audio signals and their corresponding scene description information is limited
Solution Approach 1:
The patent separates scene description data and object audio compression data into distinct streams during the deformation process, allowing each to be independently decoded and processed while maintaining their associations through object identifiers.
Solution Approach 2:
The patent implements a dynamic decoding process that adaptively processes scene description and audio objects based on their individual characteristics and the desired output format, allowing flexibility in handling different codec types and object configurations.
Data Source
AI summary
Disclosed is an apparatus for playing and producing realistic object audio. The apparatus for playing realistic object audio includes: a deformatter unit individually separating scene description (SD) compression data and object audio compression data from inputted audio files; an SD decoding unit decoding the SD compression data to restore SD information; an object audio decoding unit decoding the object audio compression data to restore object audio signals which are respective audio signals of a plurality of objects; and an object audio effect unit adding an audio effect for each object to the object audio signals according to SD information for each object corresponding to the object audio signals among the SD information to produce a realistic object audio signal corresponding to each of the object audio signals.


