Multichannel Downmix Audio Object Coding Separation
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing audio object coding methods fail to effectively decode multiple objects from an encoded multi-object signal using multichannel downmix and additional control data, particularly when dealing with more than one downmix channel, limiting the separation quality and flexibility of object rendering.
Innovation Solution
A method that jointly decodes multiple downmix channels using a spatial audio object coder, which generates downmix information and object parameters to reconstruct audio objects, allowing for flexible scaling of separation quality and incorporating correlation between objects, enabling the use of multiple downmix channels for improved object separation.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Device complexity
If single-channel downmix coding is used, then device complexity is reduced, but separation quality of audio objects deteriorates
Solution Approach 1:
The patent transitions from single-channel downmix coding to multichannel downmix coding, adding spatial dimensionality to the coding process. By distributing audio objects across multiple downmix channels and using multichannel coherence parameters, the system achieves better separation quality while maintaining manageable complexity through efficient parameter representation.
2Measurement precision
If multichannel downmix coding is used, then separation quality of audio objects is improved, but device complexity increases
Solution Approach 1:
The patent introduces multichannel coherence parameters and object-specific downmix parameters as new parameter types that capture spatial relationships between multiple downmix channels. These parameter changes enable the decoder to separate audio objects effectively from multichannel downmixes without requiring complex processing, thus improving separation quality while controlling device complexity.
3Adaptability or versatility
If existing audio object coding methods are used, then compatibility with single-channel devices is maintained, but flexibility in rendering configurations is limited
Solution Approach 1:
The patent creates a universal coding framework that can represent audio objects in multichannel downmixes while maintaining compatibility with single-channel playback. The encoder can adaptively choose to encode objects into single-channel or multichannel downmixes based on the capabilities of the playback device, making the system universally applicable across different device types while providing enhanced flexibility for multichannel rendering when available.
Data Source
AI summary
An audio object coder for generating an encoded object signal using a plurality of audio objects includes a downmix information generator for generating downmix information indicating a distribution of the plurality of audio objects into at least two downmix channels, an audio object parameter generator for generating object parameters for the audio objects, and an output interface for generating the imported audio output signal using the downmix information and the object parameters. An audio synthesizer uses the downmix information for generating output data usable for creating a plurality of output channels of the predefined audio output configuration.


