Audio Signal Insertion in Downmixed Streams
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing audio processing technologies face challenges in efficiently and effectively inserting additional audio signals, such as system sounds, into downmixed audio signals that represent multiple audio objects, often resulting in audible artifacts and reduced spatial diversity during rendering.
Innovation Solution
A method and system that utilize an insertion unit to mix additional audio signals with downmix signals and modify associated metadata, ensuring seamless integration and maintaining spatial diversity by adapting upmix and object metadata, thereby avoiding artifacts and maintaining valid audio objects during decoding and rendering.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If additional audio signals are inserted into downmixed audio signals, then the audio functionality is enhanced, but audible artifacts and reduced spatial diversity occur
Solution Approach 1:
The patent applies preliminary action by detecting the insertion requirement before actually inserting the audio signal, and by pre-modifying the upmix metadata to prepare for the insertion. This advance preparation allows the system to seamlessly integrate additional audio signals while maintaining spatial diversity and avoiding audible artifacts during the actual insertion process.
2Adaptability or versatility
If additional audio signals are inserted into downmixed audio signals, then the audio functionality is enhanced, but spatial diversity is reduced
Solution Approach 1:
The patent applies parameter changes by dynamically modifying the upmix metadata parameters to reflect the insertion of additional audio signals. By adjusting these metadata parameters, the system maintains accurate spatial information and diversity characteristics even as new audio signals are integrated into the downmixed signal.
3Adaptability or versatility
If additional audio signals are inserted into downmixed audio signals, then the audio functionality is enhanced, but decoding and rendering validity is compromised
Solution Approach 1:
The patent applies feedback by detecting whether audio signal insertion is required, then modifying the downmix signal and upmix metadata accordingly, and finally ensuring the modified data maintains validity for subsequent decoding and rendering operations. This closed-loop approach ensures that the audio stream remains reliable and valid throughout the insertion process.
Data Source
Figure 1~2
Figure 3
AI summary
A method for inserting a first audio signal into a bitstream which comprises a downmix signal and associated bitstream metadata is described. The downmix signal and associated bitstream metadata are indicative of an audio program comprising a plurality of spatially diverse audio signals. The downmix signal comprises at least one audio channel and the bitstream metadata comprise upmix metadata for reproducing the plurality of spatially diverse audio signals from the at least one channel. The method comprises mixing the first audio signal with the at least one audio channel to generate a modified downmix signal. The method further comprises generating an output bitstream comprising the modified downmix signal and the associated modified bitstream metadata indicative of a modified audio program comprising a plurality of modified spatially diverse audio signals.