Audio Stream Container Metadata for Selective Decoding
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing technologies face high processing loads when transmitting multiple types of audio data, such as channel and object encoded data, which can lead to inefficiencies in audio reproduction.
Innovation Solution
A transmission device and method that includes a transmission unit for sending a predetermined format container with audio streams and an information insertion unit to insert attribute information into the container, allowing selective decoding of necessary group encoded data based on attribute information and stream correspondence.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If multiple types of audio data (channel encoded data and object encoded data) are transmitted together, then acoustic reproduction with enhanced realistic feeling is achieved, but processing load at the reception side increases
Solution Approach 1:
The patent segments audio data into multiple groups (channel encoded data and object encoded data) and transmits them as separate group encoded data within the container. This segmentation allows the reception device to selectively decode only the necessary groups based on reproduction requirements, thereby maintaining high acoustic reproduction quality while reducing processing load by avoiding unnecessary decoding operations.
2Reliability
If all audio streams are decoded at the reception side, then complete audio reproduction is achieved, but processing time and computational resources increase
Solution Approach 1:
The patent performs preliminary organization of audio data by grouping related encoded data together and inserting identification information into the container structure before transmission. This preliminary action enables the reception device to quickly identify and select only the necessary group encoded data for decoding, significantly reducing processing time while ensuring complete audio reproduction through selective decoding of required groups.
Data Source
Figure 1
Figure 2
Figure 3
AI summary
A processing load of a reception side is reduced when a plurality of types of audio data is transmitted. A predetermined format container is transmitted having a predetermined number of audio streams including a plurality of group encoded data. For example, the plurality of group encoded data includes either or both of channel encoded data and object encoded data. Attribute information indicating an attribute of each of the plurality of group encoded data is inserted into a layer of the container. For example, stream correspondence information indicating an audio stream including each of the plurality of group encoded data is further inserted into the layer of the container.