Volumetric Video Representation Grouping for Adaptive Streaming
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing technologies face challenges in efficiently encoding and decoding volumetric video content for applications like AR, VR, and MR, particularly in handling the large data volumes and complex interactions of 3D scenes, which require innovative methods to manage viewer interactions and data transmission.
Innovation Solution
The proposed solution involves generating a description file that groups media component bitstreams based on identified relationships, allowing for adaptive delivery of volumetric media content to receivers, with mechanisms to handle misunderstandings in the grouping description and ensuring efficient decoding and display of volumetric video.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If multiple representations of volumetric media content are provided with detailed grouping information, then the adaptability and precision of content delivery are improved, but the complexity of the description file and processing requirements increase
Solution Approach 1:
The patent segments the volumetric media content into multiple representations (e.g., different resolutions, quality levels) and organizes them into groups with hierarchical relationships. Each group is described using structured information that breaks down the complex content into manageable units, enabling selective delivery based on viewer capabilities while maintaining organization through systematic grouping.
Solution Approach 2:
The patent performs preliminary organization of content by pre-establishing grouping relationships and description files before actual delivery. The description file is generated in advance with all necessary grouping information, allowing the receiver to understand the content structure without real-time processing complexity, thus preparing the groundwork for efficient adaptive delivery.
2Measurement precision
If comprehensive grouping information is provided for multiple representations, then the precision of representation selection is improved, but the loss of time in processing and negotiation increases
Solution Approach 1:
The grouping information and representation relationships are pre-organized in the description file before delivery. This preliminary structuring allows the receiver to quickly parse and understand the content hierarchy without time-consuming real-time analysis, enabling precise selection of appropriate representations while minimizing negotiation time.
Solution Approach 2:
The patent uses standardized description formats and pre-defined grouping structures that act as templates or copies of proven effective organization schemes. This allows receivers to familiarly process the information using established parsing logic, reducing processing time while maintaining precise interpretation of the volumetric content relationships.
Data Source
AI summary
The embodiments relate to methods for encoding and decoding, and technical equipment for the same. The method for encoding obtaining two or more representations of volumetric media content, each of the two or more representations of volumetric media content including multiple media component bitstreams; identifying a relationship between the two or more representations; generating a description file containing a media component description for each media component bitstream; group descriptions grouping the media component descriptions for each media component bitstream; grouping information descriptions providing information on the identified relationship between group descriptions representing representations. The description file is provided to a receiver.


