MPEG-2 Systems Hierarchy Extension Descriptor for HEVC
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
The existing MPEG-2 Systems specification lacks support for HEVC extension bitstreams, particularly in signaling multiple direct dependent layers and is not generic enough to accommodate various scalability types, limiting its applicability to extensions beyond H.264/AVC.
Innovation Solution
The introduction of a hierarchy extension descriptor and HEVC extension descriptor that signal data for HEVC layers, enabling the description of multiple enhancement layers across different scalability dimensions, and the modification of the HEVC video descriptor to include operation points, profile, tier, and level indicators, frame packing information, and bitrate details.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If the MPEG-2 Systems specification is used for transporting video data, then compatibility with existing systems is maintained, but it cannot support HEVC extension bitstreams with multiple direct dependent layers across different scalability dimensions
Solution Approach 1:
The patent implements a nested descriptor structure where a hierarchy extension descriptor contains multiple HEVC extension descriptors, each representing different scalability dimensions. This nested organization allows the system to support complex multi-layered HEVC extensions while maintaining a structured approach that builds upon the existing MPEG-2 Systems framework, thereby improving adaptability without proportionally increasing overall system complexity.
Solution Approach 2:
The patent segments the video data structure into distinct descriptors: a hierarchy extension descriptor that organizes multiple scalability dimensions, and individual HEVC extension descriptors for each dimension. This segmentation allows each descriptor to handle specific aspects of HEVC extension support independently, making the system more adaptable while keeping individual components manageable in complexity.
2Adaptability or versatility
If the existing MPEG-2 Systems specification is used, then system structure remains stable, but it lacks the capability to signal multiple enhancement layers across different scalability dimensions
Solution Approach 1:
The patent extends the existing MPEG-2 descriptor framework by adding new dimensional capabilities through hierarchy extension descriptors that can represent multiple scalability dimensions (spatial, temporal, quality). This dimensional extension allows the system to signal enhancement layers across different scalability dimensions without disrupting the stable underlying MPEG-2 structure, thereby improving adaptability while preserving existing information structures.
3Adaptability or versatility
If HEVC extension descriptors are added to signal multiple scalability dimensions, then support for advanced video coding is improved, but the descriptor structure becomes more complex
Solution Approach 1:
The hierarchy extension descriptor serves as a universal container that can represent multiple different scalability dimensions (spatial, temporal, quality, multiview, 3D) through a unified structure. This multi-functional approach allows the same descriptor framework to support various HEVC extensions without requiring separate specialized structures for each type, thereby improving versatility while managing complexity through standardization.
Data Source
AI summary
In one example, a device for processing video data includes a memory for storing an enhancement layer of video data coded according to an extension of a video coding standard, and one or more processors configured to decode a hierarchy extension descriptor for an elementary stream including the enhancement layer, wherein the hierarchy extension descriptor includes data representative of two or more reference layers on which the enhancement layer depends, wherein the two or more reference layers include a first enhancement layer, conforming to a first scalability dimension, and a second enhancement layer, conforming to a second scalability dimension, and wherein the first scalability dimension is different than the second scalability dimension, and to process the video data based at least in part on the data representative of the two or more reference layers.


