Output Layer Set Signaling for Scalable Video Layer Selection
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing video coding standards, such as HEVC and VVC, lack efficient mechanisms for signaling output layer sets in scalable video streams, which hinders optimal decoding and rendering of video data.
Innovation Solution
A method and device for decoding an encoded video bitstream that involves obtaining a coded video sequence with output layer sets, using flags and syntax elements to determine the layer set mode and select appropriate output layers, enabling efficient decoding and rendering.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Productivity
If existing video coding standards (HEVC, VVC) are used without enhanced signaling mechanisms, then the decoding process is simpler, but the efficiency of decoding and rendering scalable video streams is reduced
Solution Approach 1:
The patent segments the scalable video stream into multiple output layer sets, where each layer set contains layers with specific spatio-temporal or quality characteristics. This segmentation allows the decoder to process and render different layer sets independently, improving decoding efficiency by enabling parallel processing and selective rendering based on device capabilities and network conditions.
Solution Approach 2:
The patent introduces dynamic signaling mechanisms including flags (e.g., output_layer_set_flag, layer_set_runtime_selection_enabled_flag) and syntax elements that allow the encoder to adaptively indicate which layers should be output at different times. This dynamic approach enables real-time adjustment of output layers based on runtime conditions, improving productivity without requiring fixed complex structures.
2Manufacturing precision
If multiple output layers are selected and processed, then video quality is improved, but computational resources increase
Solution Approach 1:
The patent enables partial processing by allowing the decoder to select and process only certain output layers from each layer set based on capabilities and requirements. The runtime selection mechanism allows the system to process exactly the necessary number of layers (neither too few nor too many), optimizing the balance between video quality and computational resource usage by avoiding excessive processing of all possible layers.
Solution Approach 2:
The patent changes the parameter of layer selection dynamically through syntax elements and flags that indicate which layers should be output. By modifying these selection parameters at runtime based on device capabilities, network bandwidth, and quality requirements, the system can adjust computational resource usage while maintaining optimal video quality for different operating conditions.
3Adaptability or versatility
If runtime selection of output layers is enabled, then adaptability to different devices and conditions is improved, but the complexity of the decoding process increases
Solution Approach 1:
The patent applies preliminary action by having the encoder pre-organize the scalable video stream into structured output layer sets with clear syntax elements and flags indicating layer relationships and selection criteria. This preliminary organization at encoding time reduces the complexity of runtime selection, as the decoder receives pre-processed information about which layers can be output and under what conditions, rather than having to make complex decisions without prior guidance.
Data Source
AI summary
A method of decoding an encoded video bitstream using at least one processor includes obtaining a coded video sequence including a plurality of output layer sets from the encoded video bitstream; obtaining a first flag indicating whether each output layer set of the plurality of output layer sets includes more than one layer; based on the first flag indicating that the each output layer set includes more than the one layer, obtaining a first syntax element indicating an output layer set mode; selecting at least one layer from among layers included in the plurality of output layer sets as at least one output layer based on at least one of the first flag and the first syntax element; and outputting the at least one output layer.


