Output Layer Sets for Spatial/SNR Video Layer Selection
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing video coding systems require decoders to always support the highest encoded layer, leading to scalability issues as they cannot adapt to intermediate layers based on hardware and network requirements, resulting in errors and inefficient resource utilization.
Innovation Solution
The use of Output Layer Sets (OLSs) to support spatial and SNR scalability, where an ols_mode_idc syntax element in the Video Parameter Set (VPS) indicates the number of OLSs, with each OLS containing specific layers, allowing the decoder to determine the output layer quickly and decode only the necessary layers.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Manufacturing precision
If decoders are configured to always support the highest encoded layer, then video quality is maximized, but system adaptability to intermediate layers deteriorates
Solution Approach 1:
The patent segments the video data into multiple independent layers with different quality levels. Each layer can be decoded independently or in combination with other layers, allowing decoders to select appropriate layers based on hardware capabilities and network conditions. This segmentation enables intermediate layer decoding without requiring support for the highest layer.
Solution Approach 2:
The patent introduces dynamic layer selection mechanisms where decoders can adaptively choose which layers to decode based on real-time hardware capabilities and network conditions. The system dynamically adjusts the decoding configuration to match available resources, enabling flexible adaptation to intermediate layers while maintaining optimal video quality when resources permit.
2Adaptability or versatility
If multiple layers are encoded to support different quality levels, then scalability is improved, but bitstream complexity increases
Solution Approach 1:
The patent merges multiple layers into a unified bitstream structure with standardized syntax elements. The VPS (Video Parameter Set) contains consolidated layer configuration information, and OLS (Output Layer Set) structures group layers systematically. This merging reduces bitstream complexity by eliminating redundant signaling while maintaining scalability across different quality levels.
Solution Approach 2:
The patent uses parameter-based layer differentiation where layers are defined by specific parameter sets (resolution, SNR, frame rate) rather than complex structural variations. The ols_mode_idc parameter and related syntax elements provide compact signaling that changes parameters efficiently to indicate layer configurations, reducing bitstream overhead while supporting multiple quality levels.
3Manufacturing precision
If all encoded layers are decoded, then maximum video quality is achieved, but resource utilization efficiency deteriorates
Solution Approach 1:
The patent extracts and processes only the necessary layers required for the desired output quality. Instead of decoding all encoded layers, the system selectively extracts the appropriate layer or combination of layers needed to meet quality requirements. This extraction approach reduces computational resources and energy consumption while maintaining adequate video quality.
Solution Approach 2:
The patent implements partial decoding where only the necessary portion of the multi-layer bitstream is processed. Decoders can perform partial action by decoding only up to the required quality level without processing higher quality layers. This partial action approach optimizes resource utilization by avoiding unnecessary decoding operations while still achieving the target video quality.
Data Source
AI summary
A video coding mechanism is disclosed. The mechanism includes encoding a bitstream comprising one or more layers of coded pictures. A video parameter set (VPS) is also encoded into the bitstream. The VPS includes an output layer set (OLS) mode identification code (ols_mode_idc) specifying that a total number of OLSs specified by the VPS is equal to a number of layers specified by the VPS. The bitstream is stored for communication toward a decoder.


