Sub-bitstream Extraction for Scalable Image Coding
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
The increasing demand for high-resolution and high-quality images and videos, such as UHD and immersive media, poses challenges in efficient compression, transmission, and storage due to the increased bit rate, necessitating improved image/video coding techniques for scalability and compression efficiency.
Innovation Solution
A multilayer-based coding method that configures a target output layer and derives a sub-bitstream for scalability through sub-bitstream extraction, using external means or signaled information to determine the target output layer and highest temporal layer, and applies temporal scalability, thereby optimizing image/video coding efficiency.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If high-resolution and high-quality image/video data are transmitted using existing wired or wireless wideband lines or storage media, then transmission and storage costs are increased
Solution Approach 1:
The patent divides the bitstream into multiple layers (base layer and enhancement layers) with different quality levels. This segmentation allows receivers to select and process only the necessary layers based on their capabilities and network conditions, reducing unnecessary transmission and storage costs while maintaining the option for high quality when needed.
Solution Approach 2:
The patent changes the quality parameter by providing multiple representation qualities through different enhancement layers. The same base layer can be combined with different enhancement layers to produce multiple quality outputs, allowing flexible adaptation to different transmission and storage requirements without permanently increasing costs.
2Productivity
If multi-layer coding techniques are applied to satisfy compression/transmission efficiency and scalability requirements, then information signaling complexity is increased
Solution Approach 1:
The patent extracts and separates the layer selection information from the main bitstream by using specific NAL unit types (e.g., VPS, SPS, PPS) that carry scalability parameters. This extraction allows the multi-layer coding structure to be established without burdening the main decoding process with excessive signaling complexity, as the layer information is organized in dedicated structures.
Solution Approach 2:
The patent performs preliminary organization of layer information through VPS (Video Parameter Set), SPS (Sequence_parameter Set), and PPS (Picture_parameter Set) structures before actual decoding. These pre-configured parameter sets contain all necessary scalability information, allowing receivers to understand the layer structure in advance and process only relevant layers, thereby reducing runtime signaling complexity.
Data Source
AI summary
According to embodiment(s) of the present document, multilayer-based coding can be performed. The multilayer-based coding can include a sub-bitstream extraction process. Inputs of the sub-bitstream extraction process can include a value related to a target OLS index and a value related to a highest temporal identifier.


