Scalable Video Encoding Inter-Layer Prediction Reference Selection
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current video encoding and decoding technologies face challenges in providing scalable video quality, particularly in efficiently performing inter-layer prediction in multilayer structures to accommodate varying video qualities and environments.
Innovation Solution
A method and apparatus for encoding and decoding scalable video in a multilayer structure, where an inter-layer reference picture is derived from a decoded picture in a reference layer, used to predict a current block, and reference information is transmitted to indicate available pictures for inter-layer prediction, allowing for effective inter-layer prediction and selective processing.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If inter-layer prediction is performed using all pictures in the reference layer, then prediction accuracy is improved, but processing complexity and memory requirements increase
Solution Approach 1:
The patent applies local quality by selectively marking specific pictures in the reference layer as available for inter-layer prediction using reference information, rather than uniformly processing all pictures. This allows the system to focus computational resources on only those reference pictures that are actually useful for prediction, improving efficiency while maintaining prediction accuracy.
Solution Approach 2:
The patent implements partial action by performing inter-layer prediction only on pictures that are marked as available through reference information. Instead of processing all pictures in the reference layer, the system selectively applies prediction operations only where needed, reducing overall processing complexity while maintaining effective prediction performance.
2Productivity
If reference information is transmitted to indicate available pictures for inter-layer prediction, then encoding efficiency is improved, but data transmission overhead increases
Solution Approach 1:
The patent extracts only the essential reference information needed to identify available pictures for inter-layer prediction, rather than transmitting complete picture data or exhaustive reference lists. This extraction approach provides sufficient information to enable efficient encoding while minimizing the overhead of transmitted data.
3Speed
If selective processing is performed on pictures in the reference layer, then processing speed is improved, but prediction accuracy may deteriorate
Solution Approach 1:
The patent applies preliminary action by pre-marking pictures in the reference layer with reference information that indicates their availability for inter-layer prediction. This preliminary classification allows the decoding process to quickly identify and process only the relevant reference pictures, improving processing speed while ensuring that prediction accuracy is maintained by selecting appropriate reference pictures in advance.
Data Source
AI summary
The present invention relates to a method for coding a scalable video in a multilayer structure, and a method for encoding a video according to the present invention comprises the steps of: decoding and saving a picture of a reference layer; inducing an interlayer reference picture which is referenced for predicting a current block of a current layer; producing a reference picture list including the interlayer reference picture and a reference picture of the current layer; conducting a prediction on the current block of the current layer with the reference picture list to induce a prediction sample with respect to the current block; inducing a recovery sample with respect to the current block based on the prediction sample and the prediction block with respect to the current block; and transmitting reference information for indicating a picture that can be used for interlayer prediction from the pictures of the reference layer.


