Scalable Video Coding Inter-Layer Prediction Restriction
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Scalable video coding methods, such as single-loop decoding, face inefficiencies due to the need to transmit entire inter base layer pictures for enhancement layer decoding, even if only a small portion of data is required, and struggle with standard scalability cases where different coding standards are used for base and enhancement layers, like H.264/AVC and HEVC.
Innovation Solution
A sequence level indication is provided to restrict inter-layer prediction in scalable video coding to only intra-coded pictures or random access point pictures in the base layer, allowing for efficient encoding and decoding by discarding unnecessary pictures and simplifying motion vector prediction processes.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If entire inter base layer pictures are transmitted for enhancement layer decoding, then decoding accuracy is improved, but data transmission volume and processing complexity increase
Solution Approach 1:
The patent extracts and transmits only the necessary portions of base layer pictures (intra-coded pictures or RAP pictures) that are required for enhancement layer decoding, rather than transmitting entire inter base layer pictures. This selective extraction reduces data transmission volume while maintaining the decoding accuracy needed for enhancement layer reconstruction.
Solution Approach 2:
The patent segments the base layer pictures into different types (intra-coded pictures and RAP pictures versus other inter-coded pictures) and selectively processes only the relevant segments for enhancement layer decoding. This segmentation allows the system to discard unnecessary picture data while preserving the essential components needed for accurate enhancement layer reconstruction.
2Measurement precision
If entire inter base layer pictures are transmitted for enhancement layer decoding, then decoding accuracy is improved, but processing complexity increases
Solution Approach 1:
The patent extracts only the necessary base layer picture types (intra-coded and RAP pictures) required for enhancement layer decoding, eliminating the need to process entire inter base layer pictures. This extraction approach reduces processing complexity by removing unnecessary decoding and motion compensation operations for irrelevant picture data.
Solution Approach 2:
The patent divides base layer pictures into different categories and selectively processes only the relevant segments (intra-coded and RAP pictures) for enhancement layer decoding. This segmentation simplifies the processing pipeline by avoiding complex motion compensation and prediction operations on unnecessary inter-coded pictures.
3Adaptability or versatility
If syntax elements from all base layer pictures are decoded, then compatibility with different coding standards is improved, but memory requirements increase
Solution Approach 1:
The patent extracts and retains only the syntax elements from intra-coded base layer pictures and RAP pictures that are necessary for enhancement layer decoding and cross-standard compatibility. This selective extraction reduces memory requirements by discarding syntax elements from other inter-coded pictures that are not needed for enhancement layer reconstruction or standard compatibility.
Data Source
AI summary
There are disclosed various methods, apparatuses and computer program products for video encoding and decoding. In other embodiments, there is provided a method, an apparatus, a computer readable storage medium stored with code thereon for use by an apparatus, and a video encoder, for encoding a scalable bitstream, to provide indicating an encoding configuration, where only samples and syntax from intra coded pictures of base layer is used for coding the enhancement layer pictures. In other embodiments, there is provided an apparatus, a computer readable storage medium stored with code thereon for use by an apparatus, and a video decoder, for decoding a scalable bitstream, to receive indications of an encoding configuration, where only samples and syntax from intra coded pictures of base layer is used for coding the enhancement.


