Scalable Video Coding Inter-Layer Prediction Restriction

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Scalable video coding methods, such as single-loop decoding, face inefficiencies due to the need to transmit entire inter base layer pictures for enhancement layer decoding, even if only a small portion of data is required, and struggle with standard scalability cases where different coding standards are used for base and enhancement layers, like H.264/AVC and HEVC.

Innovation Solution

A sequence level indication is provided to restrict inter-layer prediction in scalable video coding to only intra-coded pictures or random access point pictures in the base layer, allowing for efficient encoding and decoding by discarding unnecessary pictures and simplifying motion vector prediction processes.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If entire inter base layer pictures are transmitted for enhancement layer decoding, then decoding accuracy is improved, but data transmission volume and processing complexity increase

Engineering Contradiction:
Improvedecoding accuracyVSAvoiddata transmission volume
Core Design Contradiction:
Measurement precisionVSQuantity of substance

Solution Approach 1:

The patent extracts and transmits only the necessary portions of base layer pictures (intra-coded pictures or RAP pictures) that are required for enhancement layer decoding, rather than transmitting entire inter base layer pictures. This selective extraction reduces data transmission volume while maintaining the decoding accuracy needed for enhancement layer reconstruction.

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The patent segments the base layer pictures into different types (intra-coded pictures and RAP pictures versus other inter-coded pictures) and selectively processes only the relevant segments for enhancement layer decoding. This segmentation allows the system to discard unnecessary picture data while preserving the essential components needed for accurate enhancement layer reconstruction.

Inventive Principle:
Principle #1Segmentation

2Measurement precision

If entire inter base layer pictures are transmitted for enhancement layer decoding, then decoding accuracy is improved, but processing complexity increases

Engineering Contradiction:
Improvedecoding accuracyVSAvoidprocessing complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The patent extracts only the necessary base layer picture types (intra-coded and RAP pictures) required for enhancement layer decoding, eliminating the need to process entire inter base layer pictures. This extraction approach reduces processing complexity by removing unnecessary decoding and motion compensation operations for irrelevant picture data.

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The patent divides base layer pictures into different categories and selectively processes only the relevant segments (intra-coded and RAP pictures) for enhancement layer decoding. This segmentation simplifies the processing pipeline by avoiding complex motion compensation and prediction operations on unnecessary inter-coded pictures.

Inventive Principle:
Principle #1Segmentation

3Adaptability or versatility

If syntax elements from all base layer pictures are decoded, then compatibility with different coding standards is improved, but memory requirements increase

Engineering Contradiction:
Improvecompatibility with different coding standardsVSAvoidmemory requirements
Core Design Contradiction:
Adaptability or versatilityVSQuantity of substance

Solution Approach 1:

The patent extracts and retains only the syntax elements from intra-coded base layer pictures and RAP pictures that are necessary for enhancement layer decoding and cross-standard compatibility. This selective extraction reduces memory requirements by discarding syntax elements from other inter-coded pictures that are not needed for enhancement layer reconstruction or standard compatibility.

Inventive Principle:
Principle #2Taking out (Extraction)

Data Source

PatentUS10771805B2Apparatus, a method and a computer program for video coding and decoding
Publication Date: 2020.09.08 NOKIA TECHNOLOGIES OY
  • US10771805B2 patent drawing
  • US10771805B2 patent drawing
  • US10771805B2 patent drawing

AI summary

There are disclosed various methods, apparatuses and computer program products for video encoding and decoding. In other embodiments, there is provided a method, an apparatus, a computer readable storage medium stored with code thereon for use by an apparatus, and a video encoder, for encoding a scalable bitstream, to provide indicating an encoding configuration, where only samples and syntax from intra coded pictures of base layer is used for coding the enhancement layer pictures. In other embodiments, there is provided an apparatus, a computer readable storage medium stored with code thereon for use by an apparatus, and a video decoder, for decoding a scalable bitstream, to receive indications of an encoding configuration, where only samples and syntax from intra coded pictures of base layer is used for coding the enhancement.