Scalable Video Encoding Inter-Layer Prediction Reference Selection

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current video encoding and decoding technologies face challenges in providing scalable video quality, particularly in efficiently performing inter-layer prediction in multilayer structures to accommodate varying video qualities and environments.

Innovation Solution

A method and apparatus for encoding and decoding scalable video in a multilayer structure, where an inter-layer reference picture is derived from a decoded picture in a reference layer, used to predict a current block, and reference information is transmitted to indicate available pictures for inter-layer prediction, allowing for effective inter-layer prediction and selective processing.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If inter-layer prediction is performed using all pictures in the reference layer, then prediction accuracy is improved, but processing complexity and memory requirements increase

Engineering Contradiction:
Improveprediction accuracyVSAvoidprocessing complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The patent applies local quality by selectively marking specific pictures in the reference layer as available for inter-layer prediction using reference information, rather than uniformly processing all pictures. This allows the system to focus computational resources on only those reference pictures that are actually useful for prediction, improving efficiency while maintaining prediction accuracy.

Inventive Principle:
Principle #3Local quality

Solution Approach 2:

The patent implements partial action by performing inter-layer prediction only on pictures that are marked as available through reference information. Instead of processing all pictures in the reference layer, the system selectively applies prediction operations only where needed, reducing overall processing complexity while maintaining effective prediction performance.

Inventive Principle:
Principle #16Partial or excessive action

2Productivity

If reference information is transmitted to indicate available pictures for inter-layer prediction, then encoding efficiency is improved, but data transmission overhead increases

Engineering Contradiction:
Improveencoding efficiencyVSAvoiddata transmission overhead
Core Design Contradiction:
ProductivityVSQuantity of substance

Solution Approach 1:

The patent extracts only the essential reference information needed to identify available pictures for inter-layer prediction, rather than transmitting complete picture data or exhaustive reference lists. This extraction approach provides sufficient information to enable efficient encoding while minimizing the overhead of transmitted data.

Inventive Principle:
Principle #2Taking out (Extraction)

3Speed

If selective processing is performed on pictures in the reference layer, then processing speed is improved, but prediction accuracy may deteriorate

Engineering Contradiction:
Improveprocessing speedVSAvoidprediction accuracy
Core Design Contradiction:
SpeedVSMeasurement precision

Solution Approach 1:

The patent applies preliminary action by pre-marking pictures in the reference layer with reference information that indicates their availability for inter-layer prediction. This preliminary classification allows the decoding process to quickly identify and process only the relevant reference pictures, improving processing speed while ensuring that prediction accuracy is maintained by selecting appropriate reference pictures in advance.

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS10873750B2Method for encoding video, method for decoding video, and apparatus using same
Publication Date: 2020.12.22 LG ELECTRONICS INC
  • US10873750B2 patent drawing
  • US10873750B2 patent drawing
  • US10873750B2 patent drawing

AI summary

The present invention relates to a method for coding a scalable video in a multilayer structure, and a method for encoding a video according to the present invention comprises the steps of: decoding and saving a picture of a reference layer; inducing an interlayer reference picture which is referenced for predicting a current block of a current layer; producing a reference picture list including the interlayer reference picture and a reference picture of the current layer; conducting a prediction on the current block of the current layer with the reference picture list to induce a prediction sample with respect to the current block; inducing a recovery sample with respect to the current block based on the prediction sample and the prediction block with respect to the current block; and transmitting reference information for indicating a picture that can be used for interlayer prediction from the pictures of the reference layer.