VVC Planar Intra Prediction with Adaptive Reference Samples

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

The existing video coding standards, such as HEVC, face challenges in achieving efficient compression of video data, particularly in capturing complex edge directions and textures, leading to suboptimal prediction modes and increased computational complexity.

Innovation Solution

The VVC standard introduces extended angular intra prediction modes, including Wide-Angle Intra Prediction (WAIP) and planar modes, along with advanced filtering and interpolation techniques, to enhance prediction accuracy and reduce complexity through methods like MPM lists, TIMD, and PDPC, while incorporating new transforms like LFNST and MTS.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If extended angular intra prediction modes (WAIP, planar modes) and advanced filtering techniques are introduced, then prediction accuracy is improved, but device complexity increases

Engineering Contradiction:
Improveprediction accuracyVSAvoiddevice complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The patent segments the prediction process into multiple independent components: angular prediction modes (WAIP, planar, DC), filtering operations (TMDF, PDPC), and transformation techniques (LFNST, MTS). Each component can be processed separately and combined, allowing the system to achieve high prediction accuracy through modular operations rather than a single complex algorithm.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent implements dynamic adaptation by allowing the encoder and decoder to select from multiple prediction modes (angular, planar, DC) based on the specific characteristics of each block. The system dynamically switches between different filtering techniques (TMDF, PDPC) and transformation methods (LFNST, MTS) to optimize performance for different video content types, thereby achieving high accuracy without requiring all operations to be performed simultaneously.

Inventive Principle:
Principle #15Dynamics

2Productivity

If multiple prediction modes and transforms are implemented, then compression efficiency is improved, but computational complexity increases

Engineering Contradiction:
Improvecompression efficiencyVSAvoidcomputational complexity
Core Design Contradiction:
ProductivityVSDevice complexity

Solution Approach 1:

The patent applies partial action by implementing multiple prediction modes and transforms but allowing the system to use only the necessary ones for each specific block. Rather than always performing all available operations, the encoder selects and applies only the relevant modes (angular, planar, DC) and transforms (LFNST, MTS) needed for that particular block, reducing overall computational complexity while maintaining high compression efficiency when needed.

Inventive Principle:
Principle #16Partial or excessive action

Solution Approach 2:

The patent utilizes parameter changes by varying the prediction mode indices, filter coefficients, and transformation parameters based on block characteristics. The system adjusts these parameters dynamically to optimize compression efficiency for different content types while controlling computational complexity through parameter adaptation rather than fixed complex operations.

Inventive Principle:
Principle #35Parameter changes

3Manufacturing precision

If advanced filtering (TMDF, PDPC) and transforms (LFNST, MTS) are applied, then video quality is enhanced, but processing time increases

Engineering Contradiction:
Improvevideo qualityVSAvoidprocessing time
Core Design Contradiction:
Manufacturing precisionVSLoss of time

Solution Approach 1:

The patent implements preliminary action by pre-defining multiple prediction modes, filter coefficients, and transformation parameters that can be quickly selected and applied. The encoder and decoder have pre-established lists of available modes and parameters, allowing them to rapidly switch between different operations without performing complex calculations from scratch, thereby enhancing video quality while minimizing processing time through prepared options.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent applies dynamics by enabling the system to adaptively select the most appropriate filtering and transformation operations based on block characteristics. Rather than always applying the most complex operations, the system dynamically chooses the optimal balance between quality enhancement and processing time for each specific block, using modes like TMDF and PDPC only when necessary to achieve significant quality improvements.

Inventive Principle:
Principle #15Dynamics

Data Source

PatentUS12457325B2On planar intra prediction mode
Publication Date: 2025.10.28 ALIBABA (CHINA) CO LTD
  • US12457325B2 patent drawing
  • US12457325B2 patent drawing
  • US12457325B2 patent drawing

AI summary

An input video or video stream may be obtained or received. The input video or video stream may include a plurality of video frames, and each frame may be divided into a plurality of blocks. A current block of the plurality of blocks may be predicted using a planar mode. Depending on which planar mode is used, different reference samples may be used for predicting a current sample in the current block.