Video Coding Profile Tier Level Signaling for High Luma Rates

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current video coding standards face challenges in efficiently signaling profile and level information, which affects video decoding and compression capabilities, particularly in next-generation video coding standards like Versatile Video Coding (VVC).

Innovation Solution

The proposed techniques involve receiving and parsing profile tier level syntax to determine the appropriate decoding level, specifically using a value of 105 to indicate a maximum luma sample rate of 4812963840 samples per second, and performing video decoding accordingly, allowing for flexible and efficient video block structures, prediction techniques, and entropy coding.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If profile and level information is signaled using traditional methods in video coding standards, then compatibility with existing decoders is maintained, but signaling efficiency and support for next-generation video capabilities are limited

Engineering Contradiction:
Improvesupport for next-generation video capabilitiesVSAvoidsignaling efficiency
Core Design Contradiction:
Adaptability or versatilityVSLoss of information

Solution Approach 1:

The patent introduces a new level syntax element with specific values (e.g., level 105 indicating maximum luma sample rate of 4812963840 samples per second) to encode advanced video capabilities. This changes the parameter representation method to support higher resolution and frame rate video while maintaining backward compatibility through selective parsing based on the syntax element value.

Inventive Principle:
Principle #35Parameter changes

2Reliability

If a compliant bitstream format is used to ensure decoder compatibility, then reliable video decoding is achieved, but flexibility in supporting diverse video compression capabilities is reduced

Engineering Contradiction:
Improvevideo decoding reliabilityVSAvoidvideo compression capabilities
Core Design Contradiction:
ReliabilityVSAdaptability or versatility

Solution Approach 1:

The patent implements dynamic syntax element parsing where the decoder selectively processes profile_tier_level syntax based on the value of level syntax elements. When advanced level values are detected (e.g., level 105), the decoder activates corresponding advanced decoding capabilities. This dynamic approach allows the same bitstream format to reliably support both traditional and next-generation video capabilities.

Inventive Principle:
Principle #15Dynamics

3Loss of information

If detailed profile and level information is always signaled, then complete video capability information is provided, but bitstream complexity and processing overhead increase

Engineering Contradiction:
Improvevideo capability informationVSAvoidbitstream processing complexity
Core Design Contradiction:
Loss of informationVSDevice complexity

Solution Approach 1:

The patent extracts only the necessary profile_tier_level syntax elements based on the detected level value. Instead of always processing complete profile information, the decoder extracts and processes only the relevant capability information needed for the specific video stream, reducing processing complexity while maintaining complete capability awareness when needed.

Inventive Principle:
Principle #2Taking out (Extraction)

Data Source

PatentUS11792433B2Systems and methods for signaling profile and level information in video coding
Publication Date: 2023.10.17 SHARP KK
  • US11792433B2 patent drawing
  • US11792433B2 patent drawing
  • US11792433B2 patent drawing

AI summary

A method of decoding video data comprises: receiving profile tier level syntax; parsing a syntax element, from the profile tier level syntax, indicating a level to which an output layer set conforms, wherein a value of 105 indicates a level where a maximum luma sample rate of 4812963840 samples per second is supported; and performing video decoding based on the indicated level.