Text Description Signaling in Video Coding Bitstreams
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing video coding standards like ITU-T H.264, ITU-T H.265, and ITU-T H.266 lack efficient methods for signaling text description information, which is crucial for enhancing video coding capabilities beyond their limitations.
Innovation Solution
Incorporating techniques for signaling a text description information message with specific syntax elements to indicate the role and specify a string value, enabling effective communication of text description information in video coding processes.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Loss of information
If existing video coding standards (ITU-T H.264, H.265, H.266) are used, then video compression is achieved, but text description information cannot be efficiently signaled
Solution Approach 1:
The text description information is nested within existing video coding structures by embedding it in SEI (Supplemental Enhancement Information) messages or VUI (Video Usability Information) parameters. This allows text information to be carried within the existing bitstream framework without requiring a separate signaling channel, thus preventing information loss while avoiding excessive complexity increase.
Solution Approach 2:
The invention creates a universal text description signaling mechanism that can handle multiple types of text information (captions, subtitles, metadata, etc.) through a unified syntax structure. This multi-functional approach allows a single signaling mechanism to serve various text description needs, reducing the overall system complexity compared to having separate mechanisms for each text type.
2Adaptability or versatility
If text description information is added to video coding, then video coding capabilities are enhanced, but data requirements increase
Solution Approach 1:
The invention uses efficient parameter-based encoding for text description information, where text data is represented compactly using optimized syntax elements and entropy coding techniques. By changing the parameter representation method to be more compact, the system enhances video coding capabilities with text information while minimizing the increase in data requirements.
Solution Approach 2:
The text description information is signaled selectively based on what is necessary for the specific video content and application requirements. The signaling mechanism allows for partial inclusion of text information only when needed, rather than always transmitting full text descriptions, thus enhancing capabilities where required while keeping data requirements minimal when text information is not necessary.
Data Source
AI summary
A device may be configured to decode video data based on information included in a text description information message. In one example, a text description information message includes a syntax element indicating a role of the text description information message. In one example, a text description information message includes a syntax element specifying a text description information string having a value which is interpreted as indicated by a role.


