Tile Group Identification in Video Streams
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing video coding technologies lack easily identifiable and parseable syntax elements for identifying tile groups and other picture segments, making it complex for Media-Aware Network Elements (MANEs) to efficiently handle and process video streams, especially in scenarios requiring selective forwarding of specific tile groups to multiple receivers.
Innovation Solution
Incorporating a flag or indicator in the parameter set of the coded video stream to indicate whether a tile group has a rectangular shape, along with syntax elements for corner identification and tile group ID, allowing for efficient identification and reconstruction, forwarding, or discarding of tile groups without requiring complex parsing of variable length codewords or parameter set context.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Productivity
If traditional video coding standards (H.264, H.265) are used without tile group identification syntax elements, then video compression efficiency is maintained, but Media-Aware Network Elements cannot efficiently identify and selectively forward specific tile groups, increasing processing complexity and time
Solution Approach 1:
The patent divides the video picture into tile groups and assigns unique identification syntax elements to each tile group header. This segmentation allows MANEs to independently identify and process specific tile groups without parsing the entire bitstream, resolving the contradiction between identification efficiency and parsing complexity
Solution Approach 2:
The patent introduces tile group identification syntax elements as an intermediary mechanism between the video encoder and MANEs. These syntax elements serve as a bridge that enables MANEs to efficiently identify tile groups without requiring complex parsing of variable length codewords or parameter set context, thus resolving the technical contradiction
2Quantity of substance
If variable length codewords and parameter set context are used for tile group identification, then compression ratio is improved, but parsing complexity increases significantly for network elements
Solution Approach 1:
The patent places tile group identification syntax elements in the tile group header, which is decoded early in the bitstream processing. This preliminary placement allows MANEs to identify tile groups before needing to parse complex variable length codewords or parameter set context, reducing identification difficulty while maintaining compression efficiency
Solution Approach 2:
The patent applies different syntax element strategies to different parts of the bitstream: simple identification syntax elements are placed in tile group headers for easy MANE processing, while more complex variable length codewords are used in other parts of the bitstream for compression. This local differentiation resolves the contradiction between compression ratio and identification difficulty
Data Source
AI summary
Methods and systems for decoding a video stream are provided, a method comprises receiving a coded video stream comprising a picture partitioned into a plurality of tile groups, each of the plurality of tile groups include at least one tile, the coded video stream further comprising a first indicator that indicates whether a tile group of the plurality of tile groups has a rectangular shape; identifying whether the tile group of the picture has the rectangular shape based on the first indicator; and reconstructing, forwarding, or discarding the tile group.


