Tile Group Identification in Video Streams

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing video coding technologies lack easily identifiable and parseable syntax elements for identifying tile groups and other picture segments, making it complex for Media-Aware Network Elements (MANEs) to efficiently handle and process video streams, especially in scenarios requiring selective forwarding of specific tile groups to multiple receivers.

Innovation Solution

Incorporating a flag or indicator in the parameter set of the coded video stream to indicate whether a tile group has a rectangular shape, along with syntax elements for corner identification and tile group ID, allowing for efficient identification and reconstruction, forwarding, or discarding of tile groups without requiring complex parsing of variable length codewords or parameter set context.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Productivity

If traditional video coding standards (H.264, H.265) are used without tile group identification syntax elements, then video compression efficiency is maintained, but Media-Aware Network Elements cannot efficiently identify and selectively forward specific tile groups, increasing processing complexity and time

Engineering Contradiction:
Improvetile group identification efficiencyVSAvoidparsing complexity for MANEs
Core Design Contradiction:
ProductivityVSDevice complexity

Solution Approach 1:

The patent divides the video picture into tile groups and assigns unique identification syntax elements to each tile group header. This segmentation allows MANEs to independently identify and process specific tile groups without parsing the entire bitstream, resolving the contradiction between identification efficiency and parsing complexity

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent introduces tile group identification syntax elements as an intermediary mechanism between the video encoder and MANEs. These syntax elements serve as a bridge that enables MANEs to efficiently identify tile groups without requiring complex parsing of variable length codewords or parameter set context, thus resolving the technical contradiction

Inventive Principle:
Principle #24Intermediary (Mediator)

2Quantity of substance

If variable length codewords and parameter set context are used for tile group identification, then compression ratio is improved, but parsing complexity increases significantly for network elements

Engineering Contradiction:
Improvebitstream compression ratioVSAvoidtile group identification difficulty
Core Design Contradiction:
Quantity of substanceVSDifficulty of detecting and measuring

Solution Approach 1:

The patent places tile group identification syntax elements in the tile group header, which is decoded early in the bitstream processing. This preliminary placement allows MANEs to identify tile groups before needing to parse complex variable length codewords or parameter set context, reducing identification difficulty while maintaining compression efficiency

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent applies different syntax element strategies to different parts of the bitstream: simple identification syntax elements are placed in tile group headers for easy MANE processing, while more complex variable length codewords are used in other parts of the bitstream for compression. This local differentiation resolves the contradiction between compression ratio and identification difficulty

Inventive Principle:
Principle #3Local quality

Data Source

PatentUS12113997B2Method for tile group identification
Publication Date: 2024.10.08 TENCENT AMERICA LLC
  • US12113997B2 patent drawing
  • US12113997B2 patent drawing
  • US12113997B2 patent drawing

AI summary

Methods and systems for decoding a video stream are provided, a method comprises receiving a coded video stream comprising a picture partitioned into a plurality of tile groups, each of the plurality of tile groups include at least one tile, the coded video stream further comprising a first indicator that indicates whether a tile group of the plurality of tile groups has a rectangular shape; identifying whether the tile group of the picture has the rectangular shape based on the first indicator; and reconstructing, forwarding, or discarding the tile group.