Video Encoding Tile Merging for 360-Degree ROI Quality
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional video encoding and decoding techniques are inefficient for 360° video due to limitations in setting a region of interest (ROI) and maintaining image quality across tile boundaries, especially when dealing with non-uniform tile structures.
Innovation Solution
A method and apparatus for video encoding and decoding that allow for flexible tile structures by merging tiles to form merge tiles, enabling improved encoding efficiency by eliminating restrictions on encoding dependencies between tiles, and allowing for higher image quality within the ROI.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Productivity
If a picture is split into multiple tiles for independent encoding, then encoding parallelism and processing efficiency are improved, but image quality degradation at tile boundaries occurs and encoding dependencies between tiles must be restricted
Solution Approach 1:
The patent merges multiple tiles into a single tile structure when they share encoding dependencies, eliminating tile boundary artifacts and allowing unrestricted inter-tile prediction. This combining approach resolves the contradiction by maintaining encoding parallelism benefits while removing quality degradation at artificial boundaries.
2Adaptability or versatility
If conventional tile structures are used with encoding dependency restrictions, then independent tile processing is enabled, but flexibility in setting regions of interest (ROI) is limited
Solution Approach 1:
The patent introduces dynamic tile merging where the tile structure adapts based on ROI requirements and encoding dependencies. Tiles can be merged or kept separate depending on the specific encoding scenario, providing flexibility for ROI setting while managing structural complexity through conditional merging rules.
Data Source
AI summary
The present disclosure relates to video encoding or decoding for splitting a picture into a plurality of tiles in order to encode video efficiently. In one aspect of the present disclosure, a video encoding method for encoding a picture split into a plurality of tiles includes encoding first information indicating whether to merge some of the plurality of tiles; when the first information is encoded to indicate tile merging, generating one or more merge tiles by merging some of the plurality of tiles, each of the merge tiles being defined as one tile; encoding second information indicating tiles merged into each of the merge tiles among the plurality of tiles; and encoding each of the merge tiles as one tile without restriction on encoding dependencies between the tiles merged into each of the merge tiles.


