Video Coding Merge Syntax Signaling for Efficient Inter Prediction
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
The increasing demand for high-resolution, high-quality images and videos, particularly in virtual reality and augmented reality, leads to higher data transmission and storage costs due to the increased amount of information required, necessitating a more efficient compression technique.
Innovation Solution
A method and apparatus for enhancing image coding efficiency by optimizing inter prediction through combined inter-picture merge and intra-picture prediction (CIIP), which includes determining a prediction mode, deriving motion information, and generating prediction samples, while efficiently signaling information on merge modes.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Manufacturing precision
If high resolution and high quality image/video data are transmitted, then image/video quality is improved, but transmission and storage costs increase due to increased amount of information
Solution Approach 1:
The patent changes the parameter of data representation by introducing merge mode that reuses motion information from neighboring blocks, reducing the number of bits needed to represent motion vectors while maintaining prediction accuracy for high-quality image/video compression
2Ease of operation
If conventional inter prediction is used, then prediction functionality is provided, but unnecessary signaling increases device complexity and reduces coding efficiency
Solution Approach 1:
The patent extracts and removes unnecessary signaling elements from the inter prediction process by introducing a flag-based mechanism that selectively signals merge mode only when needed, eliminating redundant syntax elements and simplifying the decoding process
Solution Approach 2:
Instead of always signaling complete motion information, the patent inverts the approach by assuming motion information is available and only signaling when merge mode is active, reversing the default behavior to reduce overhead
Data Source
AI summary
A decoding method performed by a decoding device according to the present document comprises: a step for determining the prediction mode of a current block on the basis of information on a prediction mode acquired from a bitstream; a step for deriving movement information on the current block on the basis of the prediction mode; a step for generating prediction samples of the current block on the basis of the movement information; and a step for generating reconstructed samples on the basis of the prediction samples, wherein the step for determining the prediction mode may include a step for acquiring a regular merge flag from the bitstream on the basis of a combined inter-picture merge and intra-picture prediction (CIIP) available flag that indicates whether or not a CIIP is available.


