Video Transform Block Processing With Selective Secondary Transform
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
The increasing complexity in processing high spatial resolution, high frame rate video content requires efficient transform technologies to reduce computational complexity and memory requirements.
Innovation Solution
A method and apparatus for encoding and decoding video signals that selectively apply a non-separable secondary transform based on the size of the transform block, omitting it for blocks smaller than or equal to 4×4 and applying it to larger blocks, using transforms like RST, SOT, Givens rotation-based transforms, and permutations.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Manufacturing precision
If a non-separable secondary transform is applied to all transform blocks to improve video quality, then video quality is improved, but computational complexity increases
Solution Approach 1:
The patent applies the non-separable secondary transform selectively only to transform blocks larger than 4×4, while omitting it for 4×4 blocks. This local differentiation allows the system to maintain high video quality for larger blocks where the transform provides significant benefit, while avoiding unnecessary computational complexity for smaller blocks where the transform offers minimal improvement.
2Productivity
If transform processing is applied to high spatial resolution video content to improve compression efficiency, then compression efficiency is improved, but processing time increases
Solution Approach 1:
The patent implements partial action by applying the non-separable secondary transform only to transform blocks larger than 4×4, rather than uniformly to all blocks. This selective application provides sufficient compression efficiency improvement for high spatial resolution content while reducing overall processing time by excluding small blocks from the computationally intensive transform.
Data Source
AI summary
A video signal encoding method according to an embodiment of the present invention comprises checking a transform block including residual samples except a prediction sample from a picture of the video signal, generating transform coefficients through a transform for the residual samples of the transform block based on a size of the transform block, and performing quantization and entropy coding the transform coefficients, wherein generating the transform coefficients includes, applying a forward primary transform to each of a horizontal direction and vertical direction of the transform block including the residual samples, and not applying a forward non-separable secondary transform to the transform block to which the primary transform has been applied when the size of the transform block is smaller than or equal to 4×4, and applying the forward non-separable secondary transform to the transform block when the size of the transform block is greater than 4×4.


