Upsampling Prediction Mode Information for Scalable Video Coding
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing video coding technologies face challenges in enabling inter-layer motion prediction when the spatial resolution of an enhancement layer differs from that of the base layer, as motion information from the base layer may not be accessible without modifying the base layer system design or using different hardware/software, which limits video compression efficiency.
Innovation Solution
The approach involves upsampling prediction mode information from the base layer to facilitate inter-layer motion prediction for the enhancement layer, allowing the use of upsampled prediction mode information to determine predicted mode information for enhancement layer blocks, thereby enabling efficient compression without requiring changes to the coding unit or low-level system design.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Productivity
If inter-layer motion prediction is enabled between base layer and enhancement layer with different spatial resolutions, then video compression efficiency is improved, but the base layer system design must be modified or different hardware/software is required
Solution Approach 1:
The patent introduces an intermediary upsampling process that converts prediction mode information from base layer resolution to enhancement layer resolution. This mediator enables compatibility between layers with different spatial resolutions without requiring modifications to the base layer system design or specialized hardware, thus resolving the technical contradiction by maintaining system simplicity while improving compression efficiency
Solution Approach 2:
The patent changes the resolution parameter of prediction mode information through upsampling operations. By transforming the spatial resolution of prediction mode data to match the enhancement layer resolution, the system enables inter-layer motion prediction across different resolutions without altering the base layer system architecture, thereby improving compression efficiency while avoiding increased device complexity
2Adaptability or versatility
If prediction mode information is upsampled from base layer to enhancement layer, then inter-layer motion prediction becomes possible, but additional processing steps are required
Solution Approach 1:
The patent performs upsampling of prediction mode information as a preliminary action before inter-layer motion prediction. By pre-processing the prediction mode data to match the enhancement layer resolution, the system enables subsequent prediction operations without requiring complex real-time adaptations, thus improving versatility while keeping processing complexity manageable through structured pre-computation
Data Source
Figure 1
Figure 2
Figure 3
AI summary
In one embodiment, an apparatus configured to code video data includes a processor and a memory unit. The memory unit stores video data associated with a first layer having a first spatial resolution and a second layer having a second spatial resolution. The video data associated with the first layer includes at least a first layer block and first layer prediction mode information associated with the first layer block, and the first layer block includes a plurality of sub-blocks where each sub-block is associated with respective prediction mode data of the first layer prediction mode information. The processor derives the predication mode data associated with one of the plurality of sub-blocks based at least on a selection rule, upsamples the derived prediction mode data and the first layer block, and associates the upsampled prediction mode data with each upsampled sub-block of the upsampled first layer block.