Motion Matrix Video Encoding Using Low-Rank Sparse Representation

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Traditional video encoding methods fail to fully leverage temporal and spatial correlations for efficient compression, relying on non-invertible quantization and limited sparsity in transform domains.

Innovation Solution

The use of a motion matrix with low rank and sparse representation, formed from reference frames, which allows for efficient encoding and decoding by exploiting high temporal and spatial correlations through compressive sensing and low-rank matrix completion techniques.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Quantity of substance

If traditional quantization is used for compression, then data rate is reduced, but reconstruction accuracy deteriorates due to non-invertible loss

Engineering Contradiction:
Improvedata rateVSAvoidreconstruction accuracy
Core Design Contradiction:
Quantity of substanceVSMeasurement precision

Solution Approach 1:

The patent applies preliminary action by performing motion compensation and predicting the current frame from reference frames before quantization. This preliminary reconstruction step preserves essential temporal and spatial information, allowing subsequent quantization to operate on already-compressed data rather than raw pixel values, thereby maintaining reconstruction accuracy while achieving compression.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent introduces an intermediary prediction residue as a mediator between the reference frames and the current frame. Instead of directly quantizing and transmitting full frame data, the system computes the difference (residue) between predicted and actual frames, then quantizes only this residue. This intermediary representation captures only the necessary new information, reducing data rate while preserving reconstruction fidelity.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Quantity of substance

If transform coding is applied for compression, then bit rate is reduced, but sparsity is limited and compression efficiency deteriorates

Engineering Contradiction:
Improvebit rateVSAvoidcompression efficiency
Core Design Contradiction:
Quantity of substanceVSProductivity

Solution Approach 1:

The patent applies segmentation by dividing the video sequence into independent frame predictions and residue components. Each frame is segmented into a predicted portion (from reference frames via motion compensation) and a residue portion (the difference). This segmentation allows transform coding to be applied selectively to the residue, which has higher sparsity, thereby improving compression efficiency while reducing bit rate.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent changes the parameter representation from direct pixel values to motion-compensated prediction residues. By transforming the data representation from spatial domain pixels to temporal-domain residues, the system exploits temporal correlations more effectively, increasing sparsity and improving compression efficiency without significantly increasing bit rate.

Inventive Principle:
Principle #35Parameter changes

3Device complexity

If simple motion compensation is used, then decoding complexity is reduced, but temporal and spatial correlations are not fully exploited and compression deteriorates

Engineering Contradiction:
Improvedecoding complexityVSAvoidcompression ratio
Core Design Contradiction:
Device complexityVSQuantity of substance

Solution Approach 1:

The patent applies dynamics by implementing motion compensation that adapts to varying temporal and spatial correlations in different video sequences and scenes. The motion estimation and compensation parameters are dynamically adjusted based on the content characteristics, allowing the system to fully exploit temporal and spatial correlations when present while maintaining manageable decoding complexity through standardized algorithms.

Inventive Principle:
Principle #15Dynamics

Data Source

PatentUS9894354B2Methods and apparatus for video encoding and decoding using motion matrix
Publication Date: 2018.02.13 INTERDIGITAL MADISON PATENT HLDG
  • US9894354B2 patent drawing
  • US9894354B2 patent drawing
  • US9894354B2 patent drawing

AI summary

Methods and apparatus are provided for video encoding and decoding using a motion matrix. An apparatus includes a video encoder for encoding a picture in a video sequence using a motion matrix. The motion matrix has a rank below a given threshold and a sparse representation with respect to a dictionary. The dictionary includes a set of atoms and basis vectors for representing the picture and for permitting the picture to be derived at a corresponding decoder using only the set. The dictionary formed from a set of reference pictures in the video sequence.