Video Encoding Character Key Point Extraction

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current video encoding standards, particularly those using a hybrid framework of "prediction+transformation", are not optimized for video applications where characters are the main subject, leading to inefficient compression and increased storage and transmission costs.

Innovation Solution

The proposed method involves identifying character elements in a video sequence, extracting key point data, fitting this data with a standard character model to obtain fitting parameters, and encoding these parameters to form a structured information code stream, which can be decoded to reconstruct the video.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Productivity

If a hybrid encoding framework of 'prediction+transformation' is used, then general video encoding is achieved, but encoding efficiency for character-focused videos is insufficient

Engineering Contradiction:
Improveencoding efficiencyVSAvoidadaptability to character-focused videos
Core Design Contradiction:
ProductivityVSAdaptability or versatility

Solution Approach 1:

The video sequence is segmented into character elements and background elements. The encoder identifies character elements, extracts key point data from them, and processes these separately from the background, allowing optimized encoding for character-focused content while maintaining general video encoding capabilities.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

Different encoding strategies are applied to different parts of the video. Character key point data is extracted and encoded with higher precision using fitting parameters, while background areas receive standard encoding treatment. This local differentiation improves overall encoding efficiency for character-focused videos.

Inventive Principle:
Principle #3Local quality

2Reliability

If full video data is encoded and transmitted, then complete video quality is maintained, but storage and transmission costs increase

Engineering Contradiction:
Improvevideo qualityVSAvoiddata transmission amount
Core Design Contradiction:
ReliabilityVSQuantity of substance

Solution Approach 1:

The encoder extracts only the essential character key point data from the full video sequence. Instead of encoding and transmitting complete video frames, the system identifies character elements, extracts their key point coordinates, and encodes only this extracted data along with fitting parameters, dramatically reducing data transmission amount while preserving character action information.

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The system creates a simplified representation (copy) of the character data using key point coordinates and fitting parameters rather than transmitting the original full-resolution video data. This copied representation maintains the essential character information needed for reconstruction while occupying minimal storage and bandwidth.

Inventive Principle:
Principle #26Copying

Data Source

PatentUS20250080745A1Method for video encoding, method for video decoding, and related product
Publication Date: 2025.03.06 CAMBRICON TECH CO LTD
  • US20250080745A1 patent drawing
  • US20250080745A1 patent drawing
  • US20250080745A1 patent drawing

AI summary

The present disclosure provides video encoding and decoding methods and related products, where the methods are included in a combined processing apparatus, and the combined processing apparatus further includes an interface apparatus and other processing apparatus. A computing apparatus interacts with other processing apparatus to jointly complete a computing operation specified by a user. The combined processing apparatus further includes a storage apparatus. The storage apparatus is connected to the apparatus and other processing apparatus, respectively. The storage apparatus is configured to store data of the apparatus and other processing apparatus. A technical scheme of the present disclosure may significantly improve video compression efficiency.