Video Encoding Character Key Point Extraction
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current video encoding standards, particularly those using a hybrid framework of "prediction+transformation", are not optimized for video applications where characters are the main subject, leading to inefficient compression and increased storage and transmission costs.
Innovation Solution
The proposed method involves identifying character elements in a video sequence, extracting key point data, fitting this data with a standard character model to obtain fitting parameters, and encoding these parameters to form a structured information code stream, which can be decoded to reconstruct the video.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Productivity
If a hybrid encoding framework of 'prediction+transformation' is used, then general video encoding is achieved, but encoding efficiency for character-focused videos is insufficient
Solution Approach 1:
The video sequence is segmented into character elements and background elements. The encoder identifies character elements, extracts key point data from them, and processes these separately from the background, allowing optimized encoding for character-focused content while maintaining general video encoding capabilities.
Solution Approach 2:
Different encoding strategies are applied to different parts of the video. Character key point data is extracted and encoded with higher precision using fitting parameters, while background areas receive standard encoding treatment. This local differentiation improves overall encoding efficiency for character-focused videos.
2Reliability
If full video data is encoded and transmitted, then complete video quality is maintained, but storage and transmission costs increase
Solution Approach 1:
The encoder extracts only the essential character key point data from the full video sequence. Instead of encoding and transmitting complete video frames, the system identifies character elements, extracts their key point coordinates, and encodes only this extracted data along with fitting parameters, dramatically reducing data transmission amount while preserving character action information.
Solution Approach 2:
The system creates a simplified representation (copy) of the character data using key point coordinates and fitting parameters rather than transmitting the original full-resolution video data. This copied representation maintains the essential character information needed for reconstruction while occupying minimal storage and bandwidth.
Data Source
AI summary
The present disclosure provides video encoding and decoding methods and related products, where the methods are included in a combined processing apparatus, and the combined processing apparatus further includes an interface apparatus and other processing apparatus. A computing apparatus interacts with other processing apparatus to jointly complete a computing operation specified by a user. The combined processing apparatus further includes a storage apparatus. The storage apparatus is connected to the apparatus and other processing apparatus, respectively. The storage apparatus is configured to store data of the apparatus and other processing apparatus. A technical scheme of the present disclosure may significantly improve video compression efficiency.


