Video Object Recognition and Synthesis for Efficient Processing
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current video processing methods require manual extraction and editing of video images, which is time-consuming and lacks efficiency in isolating specific objects such as faces or bodies for generating new videos.
Innovation Solution
A method that automatically recognizes faces or body areas in source videos, assigns object identifiers, and synthesizes target videos based on these identifiers, allowing for efficient extraction and combination of video images featuring specific objects.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Productivity
If manual extraction and editing of video images is used, then users can process videos with video editing programs, but the process is time-consuming and lacks efficiency
Solution Approach 1:
The system performs automatic object recognition, identification, and video generation without requiring manual user intervention for each step. The processor automatically identifies objects in video frames, tracks them across frames, and synthesizes target videos based on user selections, eliminating the need for manual extraction and editing operations
Solution Approach 2:
The patent replaces manual mechanical editing operations with automated computer vision and image processing algorithms. Object recognition algorithms automatically identify and extract objects from video frames, replacing the manual process of selecting and extracting video images with a video editing program
2Measurement precision
If automatic object recognition is implemented, then video processing accuracy is improved, but system complexity increases
Solution Approach 1:
The system segments the video processing task into distinct functional modules: object recognition module that identifies objects in video frames, object identification module that assigns unique identifiers to detected objects, object tracking module that follows objects across frames, and video synthesis module that generates target videos. This modular segmentation manages system complexity while maintaining high recognition accuracy
Solution Approach 2:
The patent introduces an object identification layer that acts as an intermediary between object detection and video synthesis. This layer assigns unique object identifiers to detected objects and maintains tracking information, serving as a mediator that connects the recognition system with the video generation system while managing complexity
Data Source
AI summary
The present disclosure relates to a video processing method, and relates to the technical field of multimedia. The method can include: acquiring a source image from a source video, and obtaining a first object by recognizing the source image; adding an object identifier of the first object to the source image; and generating a target video based on the source image with a same object identifier.


