Video Processing Method for Dynamic Special Effect Animation Superimposition
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current video applications do not enhance the display effect of video pictures during playback, despite offering interactive features like barrages of text and emoticons.
Innovation Solution
A video processing method that extracts relevant words or phrases from audio data as tags, determines corresponding special effect animations, and superimposes them on the video picture during playback, improving the display effect and user engagement.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If traditional barrage interaction is used, then user interaction is enhanced, but the display effect of video picture itself cannot be improved
Solution Approach 1:
The patent segments the video processing into multiple independent modules: audio data processing module, tag extraction module, animation determination module, and superimposition display module. Each module handles a specific task, allowing the system to enhance video display effects through structured processing without compromising user interaction capabilities
Solution Approach 2:
The patent transitions from traditional 2D video display to enhanced display by adding animation layers in a new dimension. Special effect animations are generated and superimposed on the video picture, creating a multi-layered display effect that enriches the visual experience without interfering with the original video content or barrage interactions
2Loss of information
If audio data processing is performed for all video segments, then comprehensive tag extraction is achieved, but system resource consumption increases
Solution Approach 1:
The patent extracts only the essential and relevant information from audio data through tag extraction. By identifying and extracting key tags that represent important audio content, the system avoids processing all audio data comprehensively, thereby reducing computational resource consumption while still achieving effective video enhancement
Solution Approach 2:
The patent applies partial processing by selecting specific audio segments or keyframes for tag extraction rather than processing the entire audio stream. This partial action approach maintains adequate tag extraction quality while significantly reducing the computational burden and resource consumption
3Adaptability or versatility
If special effect animations are added to enhance video display, then video experience is enriched, but device complexity increases
Solution Approach 1:
The patent implements a universal animation determination module that can generate various types of special effect animations based on different tag types. This multi-functional module handles diverse animation requirements through a unified processing framework, enriching video experience without proportionally increasing system complexity
Solution Approach 2:
The patent performs preliminary processing by pre-defining tag types and their corresponding animation mappings. By establishing this framework in advance, the system reduces real-time processing complexity when generating animations, as the determination logic is already prepared and structured
Data Source
AI summary
A video processing method, an electronic device and a storage medium, which relates to the field of video recognition and understanding and deep learning, are disclosed. The method may include: during video play, for to-be-processed audio data, which has not been played, determined according to a predetermined policy, performing the following processing: extracting a word/phrase meeting a predetermined requirement from text content corresponding to the audio data, as a tag of the audio data; determining a special effect animation corresponding to the audio data according to the tag; and superimposing the special effect animation on a corresponding video picture for display when the audio data begins to be played.


