Video Processing Method for Dynamic Special Effect Animation Superimposition

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current video applications do not enhance the display effect of video pictures during playback, despite offering interactive features like barrages of text and emoticons.

Innovation Solution

A video processing method that extracts relevant words or phrases from audio data as tags, determines corresponding special effect animations, and superimposes them on the video picture during playback, improving the display effect and user engagement.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If traditional barrage interaction is used, then user interaction is enhanced, but the display effect of video picture itself cannot be improved

Engineering Contradiction:
Improveuser interactionVSAvoiddisplay effect of video picture
Core Design Contradiction:
Ease of operationVSIllumination intensity

Solution Approach 1:

The patent segments the video processing into multiple independent modules: audio data processing module, tag extraction module, animation determination module, and superimposition display module. Each module handles a specific task, allowing the system to enhance video display effects through structured processing without compromising user interaction capabilities

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent transitions from traditional 2D video display to enhanced display by adding animation layers in a new dimension. Special effect animations are generated and superimposed on the video picture, creating a multi-layered display effect that enriches the visual experience without interfering with the original video content or barrage interactions

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

2Loss of information

If audio data processing is performed for all video segments, then comprehensive tag extraction is achieved, but system resource consumption increases

Engineering Contradiction:
Improvecomprehensive tag extractionVSAvoidsystem resource consumption
Core Design Contradiction:
Loss of informationVSUse of energy by moving object

Solution Approach 1:

The patent extracts only the essential and relevant information from audio data through tag extraction. By identifying and extracting key tags that represent important audio content, the system avoids processing all audio data comprehensively, thereby reducing computational resource consumption while still achieving effective video enhancement

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The patent applies partial processing by selecting specific audio segments or keyframes for tag extraction rather than processing the entire audio stream. This partial action approach maintains adequate tag extraction quality while significantly reducing the computational burden and resource consumption

Inventive Principle:
Principle #16Partial or excessive action

3Adaptability or versatility

If special effect animations are added to enhance video display, then video experience is enriched, but device complexity increases

Engineering Contradiction:
Improvevideo experienceVSAvoidprocessing system complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent implements a universal animation determination module that can generate various types of special effect animations based on different tag types. This multi-functional module handles diverse animation requirements through a unified processing framework, enriching video experience without proportionally increasing system complexity

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The patent performs preliminary processing by pre-defining tag types and their corresponding animation mappings. By establishing this framework in advance, the system reduces real-time processing complexity when generating animations, as the determination logic is already prepared and structured

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS11800042B2Video processing method, electronic device and storage medium thereof
Publication Date: 2023.10.24 BAIDU ONLINE NETWORK TECH (BEIJIBG) CO LTD
  • US11800042B2 patent drawing
  • US11800042B2 patent drawing
  • US11800042B2 patent drawing

AI summary

A video processing method, an electronic device and a storage medium, which relates to the field of video recognition and understanding and deep learning, are disclosed. The method may include: during video play, for to-be-processed audio data, which has not been played, determined according to a predetermined policy, performing the following processing: extracting a word/phrase meeting a predetermined requirement from text content corresponding to the audio data, as a tag of the audio data; determining a special effect animation corresponding to the audio data according to the tag; and superimposing the special effect animation on a corresponding video picture for display when the audio data begins to be played.