Video Key Identifier Recognition via Frame Difference Masking

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current image recognition methods for video identifiers are weak in fault tolerance and perform poorly in scenarios with low resolution and definition, failing to accurately recognize key identifiers.

Innovation Solution

A method and apparatus for recognizing key identifiers in videos by extracting key frames, generating masks using frame differences, and determining key identifier area images to improve recognition accuracy and fault tolerance, involving modules for extraction, generation, and recognition using deep learning algorithms.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If traditional image recognition methods are used for video identifiers, then the recognition process is simple, but the recognition accuracy is low and fault tolerance is weak

Engineering Contradiction:
Improverecognition accuracyVSAvoidrecognition process complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The patent segments the video processing into distinct stages: key frame extraction, mask generation through frame differencing, candidate region identification, and final recognition. This segmentation allows each module to be optimized independently, improving overall recognition accuracy while maintaining manageable system complexity

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent performs preliminary actions by extracting key frames and generating masks before the actual recognition process. These preprocessing steps eliminate redundant information and highlight potential identifier locations, thereby improving recognition accuracy while reducing the complexity of the subsequent recognition task

Inventive Principle:
Principle #10Preliminary action

2Measurement precision

If full video frames are processed for identifier recognition, then comprehensive analysis is achieved, but data processing requirements and time consumption increase

Engineering Contradiction:
Improveidentifier recognition accuracyVSAvoidrecognition speed
Core Design Contradiction:
Measurement precisionVSProductivity

Solution Approach 1:

The patent extracts only the essential components needed for identifier recognition: key frames containing potential identifiers and masks highlighting difference regions. By extracting and processing only these critical elements rather than entire video frames, the system achieves accurate identifier recognition while significantly reducing data processing requirements and improving recognition speed

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The patent applies partial action by processing only key frames and their difference masks rather than all video frames. This selective processing approach provides sufficient information for accurate identifier recognition while minimizing computational overhead, thereby improving recognition speed without sacrificing accuracy

Inventive Principle:
Principle #16Partial or excessive action

3Reliability

If traditional recognition methods are used, then processing is faster, but fault tolerance and performance in low-resolution scenarios are poor

Engineering Contradiction:
Improvefault toleranceVSAvoidrecognition system complexity
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

The patent prepares masks through frame differencing before the recognition process, which cushions against variations in video quality, resolution, and formatting. This preliminary preparation creates a standardized input format that improves fault tolerance and reliability across different video scenarios, including low-resolution cases, while keeping the added complexity manageable

Inventive Principle:
Principle #11Beforehand cushioning (Prior cushioning)

Data Source

PatentUS11748986B2Method and apparatus for recognizing key identifier in video, device and storage medium
Publication Date: 2023.09.05 BEIJING BAIDU NETCOM SCI & TECH CO LTD
  • US11748986B2 patent drawing
  • US11748986B2 patent drawing
  • US11748986B2 patent drawing

AI summary

A method and an apparatus for recognizing a key identifier in a video, a device and a storage medium are disclosed. The method includes: extracting a plurality of key frames from the video; generating a mask of the key identifier by using a difference between the plurality of key frames; determining, in video frames of the video, a key identifier area image by using the mask; and recognizing the key identifier area image to obtain a key identifier category included in the video.