Video Search System for Frame-Based Content Identification

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current image identifying technologies only support photo-based searches and lack a solution for video-based searches, limiting the ability to identify and retrieve information from video content.

Innovation Solution

A method and device for video search that involves receiving an event triggered in a video playback page, acquiring a current video image frame, identifying and displaying a target within the frame, and providing a recommendation result on a search result page, enabling the classification and identification of persons and items within video frames.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If image identifying technology is used for photo-based searches, then identification accuracy is improved, but video-based search capability is lost

Engineering Contradiction:
Improveidentification accuracyVSAvoidvideo-based search capability
Core Design Contradiction:
Measurement precisionVSAdaptability or versatility

Solution Approach 1:

The existing image identifying technology is extended to handle both photo-based and video-based searches. The system maintains the accurate identification capabilities for photos while adding video frame extraction and analysis functions, allowing the same identification engine to work with both static images and video content through a unified interface.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The video search functionality is implemented by segmenting video content into individual frames that can be processed by the existing image identification system. The video is divided into discrete searchable units (frames), allowing the photo-based identification technology to be applied to video content without requiring complete redesign of the identification engine.

Inventive Principle:
Principle #1Segmentation

2Adaptability or versatility

If video-based search is implemented, then search versatility is improved, but system complexity increases

Engineering Contradiction:
Improvesearch versatilityVSAvoidsystem complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

An intermediary layer is introduced between the video content and the image identification system. This intermediary handles video-specific operations such as frame extraction, timestamp management, and video metadata processing, while presenting standardized image data to the existing identification engine. This mediator approach adds video search versatility without significantly increasing the complexity of the core identification system.

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

Video frames are pre-processed and prepared before being submitted to the identification system. Frames are extracted, resized, and formatted in advance, so that when search queries are executed, the identification engine receives ready-to-process images rather than having to handle raw video data. This preliminary preparation reduces the computational complexity during actual search operations.

Inventive Principle:
Principle #10Preliminary action

3Loss of information

If video frames are analyzed for search, then information retrieval capability is improved, but processing time increases

Engineering Contradiction:
Improveinformation retrieval capabilityVSAvoidprocessing time
Core Design Contradiction:
Loss of informationVSLoss of time

Solution Approach 1:

Instead of analyzing every single video frame, the system selectively processes only relevant frames based on search criteria, user interactions, or frame importance metrics. This partial action approach maintains comprehensive information retrieval capability by focusing computational resources on key frames that are most likely to contain the searched content, thereby reducing overall processing time while preserving search effectiveness.

Inventive Principle:
Principle #16Partial or excessive action

Data Source

PatentUS11630861B2Method and apparatus for video searching, terminal and storage medium
Publication Date: 2023.04.18 DOUYIN VISION CO LTD
  • US11630861B2 patent drawing
  • US11630861B2 patent drawing
  • US11630861B2 patent drawing

AI summary

Provided are a method and device for video search, a terminal and a storage medium. The method includes: receiving a first event generated by triggering a first control in a video playback page; acquiring, in response to the first event, a current video image frame played in the video playback page when the first event is triggered; acquiring a first to-be-searched target positioned by a second control in the current video image frame and a first display position of the first to-be-searched target in the current video image frame, and displaying the second control on the first display position; and acquiring a first recommendation result corresponding to the first to-be-searched target, and displaying the first recommendation result in a search result page.