Video Player Integrating Offline and Online Object Recognition
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current video-based information display methods are limited in their ability to facilitate seamless image recognition and search within videos, requiring users to capture images and upload them for recognition, leading to a complex operation and poor user experience.
Innovation Solution
A video-based information display method that displays resource information corresponding to target objects in a video, combining offline and online recognition, allowing users to view pre-recognized information during playback and enabling quick online recognition and search upon user interaction, such as by clicking or taking a screenshot.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If users capture images from video and upload to image recognition application, then image recognition and search function is provided, but operation process becomes complicated and user experience deteriorates
Solution Approach 1:
The patent merges the video player and image recognition application into a single integrated system. The video playing page directly displays resource information related to target objects in the video frames, eliminating the need for separate image capture and upload operations. This integration allows users to obtain recognition results while watching the video, significantly simplifying the operation process.
Solution Approach 2:
The system performs preliminary recognition on video frames during video playback. By pre-processing and identifying target objects in video frames in advance, the system prepares recognition results before user interaction is needed, enabling immediate display of resource information without requiring users to manually capture and upload images.
2Adaptability or versatility
If comprehensive image recognition and search function is provided for video, then functionality is enriched, but operation complexity increases
Solution Approach 1:
The patent combines multiple functions (video playback, image recognition, and search) into a single integrated video playing page. The system automatically performs recognition on video frames and displays resource information directly on the playing page, merging what would otherwise be separate operations into one seamless process.
Solution Approach 2:
The system automatically performs image recognition and search operations without requiring user intervention for image capture or upload. The video playing page itself provides the recognition function by automatically processing video frames and displaying results, making the system self-serving and reducing operational complexity.
3Adaptability or versatility
If image capture and upload process is used for video recognition, then recognition and search can be performed, but user experience becomes poor
Solution Approach 1:
The patent merges the video player interface with the image recognition application interface. The video playing page directly displays resource information related to target objects in video frames, combining what would otherwise be separate applications into one unified interface, thereby improving user experience.
Solution Approach 2:
The system performs preliminary recognition operations during video playback, preparing recognition results in advance. This allows the system to immediately display resource information when users interact with the video, eliminating the need for users to manually capture and upload images, thus significantly improving user experience.
Data Source
AI summary
A video-based information display method and apparatus, an electronic device and a storage medium. The video-based information display method includes: in a process of playing a target video, displaying, on a playing page of the target video, first resource information corresponding to a target object in the target video, in which the target video includes M image frames, and the first resource information is obtained in advance by matching based on target objects in N image frames; in response to triggering a first event in the process of playing the target video, acquiring second resource information corresponding to a target object in a current image frame based on at least one current image frame played by the playing page in a process of triggering the first event; and displaying the second resource information.


