AI Video Object Search Without Playback Interruption
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Users are limited in their ability to search for information about objects within content, such as clothing or vehicles, while the content is being reproduced, as they must stop the content or use external devices.
Innovation Solution
An electronic device that stores video frames for a specific period, receives voice instructions, and uses AI to provide search results for objects within the frames, allowing users to query information without pausing the content.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Loss of information
If the user stops the content reproduction to search for information about objects, then the user can obtain detailed information about objects, but the viewing experience is interrupted and time is lost
Solution Approach 1:
The system pre-identifies and stores information about objects present in the content before the user requests search. This allows the search function to retrieve pre-prepared information instantly without interrupting content playback, thus resolving the contradiction between obtaining object information and maintaining continuous viewing experience
Solution Approach 2:
The system introduces an intermediary search mechanism that operates in the background during content reproduction. The intermediary process captures voice inputs and retrieves object information without requiring the user to stop the main content flow, thereby maintaining both information access and viewing continuity
2Loss of information
If the user uses an external device to search for information about objects, then the user can access comprehensive information, but the operation becomes complex and less intuitive
Solution Approach 1:
The system merges the content reproduction function and the object search function into a single integrated device and interface. Users can perform object searches directly through the same device playing the content, using simple voice commands or on-screen interactions, eliminating the need to switch to external devices and thereby simplifying the operation while maintaining access to comprehensive information
Solution Approach 2:
The system enables self-service object identification and information retrieval by automatically analyzing the content, detecting objects within it, and providing relevant information when users trigger the search function. This automated process reduces the operational burden on users compared to manually using external search tools
3Measurement precision
If the system stores and analyzes multiple video frames for object recognition, then the accuracy of object identification improves, but the device complexity and processing requirements increase
Solution Approach 1:
The system applies partial action by selectively analyzing only certain frames or portions of frames that contain objects of interest, rather than processing every pixel of every frame. This approach maintains high object identification accuracy while reducing the overall computational complexity and resource requirements of the system
Data Source
AI summary
An artificial intelligence (AI) system utilizing a machine learning algorithm such as deep learning for controlling an electronic device when a video is reproduced and a user's voice instruction is received, to acquire a frame corresponding to the time point when the input of the user's voice instruction is received, and obtain a search result for information on objects in the frame using an AI model trained according to at least one of machine learning, a neural network or a deep learning algorithm.


