Voice Directed Visual Search Object Detection
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional reverse visual search methods often return a large number of irrelevant results, making it difficult for users to find relevant information due to the lack of tailoring to their interests.
Innovation Solution
Voice directed context sensitive visual searching, where a voice query related to visual content is used to detect objects within the content, incorporating contextual information to perform a targeted search, thereby improving the accuracy of search results.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Quantity of substance
If conventional reverse visual search is used, then image-based search is enabled, but a large number of irrelevant results are returned
Solution Approach 1:
The patent applies local quality by detecting specific objects within the image and using those detected objects as search queries. Instead of searching the entire image uniformly, the system identifies particular objects (e.g., animals, plants, products) and performs targeted searches based on those specific detections, thereby improving result relevance while maintaining comprehensive search coverage
Solution Approach 2:
The patent uses an intermediary approach by introducing object detection as a mediator between the input image and the search query. The object detection system acts as a bridge that extracts meaningful objects from the image and converts them into search terms, enabling the search engine to retrieve relevant results without requiring direct user input of search terms
2Productivity
If conventional reverse visual search is used, then image-based search is enabled, but users spend excessive time finding relevant information
Solution Approach 1:
The patent applies preliminary action by performing object detection on the input image before executing the search. The system pre-processes the image to identify and extract objects of interest, preparing search queries in advance based on the detected objects. This preliminary object extraction and query generation significantly reduces the time users would otherwise spend manually identifying and typing search terms
Solution Approach 2:
The system implements self-service by automatically generating search queries from the input image without requiring user intervention to specify search terms. The object detection system autonomously identifies relevant objects and formulates search queries, enabling the search engine to serve user needs automatically and efficiently, thereby reducing the time and effort users must invest in the search process
Data Source
AI summary
Various technologies described herein pertain to voice directed context sensitive visual searching. Visual content can be rendered on a display, and a voice directed query related to the visual content can be received. Contextual information related to the visual content can also be identified. Moreover, a search word recognized from the voice directed query and/or the contextual information can be used to detect an object from the visual content, where the object can be a part of the visual content. Further, a search can be performed using the object detected from the visual content, and a result of the search can be rendered on the display.


