Screen Image Search Using Position and Content Type
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional search systems rely on text or meta information, failing to effectively utilize position information on a screen for search queries, limiting users who remember information based on screen locations.
Innovation Solution
A computer-implemented method that receives location information for a screen image region, along with a type indicator for text or image, and performs a search using this information to identify and retrieve corresponding screen images, displaying them as search results.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If conventional search systems use text or meta information for search queries, then the search can be performed using keywords, but the system fails to utilize position information on screen which limits users who remember information based on screen locations
Solution Approach 1:
The patent adds a spatial dimension to traditional text-based search by incorporating screen position coordinates (x, y) and region information as new search dimensions. This allows users to search not only by keyword content but also by where information appeared on screen, transforming a one-dimensional text search into a multi-dimensional search that includes spatial location.
Solution Approach 2:
The patent segments the screen into searchable regions and captures position information for different elements within those regions. By dividing the search space into text regions, image regions, and their respective positions, the system can selectively search based on different segment types and their locations, enabling precise retrieval based on spatial memory.
2Measurement precision
If the system captures and stores position information for screen images, then search accuracy based on spatial memory is improved, but the system complexity increases
Solution Approach 1:
The patent performs preliminary capture and storage of position information for text and image elements on screens during normal usage. By pre-processing and storing location data (x, y coordinates, region boundaries) alongside content data, the system prepares search-ready information in advance, reducing the complexity of real-time search operations while maintaining high accuracy.
Solution Approach 2:
The patent introduces an intermediary layer that bridges between the screen display system and the search system. This intermediary component captures position information and maintains a mapping between screen locations and content elements, acting as a mediator that translates spatial positions into searchable data without requiring direct integration between display and search systems.
3Ease of operation
If the system requires users to specify keywords or metadata for search, then the search query is simple to formulate, but users who remember information by location cannot effectively search
Solution Approach 1:
The patent creates a universal search system that accepts multiple types of search inputs: traditional keywords, position coordinates, region descriptions, and combinations thereof. This multi-functional search interface allows users to search using whichever method best matches their memory - whether they remember what the content was (keyword search) or where it was located (position-based search), making the system adaptable to different user recall patterns.
Data Source
AI summary
Provided are techniques for performing a search based on position information. A search request that provides location information for a region of a screen image is received. A selection of a type indicator is received, where the type indicator indicates one of a text item and an image. In response to the type indicator indicating the text item, one or more of the text item and a date and time are received. A search is performed using the location information and the one or more of the text item and the date and time to identify one or more screen image identifiers of one or more corresponding screen images of a plurality of screen images. The one or more screen image identifiers are used to retrieve the one or more corresponding screen images. The one or more corresponding screen images are displayed as search results.


