Augmented Reality Product Recognition via Optical Image Analysis
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional methods for providing users with information through computing devices are inefficient, as they require manual input and lack effective ways to recognize text or objects in real-time, leading to tedious tasks and inaccuracies in identifying products or services.
Innovation Solution
A portable computing device uses augmented reality (AR) to analyze images from a camera's field of view, recognizing text or objects and displaying related product listings or advertisements, even with partial recognition, to enable users to purchase products directly from an electronic marketplace.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Productivity
If manual input methods are used for providing information to users, then users can access information through computing devices, but the process becomes tedious and time-consuming
Solution Approach 1:
The patent replaces manual typing and mechanical input methods with optical recognition technology. The system captures images of text or objects using a camera and automatically identifies them through image processing algorithms, substituting the mechanical action of manual input with an optical recognition system that processes visual information directly.
Solution Approach 2:
The system enables self-service by allowing the device to automatically capture, recognize, and process information without requiring user intervention for typing or data entry. The camera automatically captures the field of view, the recognition system independently identifies text or objects, and the device autonomously retrieves and displays related information, creating a self-service information access workflow.
2Speed
If text or object recognition is performed in real-time, then users can receive immediate information about products or services, but recognition accuracy may be compromised with partial recognition
Solution Approach 1:
The patent applies partial recognition by identifying and processing only the most salient or recognizable portions of text or objects in the captured image. Rather than requiring complete and perfect recognition of entire text strings or objects, the system identifies partial matches and uses those to retrieve relevant information, accepting that some portions may be partially recognized rather than achieving complete accuracy.
Solution Approach 2:
The system introduces an intermediary processing layer that bridges the gap between partial recognition and complete information retrieval. The recognition system identifies what it can from the image, then uses this partial information as an intermediary step to query databases and retrieve comprehensive product or service information, mediating between incomplete recognition data and complete information delivery.
3Loss of information
If AR applications provide detailed product information, then users can make informed purchasing decisions, but the device complexity increases
Solution Approach 1:
The patent implements multi-functionality by integrating multiple capabilities into a unified AR application: the device performs camera capture, image processing, text and object recognition, database querying, and information display all through a single integrated system. This universal approach consolidates what could be separate complex functions into one cohesive application that handles the entire information retrieval workflow.
Solution Approach 2:
The system merges the recognition engine, database access, and display functions into an integrated workflow. The image capture, text/object identification, information retrieval, and presentation are combined into a seamless process where the AR application serves as a unified interface that consolidates multiple technical functions into a single user-facing system.
Data Source
AI summary
Various embodiments enable a computing device to perform tasks such as processing an image to recognize text or an object in an image to identify a particular product or related products associated with the text or object. In response to recognizing the text or the object as being associated with a product available for purchase from an electronic marketplace, one or more advertisements or product listings associated with the product can be displayed to the user. Accordingly, additional information for the associated product can be displayed, enabling the user to learn more about and purchase the product from the electronic marketplace through the portable computing device.


