Mobile Content Fingerprinting via Optical Capture
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current methods for identifying and accessing video/audio content playing on other devices lack efficiency and convenience, as users must manually recognize or remember titles, which can be cumbersome and prone to errors.
Innovation Solution
A mobile device captures a clip of the video/audio using its camera/microphone, generates a fingerprint, and transmits it to a network device for identification, allowing users to receive notifications about the original content, enabling easy viewing, purchase, or bookmarking.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If users manually recognize or remember content titles, then they can identify content, but the process becomes cumbersome and error-prone
Solution Approach 1:
The patent replaces the manual mechanical process of remembering and typing titles with an automated optical system. The camera captures visual information from the display, and image processing algorithms automatically extract and match content identifiers, eliminating the need for manual cognitive and typing efforts while improving accuracy.
Solution Approach 2:
The system creates a visual copy of the content identifier by capturing an image of the display screen. This optical copy is then processed through fingerprint extraction and matching algorithms to identify the content, replacing the need for users to manually recall or transcribe titles.
2Extent of automation
If the system captures and processes video clips for identification, then content recognition becomes automated, but the device complexity increases
Solution Approach 1:
The patent extracts only the essential visual features from captured video frames to create content fingerprints. Instead of processing entire video clips, the system identifies and extracts key identifier elements from the display, significantly reducing computational complexity while maintaining automation.
Solution Approach 2:
The system performs partial processing by capturing only the necessary portions of the display screen that contain content identifiers. Rather than processing complete video content, it focuses on extracting and matching specific fingerprint features, reducing overall system complexity.
3Loss of information
If users examine products in stores or online, then they can sample goods, but they cannot easily identify or access the original content later
Solution Approach 1:
The system performs preliminary capture and fingerprint extraction of content identifiers during the initial viewing experience. By pre-processing and storing the visual fingerprint data when the user first encounters the content, the system enables rapid later identification and access without requiring users to remember or manually record information.
Data Source
AI summary
A device may include a video camera for capturing a video clip, a processor, a transmitter, and a receiver. The processor may be configured to receive, from the video camera, the video clip that is shown on a display screen of a content presentation device. The transmitter may be configured to send the video clip or a fingerprint of the video clip to a remote device. The receiver may be configured to receive, from the remote device, an identity of content whose fingerprints match the fingerprint.


