Audio Fingerprint Program Identification Accuracy
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional methods for program identification in television and radio media often result in reduced accuracy due to users entering incorrect keywords, leading to incorrect program requests and failures in acquiring the correct program.
Innovation Solution
A method and apparatus for program identification using audio fingerprints, where a first audio fingerprint is acquired and matched against a predetermined fingerprint database to provide the associated program, improving accuracy by reducing errors caused by incorrect keyword entry.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If users manually enter keywords to identify programs, then the system can process simple text input, but the accuracy of program identification is reduced due to incorrect keyword entry
Solution Approach 1:
The patent replaces the manual keyboard input mechanism with an acoustic signal capture mechanism. Instead of requiring users to physically type keywords, the system uses the terminal's microphone to capture audio signals from the environment (such as program audio or spoken keywords), automatically converts them to text through speech recognition, and uses this for program identification. This substitution eliminates the error-prone manual entry step while maintaining ease of operation.
Solution Approach 2:
The patent introduces speech recognition technology as an intermediary between the user's spoken input and the program identification system. The user speaks the program name or keywords, the speech recognition module converts this audio signal into text, and then the text processing module uses this converted text for program identification. This intermediary layer automatically handles the conversion process, reducing errors from manual keyboard entry while keeping the interaction simple for the user.
2Loss of information
If the system requires users to input keywords through keyboard or touch screen, then text input can be obtained, but users may acquire wrong keywords leading to incorrect program requests
Solution Approach 1:
The patent replaces complex manual text input devices (keyboard, touch screen) with a simple acoustic capture process. The terminal's built-in microphone captures audio signals containing the program name or keywords, and speech recognition technology automatically converts this audio into accurate text. This substitution maintains high information accuracy while significantly simplifying the user interaction - users only need to speak rather than manually type or select from keyboards or touch screens.
3Reliability
If conventional keyword-based identification is used, then program requests can be sent to server, but the accuracy is reduced and correct program may not be acquired
Solution Approach 1:
The patent performs preliminary speech recognition and text conversion before the program identification process begins. By capturing audio signals and converting them to text in advance, the system ensures that accurate program names or keywords are ready for identification before being sent to the server. This preliminary action increases reliability by ensuring accurate input data is prepared beforehand, reducing the need for retries or corrections that would reduce overall efficiency.
Data Source
AI summary
Systems and methods are provided for program identification. For example, a first audio fingerprint corresponding to a first audio signal is acquired; whether one or more second audio fingerprints in a predetermined fingerprint database match with the first audio fingerprint is detected, a second audio fingerprint corresponding to a second audio signal; and in response to one of the second audio fingerprints matching with the first audio fingerprint, a program associated with the matching second audio signal is provided as a result for program identification associated with the first audio signal.


