Audio Fingerprint Video Capture System
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current video capture methods require users to perform multiple actions and consume significant storage space on servers by storing complete audio data of all television programs, making the process cumbersome and inefficient.
Innovation Solution
A video capture system that uses a smart device to detect user-specific actions, record timestamps, and generate audio fingerprint data, allowing the server to automatically synchronize with the television program and retrieve the desired video fragment using electronic program guide information, reducing the need for manual tagging and storing only the latest audio data.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If the server stores complete audio data of all television programs, then the audio fingerprint comparison can be performed, but the storage space of the server is consumed
Solution Approach 1:
The system pre-processes audio data into compact audio fingerprint representations and stores these fingerprints in advance. When a video capture request comes in, the server compares the incoming audio fingerprint with the pre-stored fingerprints to quickly identify the television program, avoiding the need to store and process complete audio data.
Solution Approach 2:
The patent extracts only the essential audio fingerprint features from the complete audio data and stores these extracted features on the server. This extraction process separates the critical identification information from the redundant audio content, significantly reducing storage requirements while maintaining comparison capability.
2Manufacturing precision
If the user manually controls the start tag and end tag of the video fragment, then the video capture precision can be improved, but the operation complexity increases
Solution Approach 1:
The system automatically determines the start and end tags of the video fragment by comparing audio fingerprints and using timestamp information. The server autonomously identifies the precise timing of the desired video content without requiring manual user input for tag selection, making the system self-sufficient in determining capture parameters.
Solution Approach 2:
The system uses timestamp feedback from the user's action and audio fingerprint matching results to automatically calculate and set the start and end tags. The server receives feedback about when the user performed the capturing action and uses this temporal information combined with audio content analysis to precisely determine the video fragment boundaries.
3Measurement precision
If the server compares audio fingerprint data with complete audio data, then the specific television program can be identified, but the processing time increases
Solution Approach 1:
The patent transforms the audio data into a different parameter representation - the audio fingerprint. Instead of comparing raw audio data which is large and time-consuming, the system compares compact fingerprint representations that retain the essential identification characteristics, dramatically reducing comparison time while maintaining accuracy.
Data Source
AI summary
The present disclosure illustrates a video capture system. The video capture system comprises a smart device and a first server. The smart device is configured to detect a user specific action. When the smart device detects the user specific actions, the smart device records a time stamp of the user specific action, and generate an audio fingerprint data based on a audio data of a specific television program showed on a display. The first server receives the time stamp and the audio fingerprint data, and finds the specific television program corresponding to the audio fingerprint data according to the audio fingerprint data and electronic program guide information. Then, the first server obtains a start tag based onbased on the time stamp. The start tag is a starting time of a video fragment.


