Audio Resegmentation via Dynamic Word Cloud Matching
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current audio content retrieval systems fail to efficiently search and provide precise segments of audio broadcasts in real-time, as existing technologies do not effectively utilize dynamic keyword associations to isolate relevant audio segments from large datasets.
Innovation Solution
A method and system that utilize a 'word cloud' based on search queries to dynamically identify and resegment audio content, continuously updating with time-relevant terms to match user queries, allowing for real-time retrieval of relevant audio segments by applying these terms to the audio content.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Quantity of substance
If audio broadcasts are archived and stored for public access, then the quantity of available audio content increases, but the ability to efficiently search and retrieve precise segments in real-time deteriorates
Solution Approach 1:
The system performs preliminary actions by creating word clouds from archived audio content in advance, organizing keywords and metadata before search requests arrive. This pre-processing enables rapid retrieval during real-time searches without compromising efficiency
Solution Approach 2:
The patent introduces word clouds as an intermediary data structure between archived audio content and search queries. These word clouds serve as mediators that bridge the large audio database and user searches, enabling efficient real-time retrieval by matching query terms against pre-generated keyword associations
2Loss of information
If the entire audio segment is provided to the requestor, then completeness of information is improved, but the precision and relevance of the retrieved content deteriorates
Solution Approach 1:
The system segments the audio content into distinct portions based on word cloud matches and query relevance. Instead of providing the entire audio segment, the patent divides it into relevant sub-segments that directly answer the user's query, improving precision while maintaining necessary completeness
Solution Approach 2:
The patent applies local quality by providing different levels of detail for different parts of the audio content. Highly relevant segments matching query terms are extracted and provided with greater precision, while less relevant portions are either excluded or provided in summarized form, optimizing both completeness and precision
Data Source
AI summary
Methods and systems are disclosed for providing segments of audio broadcasts in responses to user queries. The segments are provided, for example, in real time, with the segments being precise segments responsive to the search query or other search request, located, isolated, and provided to the requesting party or entity. The queried audio segments are found by matching a medium, known as a “word cloud”, obtained based on the query, with audio segments, obtained based on the query. The word cloud is continuously updated with time relevant terms, associated with the subject of the word cloud. The word cloud is applied to the audio segment to obtain the most relevant audio segments, which are resegmented from the audio segment.


