Automated Utterance Search Using Speech Analyzer
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current voice communication recording and search systems are inefficient, requiring users to manually review extensive recordings to find specific utterances, which is time-consuming and prone to errors, especially when searching for the absence of an utterance-of-interest.
Innovation Solution
Integration of a speech analyzer with an audio player that uses search criteria to identify and locate specific phonemes, words, or phrases within recorded communication sessions, generating visual results and caching processes for real-time review.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If manual review of recorded calls is performed to identify specific utterances, then the user can locate the required call or calls, but the process becomes extremely time-consuming and labor-intensive
Solution Approach 1:
The patent replaces the manual mechanical process of listening to recordings with automated speech recognition technology. The system uses speech analyzers to convert spoken utterances into text and search for them automatically, eliminating the need for human reviewers to manually listen to hours of recordings while maintaining high accuracy in identifying specific utterances.
Solution Approach 2:
The patent introduces an intermediary speech analysis system that acts as a mediator between the recorded calls and the user. This intermediary automatically analyzes recordings, transcribes speech, and presents relevant results to users, significantly reducing the time required while maintaining accurate identification of target utterances.
2Reliability
If the user listens to all identified calls to prove an utterance was not said, then complete accuracy is achieved, but the process becomes prohibitively time-consuming
Solution Approach 1:
The patent applies preliminary action by performing automated speech analysis and transcription before the user needs to verify utterance absence. The system pre-processes all identified calls, creating searchable text representations that allow rapid verification of whether specific utterances are present or absent, eliminating the need for users to listen to entire recordings.
Solution Approach 2:
The patent replaces the time-consuming mechanical process of listening to all calls with automated speech recognition and text search. The system automatically analyzes recordings, transcribes speech, and enables users to quickly verify utterance absence through text-based searching rather than audio playback.
3Productivity
If fast-forwarding is used to skip portions of calls, then review time is reduced, but the risk of missing the target utterance increases
Solution Approach 1:
The patent replaces the unreliable manual fast-forwarding process with automated speech recognition technology. The system continuously analyzes the entire recording and automatically identifies and presents the location of target utterances, maintaining high detection accuracy while significantly improving review speed compared to manual fast-forwarding.
Solution Approach 2:
The patent implements feedback by having the speech analysis system continuously monitor recordings and provide real-time information about detected utterances. The system feedbacks the precise location and context of target utterances to users, enabling rapid navigation to relevant portions without the risk of missing targets that occurs with manual fast-forwarding.
4Quantity of substance
If more detailed search criteria are used to narrow down calls, then the number of calls to review is reduced, but the complexity of the search process increases
Solution Approach 1:
The patent applies universality by creating a multi-functional search system that can handle various types of search criteria (spoken phrases, text inputs, different time ranges, multiple keywords) through a single unified interface. The speech analyzer accepts diverse input formats and automatically processes them, reducing the perceived complexity for users while effectively narrowing down the number of calls to review.
Solution Approach 2:
The patent introduces an intermediary speech analysis system that simplifies complex search operations. The intermediary handles the complexity of processing multiple search criteria, transcribing speech, and filtering results, presenting users with a simplified interface while maintaining the ability to process sophisticated search parameters that effectively reduce the number of calls to review.
Data Source
AI summary
A speech analyzer is integrated or otherwise coupled to an audio player. The speech analyzer is used to identify recorded communication sessions in accordance with a search criterion. A search criterion may be spoken or otherwise communicated to the speech analyzer. Results generated by the speech analyzer are converted into visual information that is presented to a user of the speech analyzer. Results generated by the speech analyzer can be cached for real-time user review while the speech analyzer processes additional stored conversations.


