Automated Utterance Search Using Speech Analyzer

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current voice communication recording and search systems are inefficient, requiring users to manually review extensive recordings to find specific utterances, which is time-consuming and prone to errors, especially when searching for the absence of an utterance-of-interest.

Innovation Solution

Integration of a speech analyzer with an audio player that uses search criteria to identify and locate specific phonemes, words, or phrases within recorded communication sessions, generating visual results and caching processes for real-time review.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If manual review of recorded calls is performed to identify specific utterances, then the user can locate the required call or calls, but the process becomes extremely time-consuming and labor-intensive

Engineering Contradiction:
Improveaccuracy of utterance identificationVSAvoidtime required to review recordings
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

The patent replaces the manual mechanical process of listening to recordings with automated speech recognition technology. The system uses speech analyzers to convert spoken utterances into text and search for them automatically, eliminating the need for human reviewers to manually listen to hours of recordings while maintaining high accuracy in identifying specific utterances.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Solution Approach 2:

The patent introduces an intermediary speech analysis system that acts as a mediator between the recorded calls and the user. This intermediary automatically analyzes recordings, transcribes speech, and presents relevant results to users, significantly reducing the time required while maintaining accurate identification of target utterances.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Reliability

If the user listens to all identified calls to prove an utterance was not said, then complete accuracy is achieved, but the process becomes prohibitively time-consuming

Engineering Contradiction:
Improvecertainty of utterance absenceVSAvoidtime to review all calls
Core Design Contradiction:
ReliabilityVSLoss of time

Solution Approach 1:

The patent applies preliminary action by performing automated speech analysis and transcription before the user needs to verify utterance absence. The system pre-processes all identified calls, creating searchable text representations that allow rapid verification of whether specific utterances are present or absent, eliminating the need for users to listen to entire recordings.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent replaces the time-consuming mechanical process of listening to all calls with automated speech recognition and text search. The system automatically analyzes recordings, transcribes speech, and enables users to quickly verify utterance absence through text-based searching rather than audio playback.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

3Productivity

If fast-forwarding is used to skip portions of calls, then review time is reduced, but the risk of missing the target utterance increases

Engineering Contradiction:
Improvespeed of call reviewVSAvoidaccuracy of utterance detection
Core Design Contradiction:
ProductivityVSReliability

Solution Approach 1:

The patent replaces the unreliable manual fast-forwarding process with automated speech recognition technology. The system continuously analyzes the entire recording and automatically identifies and presents the location of target utterances, maintaining high detection accuracy while significantly improving review speed compared to manual fast-forwarding.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Solution Approach 2:

The patent implements feedback by having the speech analysis system continuously monitor recordings and provide real-time information about detected utterances. The system feedbacks the precise location and context of target utterances to users, enabling rapid navigation to relevant portions without the risk of missing targets that occurs with manual fast-forwarding.

Inventive Principle:
Principle #23Feedback

4Quantity of substance

If more detailed search criteria are used to narrow down calls, then the number of calls to review is reduced, but the complexity of the search process increases

Engineering Contradiction:
Improvenumber of calls to reviewVSAvoidcomplexity of search criteria
Core Design Contradiction:
Quantity of substanceVSDevice complexity

Solution Approach 1:

The patent applies universality by creating a multi-functional search system that can handle various types of search criteria (spoken phrases, text inputs, different time ranges, multiple keywords) through a single unified interface. The speech analyzer accepts diverse input formats and automatically processes them, reducing the perceived complexity for users while effectively narrowing down the number of calls to review.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The patent introduces an intermediary speech analysis system that simplifies complex search operations. The intermediary handles the complexity of processing multiple search criteria, transcribing speech, and filtering results, presenting users with a simplified interface while maintaining the ability to process sophisticated search parameters that effectively reduce the number of calls to review.

Inventive Principle:
Principle #24Intermediary (Mediator)

Data Source

PatentUS8050923B2Automated utterance search
Publication Date: 2011.11.01 VERINT AMERICAS INC
  • US8050923B2 patent drawing
  • US8050923B2 patent drawing
  • US8050923B2 patent drawing

AI summary

A speech analyzer is integrated or otherwise coupled to an audio player. The speech analyzer is used to identify recorded communication sessions in accordance with a search criterion. A search criterion may be spoken or otherwise communicated to the speech analyzer. Results generated by the speech analyzer are converted into visual information that is presented to a user of the speech analyzer. Results generated by the speech analyzer can be cached for real-time user review while the speech analyzer processes additional stored conversations.