Mono-Recording Voice Separation for Contact Center Analysis

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Contact centers face challenges in accurately analyzing customer and agent interactions due to the lack of separated voice data in mono recordings, hindering quick and accurate analysis of customer and agent communications.

Innovation Solution

A system and method that record a mono recording of customer-agent interactions, separately record agent voice data, align and subtract it from the mono recording to isolate customer voice data, and convert it to text for linguistic analysis to determine personality types and identify distress events.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of manufacture

If mono recording is used to record customer-agent interactions, then the recording process is simple and compatible with standard telephone networks, but the voice data cannot be separated for accurate analysis

Engineering Contradiction:
Improverecording process simplicityVSAvoidvoice data separation accuracy
Core Design Contradiction:
Ease of manufactureVSMeasurement precision

Solution Approach 1:

The patent segments the mixed mono recording into separate customer voice data and agent voice data through signal processing. The system divides the composite audio signal by identifying and isolating distinct voice sources, enabling separate analysis of each party's communication while maintaining the simplicity of mono recording capture.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent introduces an intermediary signal processing system that acts as a mediator between the mono recording and the analysis requirements. This intermediary component performs voice separation, alignment, and extraction operations, allowing the system to maintain simple mono recording compatibility while achieving precise voice data separation for analysis.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Measurement precision

If separated stereo recording is used for customer and agent voices, then accurate voice separation is achieved, but compatibility with standard telephone networks is lost

Engineering Contradiction:
Improvevoice data separation accuracyVSAvoidtelephone network compatibility
Core Design Contradiction:
Measurement precisionVSAdaptability or versatility

Solution Approach 1:

The patent transitions from the traditional two-dimensional approach (separate stereo channels) to a different dimension by processing mono recording data through temporal and spectral analysis. Instead of requiring separate physical channels, the system extracts separated voice data from the single-channel mono recording by analyzing time-synchronized patterns and frequency characteristics, achieving separation accuracy without sacrificing network compatibility.

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

3Productivity

If voice data analysis is performed on unseparated mono recordings, then the analysis process is quick, but the accuracy of customer and agent communication analysis is hindered

Engineering Contradiction:
Improveanalysis speedVSAvoidcommunication analysis accuracy
Core Design Contradiction:
ProductivityVSMeasurement precision

Solution Approach 1:

The patent performs preliminary voice separation and alignment processing on the mono recording data before the actual communication analysis begins. By pre-separating the customer and agent voice streams and synchronizing them temporally, the system enables subsequent analysis operations to proceed quickly on already-separated data, achieving both speed and accuracy in the overall analysis process.

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS10276153B2Online chat communication analysis via mono-recording system and methods
Publication Date: 2019.04.30 MATTERSIGHT CORP
  • US10276153B2 patent drawing
  • US10276153B2 patent drawing
  • US10276153B2 patent drawing

AI summary

The methods, apparatus, non-transitory computer readable media, and systems described herein include recording a mono recording of a software and a customer inquiry communication using a microphone to interpret and respond to the customer inquiry communication, wherein the mono recording is unseparated and includes customer voice data and audio data generated by the software agent, separately and concurrently recording the software agent audio data in an agent recording, and subtracting agent audio data from the unseparated mono recording to provide a separated recording including only customer voice data.