Utterance Condition Determination via Personalized Backchannel Frequency

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing technologies face challenges in accurately determining emotional conditions of speakers based on backchannel feedback due to individual variations in the frequency and interval of backchannel responses, leading to incorrect assessments.

Innovation Solution

An utterance condition determination device that estimates an average backchannel frequency and calculates a satisfaction level by analyzing voice signals from both speakers, using a backchannel frequency calculation unit and a determination unit to assess the emotional state of the second speaker.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If backchannel feedback frequency is used to determine emotional conditions, then emotional state detection is enabled, but individual variations in feedback patterns lead to determination inaccuracy

Engineering Contradiction:
Improveemotional condition determination accuracyVSAvoiddetermination reliability
Core Design Contradiction:
Measurement precisionVSReliability

Solution Approach 1:

The system performs preliminary action by estimating each speaker's average backchannel frequency before the actual emotional condition determination. This preliminary estimation creates a personalized baseline for each speaker, allowing the system to account for individual variations in backchannel feedback patterns. The average backchannel frequency is calculated over a predetermined period and stored as reference data, which is then used to normalize and compare actual backchannel frequencies during emotional state detection, thereby improving determination accuracy despite individual differences.

Inventive Principle:
Principle #10Preliminary action

2Ease of operation

If fixed threshold determination is used for backchannel frequency, then determination process is simplified, but individual variations cause incorrect assessments

Engineering Contradiction:
Improvedetermination process simplicityVSAvoidemotional condition determination accuracy
Core Design Contradiction:
Ease of operationVSMeasurement precision

Solution Approach 1:

The system applies local quality by transitioning from a uniform fixed threshold applied to all speakers to personalized dynamic thresholds based on each speaker's estimated average backchannel frequency. Each speaker receives a customized determination threshold derived from their individual communication patterns. This localized approach maintains operational simplicity through automated personalization while significantly improving measurement precision by adapting to individual variations in backchannel feedback behavior.

Inventive Principle:
Principle #3Local quality

3Measurement precision

If average backchannel frequency estimation is implemented, then individual variations are accounted for, but system complexity increases

Engineering Contradiction:
Improveemotional condition determination accuracyVSAvoidsystem structure complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The system implements self-service by automatically estimating each speaker's average backchannel frequency and generating personalized determination thresholds without requiring manual configuration or intervention. The system autonomously adapts to each speaker's individual patterns by processing their communication history and computing their unique baseline metrics. This self-configuration capability improves measurement precision through personalization while minimizing the increase in operational complexity, as the system performs these calculations automatically during normal operation.

Inventive Principle:
Principle #25Self-service

Data Source

PatentEP3136388B1Utterance condition determination apparatus and method
Publication Date: 2019.11.27 FUJITSU LTD
  • EP3136388B1 patent drawingFigure 1
  • EP3136388B1 patent drawingFigure 2
  • EP3136388B1 patent drawingFigure 3

AI summary

An utterance condition determination device (5) includes an average backchannel frequency estimation unit (504, 514, 524, 536, 544), a backchannel frequency calculation unit (503, 513, 523, 534, 543), and a determination unit (505, 515, 525, 538, 546). The average backchannel frequency estimation unit estimates an average backchannel frequency that represents a backchannel frequency of the second speaker in a period of time from a voice start time of a voice signal of the second speaker to a predetermined time based on a voice signal of the first speaker and the voice signal of the second speaker. The backchannel frequency calculation unit calculates the backchannel frequency of the second speaker for each unit time based on the voice signal of the first speaker and the voice signal of the second speaker. The determination unit determines a satisfaction level of the second speaker based on the average backchannel frequency estimated in the average backchannel frequency estimation unit and the backchannel frequency calculated in the backchannel frequency calculation unit.