Impression Formation Control via Listener Stimulation

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing methods for controlling impression formation in listeners through voice synthesis alter voice features, potentially misrepresenting the speaker's intention.

Innovation Solution

An impression formation control device that acquires a speaker's speech voice signal, extracts voice features, determines the bias in listener impression based on these features, and generates a stimulation control signal to adjust the listener's perception without altering the speaker's intention.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If voice synthesis technique is used to control impression formation, then impression formation control is achieved, but voice feature is altered causing speaker intention to be misrepresented

Engineering Contradiction:
Improveimpression formation controlVSAvoidspeaker intention
Core Design Contradiction:
ReliabilityVSLoss of information

Solution Approach 1:

The system segments the voice signal processing into two independent paths: extracting acoustic features (pitch, timbre, rhythm) from the original speaker voice while separately synthesizing linguistic content. This allows selective manipulation of impression-related acoustic features without altering the core semantic message, thus maintaining speaker intention while achieving impression formation control.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The system introduces an intermediary processing layer that analyzes the original voice signal to extract acoustic characteristics, then uses these characteristics to guide synthetic voice generation. This intermediary process ensures that the synthesized voice maintains the essential intensional properties of the original speaker while allowing controlled modification of acoustic features that influence listener impression.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Adaptability or versatility

If voice features are extracted and manipulated to control listener bias, then impression formation is controlled, but the original voice signal integrity is compromised

Engineering Contradiction:
Improveimpression control capabilityVSAvoidvoice signal integrity
Core Design Contradiction:
Adaptability or versatilityVSStability of the object's composition

Solution Approach 1:

The system applies local quality modification by selectively altering specific acoustic features (such as pitch contour, spectral characteristics, or temporal patterns) that are known to influence listener impression, while preserving other aspects of the voice signal that carry the core speaker identity and intention. This localized manipulation achieves impression control without compromising overall voice signal integrity.

Inventive Principle:
Principle #3Local quality

3Reliability

If synthesized voice is used to alter listener perception, then bias control is achieved, but naturalness of communication is reduced

Engineering Contradiction:
Improvebias control effectivenessVSAvoidcommunication naturalness
Core Design Contradiction:
ReliabilityVSEase of operation

Solution Approach 1:

The system changes specific acoustic parameters of the voice signal (such as fundamental frequency, formant structure, or speech rate) that are scientifically established to influence listener perception and bias. By carefully selecting and adjusting these parameters within natural ranges, the system achieves effective bias control while maintaining the naturalness and authenticity of the communication.

Inventive Principle:
Principle #35Parameter changes

Data Source

PatentUS20250182777A1Impression formation control device, method and program
Publication Date: 2025.06.05 NT T INC
  • US20250182777A1 patent drawing
  • US20250182777A1 patent drawing
  • US20250182777A1 patent drawing

AI summary

According to an embodiment of the present invention, when impression formation for a listener with respect to a speaker is controlled, a speech voice signal of the speaker is acquired, a voice feature is extracted from the speech voice signal, a bias to an impression made on the listener by the speech voice signal is determined based on the extracted voice feature, a bias control signal for controlling the bias occurs based on a determination result of the bias and information indicating a preset control direction of the bias, and a stimulation control signal for giving external stimulation to the listener is generated in accordance with the bias control signal, and the generated stimulation control signal is output.