Blending Naturalistic and Synthetic Speech Cues for Cognitive Training

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current cognitive training methods for age-related cognitive decline and Mild Cognitive Impairment (MCI) face limitations such as lack of generalization and enduring effect, with synthesized speech stimuli often being too unnatural for participants to effectively identify sounds, leading to difficulty in progressing through training exercises.

Innovation Solution

A training program that modulates listener attention by gradually blending naturalistic cues with synthesized formant transitions in speech stimuli, allowing participants to transition from natural-sounding to formant-synthesized phonemes, enhancing auditory representation and processing efficiency through game-like exercises that engage attention and reward systems.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Productivity

If synthesized formant transitions are used in speech stimuli for training, then auditory processing efficiency is improved, but the naturalness of the speech stimuli deteriorates, making it difficult for participants to identify sounds

Engineering Contradiction:
Improveauditory processing efficiencyVSAvoidsound identification difficulty
Core Design Contradiction:
ProductivityVSEase of operation

Solution Approach 1:

The patent applies dynamics by creating a dynamic blending ratio that transitions over time. The speech stimulus starts with a higher proportion of naturalistic cues and progressively increases the synthesized formant transition component. This dynamic adjustment allows participants to adapt their auditory processing to the changing stimulus characteristics, resolving the contradiction between processing efficiency and sound naturalness.

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The patent changes the parameter of blending ratio between naturalistic and synthesized speech components across different time points. By systematically varying this parameter from 0.0 to 1.0, the system optimizes the balance between maintaining naturalness for initial recognition and enhancing synthetic cues for improved processing efficiency, thereby resolving the technical contradiction.

Inventive Principle:
Principle #35Parameter changes

2Reliability

If cognitive training methods are used to address age-related cognitive decline, then cognitive performance is improved, but the generalization and enduring effect of training deteriorates

Engineering Contradiction:
Improvecognitive performance improvementVSAvoidtraining generalization
Core Design Contradiction:
ReliabilityVSAdaptability or versatility

Solution Approach 1:

The patent applies universality by designing a training system that addresses multiple cognitive functions simultaneously through a single integrated approach. The adaptive audio processing technique enhances not only auditory processing but also attention, memory, and executive functions by engaging multiple neuromodulatory pathways, thereby achieving broad generalization of training effects across different cognitive domains.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The patent implements feedback mechanisms that continuously monitor participant performance and adjust training parameters accordingly. This feedback loop ensures that training effects are consolidated and generalized by adapting the difficulty and type of exercises based on individual progress, thereby improving both reliability of improvement and versatility of training outcomes.

Inventive Principle:
Principle #23Feedback

3Measurement precision

If attention is modulated toward synthetic formant cues, then speech distinction recognition is improved, but the complexity of stimulus processing increases

Engineering Contradiction:
Improvespeech distinction recognition accuracyVSAvoidstimulus processing complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The patent applies preliminary action by pre-processing the speech stimulus to create a blended version that already contains both naturalistic and synthesized components in optimized proportions. This preliminary blending reduces the computational complexity during actual processing, as the system doesn't need to separately analyze and weigh different components in real-time, thereby improving recognition accuracy without proportionally increasing processing complexity.

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS8210851B2Method for modulating listener attention toward synthetic formant transition cues in speech stimuli for training
Publication Date: 2012.07.03 POSIT SCIENCE CORP
  • US8210851B2 patent drawing
  • US8210851B2 patent drawing
  • US8210851B2 patent drawing

AI summary

A method on a computing device for enhancing the memory and cognitive ability of an older adult by requiring the adult to differentiate between rapidly presented stimuli. The method utilizes a sequence of phonemes from a confusable pair which are systematically manipulated to make discrimination between the phonemes less difficult or more difficult based on the success of the adult, such as processing the consonant and vowel portions of the phonemes by emphasizing the portions, stretching the portions, and/or separating the consonant and vowel portions by time intervals. As the adult improves in auditory processing, the discriminations are made progressively more difficult by reducing the amount of processing to that of normal speech. Introductory phonemes may each include a blend of a formant-synthesized phoneme and an acoustically naturalistic phoneme that substantially replicates the spectro-temporal aspects of a naturally produced phoneme, with the blends progressing from substantially natural-sounding to substantially formant-synthesized.