Speech Recognition Guidance Speech Interference Removal

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional speech recognition systems require an additional microphone to remove guidance speech, leading to high costs and complexity, and result in a decline in recognition rates when guidance speech is outputted simultaneously with recognition speech.

Innovation Solution

A speech recognition apparatus that uses the guidance speech-data as a reference to remove the guidance speech from the recognition speech-data, eliminating the need for an additional microphone by converting and synchronizing the sampling rates of guidance and reference speech-data to match the recognition speech-data, thereby preventing recognition rate decline.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If an additional microphone is used to remove guidance speech, then the recognition rate is maintained, but the device complexity and cost increase

Engineering Contradiction:
Improverecognition rateVSAvoidmicrophone configuration
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

The patent creates a reference speech signal by copying the guidance speech output and processing it to match the characteristics of the actual guidance speech picked up by the microphone. This virtual reference signal is then used for subtraction to remove guidance speech interference, eliminating the need for an additional reference microphone.

Inventive Principle:
Principle #26Copying

Solution Approach 2:

The patent introduces a reference speech generation unit that acts as an intermediary between the guidance speech output and the speech recognition process. This unit synthesizes a reference signal that mediates the subtraction operation, allowing guidance speech removal without requiring a separate physical microphone.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Reliability

If an additional microphone is used to remove guidance speech, then the recognition rate is maintained, but the system cost increases

Engineering Contradiction:
Improverecognition rateVSAvoidsystem cost
Core Design Contradiction:
ReliabilityVSEase of manufacture

Solution Approach 1:

The patent creates a reference speech signal by copying the guidance speech output and processing it to match the characteristics of the actual guidance speech picked up by the microphone. This virtual reference signal is then used for subtraction to remove guidance speech interference, eliminating the need for an additional reference microphone.

Inventive Principle:
Principle #26Copying

Solution Approach 2:

The system uses its own guidance speech output as the basis for creating the reference signal, making the system self-sufficient. The guidance speech that is already being outputted is reused and processed to generate the reference signal needed for interference removal, eliminating external hardware requirements.

Inventive Principle:
Principle #25Self-service

3Ease of operation

If guidance speech is outputted simultaneously with recognition speech, then user guidance is provided, but the recognition rate decreases due to interference

Engineering Contradiction:
Improveguidance provisionVSAvoidrecognition rate
Core Design Contradiction:
Ease of operationVSReliability

Solution Approach 1:

The patent extracts the guidance speech component from the mixed speech signal by subtracting the generated reference speech signal from the microphone input. This separation isolates the user's recognition speech from the guidance speech interference, enabling accurate recognition even during simultaneous output.

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The patent converts the harmful effect of guidance speech interference into a beneficial process. By using the guidance speech output itself to generate the reference signal for subtraction, the system transforms the interfering signal into a tool for its own removal, maintaining recognition accuracy during simultaneous operation.

Inventive Principle:
Principle #22Blessing in disguise (Convert harm into benefit)

Data Source

PatentUS10127910B2Speech recognition apparatus and computer program product for speech recognition
Publication Date: 2018.11.13 DENSO CORP
  • US10127910B2 patent drawing
  • US10127910B2 patent drawing
  • US10127910B2 patent drawing

AI summary

In a speech recognition apparatus, a speech driver fetches a guidance speech-data as a reference speech-data, and outputs the reference speech-data to a recognition core unit. A guidance speech into which the guidance speech-data is converted is outputted by a speaker to cause a microphone to receive the outputted guidance speech, which will be converted into an inputted guidance speech-data. Even in such case, a speech recognition engine removes the inputted guidance speech-data by using, as the reference speech-data, the guidance speech-data that is before being converted into the outputted guidance speech.