Speech Recognition Audio Collection Under Ambient Noise

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Speech recognition technologies face interference from ambient noise, leading to incorrect recognition of user input due to the sensitivity in collecting audio signals, which fails to distinguish between user audio and ambient noise effectively.

Innovation Solution

A device and method that utilize a processor to collect audio data, determine the presence of ambient noise by matching semantic content with a database, and adjust collection conditions such as input volume and voltage amplitude to improve recognition accuracy by prompting users to adjust their voice levels or reduce ambient noise.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If the sensitivity for collecting audio signal is increased, then the speech recognition capability is improved, but the ambient noise interference is also increased

Engineering Contradiction:
Improvespeech recognition capabilityVSAvoidambient noise interference
Core Design Contradiction:
Measurement precisionVSObject-affected harmful factors

Solution Approach 1:

The patent applies dynamics by making the voltage amplitude adjustable rather than fixed. The collecting circuit dynamically changes its collection sensitivity based on environmental conditions, switching between different voltage amplitudes to optimize between capturing user speech and rejecting ambient noise.

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The patent changes the physical parameter of voltage amplitude used for collecting audio data. By adjusting this parameter, the system can control the sensitivity of the collecting circuit, thereby resolving the contradiction between capturing weak user speech signals and rejecting strong ambient noise signals.

Inventive Principle:
Principle #35Parameter changes

2Power

If the voltage amplitude for collecting audio data is increased, then the audio signal strength is improved, but the ambient noise collection is also increased

Engineering Contradiction:
Improveaudio signal strengthVSAvoidambient noise collection
Core Design Contradiction:
PowerVSObject-affected harmful factors

Solution Approach 1:

The system dynamically adjusts the voltage amplitude based on the collected audio characteristics. When ambient noise is detected, the system reduces the voltage amplitude to minimize noise collection while maintaining sufficient signal strength for user speech recognition.

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The patent implements feedback by analyzing the collected audio data to determine whether it contains ambient noise, then using this information to adjust the collection conditions. The processor provides feedback signals to modify the voltage amplitude based on the audio environment analysis.

Inventive Principle:
Principle #23Feedback

3Measurement precision

If the semantic content matching is performed to identify ambient noise, then the recognition accuracy is improved, but the processing time is increased

Engineering Contradiction:
Improverecognition accuracyVSAvoidprocessing time
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

The patent performs partial matching by comparing semantic content against a database of known ambient noise patterns. Rather than exhaustive analysis, the system uses targeted semantic matching to efficiently identify ambient noise while maintaining reasonable processing speed.

Inventive Principle:
Principle #16Partial or excessive action

Data Source

PatentUS10861447B2Device for recognizing speeches and method for speech recognition
Publication Date: 2020.12.08 BOE TECHNOLOGY GROUP CO LTD
  • US10861447B2 patent drawing
  • US10861447B2 patent drawing
  • US10861447B2 patent drawing

AI summary

The embodiments of the present disclosure provide a device for recognizing speeches and a method for speech recognition. The device for recognizing speeches may comprise a processor, configured to execute instructions stored in the memory, to: perform speech recognition on the collected audio data to obtain a semantic content of the audio data; match the obtained semantic content with a semantic data stored in the database; determine whether the audio data contains ambient noise audio information and audio information of a user, in response to determining that the obtained semantic content does not match with the semantic data; and change conditions for collecting the audio data and control to collect the audio data with the changed conditions, in response to determining that the audio data contains the ambient noise audio information and the audio information of the user.