Speech Recognition Audio Collection Under Ambient Noise
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Speech recognition technologies face interference from ambient noise, leading to incorrect recognition of user input due to the sensitivity in collecting audio signals, which fails to distinguish between user audio and ambient noise effectively.
Innovation Solution
A device and method that utilize a processor to collect audio data, determine the presence of ambient noise by matching semantic content with a database, and adjust collection conditions such as input volume and voltage amplitude to improve recognition accuracy by prompting users to adjust their voice levels or reduce ambient noise.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If the sensitivity for collecting audio signal is increased, then the speech recognition capability is improved, but the ambient noise interference is also increased
Solution Approach 1:
The patent applies dynamics by making the voltage amplitude adjustable rather than fixed. The collecting circuit dynamically changes its collection sensitivity based on environmental conditions, switching between different voltage amplitudes to optimize between capturing user speech and rejecting ambient noise.
Solution Approach 2:
The patent changes the physical parameter of voltage amplitude used for collecting audio data. By adjusting this parameter, the system can control the sensitivity of the collecting circuit, thereby resolving the contradiction between capturing weak user speech signals and rejecting strong ambient noise signals.
2Power
If the voltage amplitude for collecting audio data is increased, then the audio signal strength is improved, but the ambient noise collection is also increased
Solution Approach 1:
The system dynamically adjusts the voltage amplitude based on the collected audio characteristics. When ambient noise is detected, the system reduces the voltage amplitude to minimize noise collection while maintaining sufficient signal strength for user speech recognition.
Solution Approach 2:
The patent implements feedback by analyzing the collected audio data to determine whether it contains ambient noise, then using this information to adjust the collection conditions. The processor provides feedback signals to modify the voltage amplitude based on the audio environment analysis.
3Measurement precision
If the semantic content matching is performed to identify ambient noise, then the recognition accuracy is improved, but the processing time is increased
Solution Approach 1:
The patent performs partial matching by comparing semantic content against a database of known ambient noise patterns. Rather than exhaustive analysis, the system uses targeted semantic matching to efficiently identify ambient noise while maintaining reasonable processing speed.
Data Source
AI summary
The embodiments of the present disclosure provide a device for recognizing speeches and a method for speech recognition. The device for recognizing speeches may comprise a processor, configured to execute instructions stored in the memory, to: perform speech recognition on the collected audio data to obtain a semantic content of the audio data; match the obtained semantic content with a semantic data stored in the database; determine whether the audio data contains ambient noise audio information and audio information of a user, in response to determining that the obtained semantic content does not match with the semantic data; and change conditions for collecting the audio data and control to collect the audio data with the changed conditions, in response to determining that the audio data contains the ambient noise audio information and the audio information of the user.


