Voice Recognition Gain Control for Variable Speaking Distance
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional electronic apparatuses with voice recognition functions often misrecognize voice commands due to inadequate consideration of the distance between the microphone and the user, as well as the intensity of the uttered voice, leading to inconvenient re-utterance of commands.
Innovation Solution
An electronic apparatus that adjusts the intensity of sound signals based on the identified gain value to ensure they fall within a predetermined range, using a processor to recognize trigger words and adjust microphone settings accordingly, thereby improving voice recognition accuracy.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If the electronic apparatus performs voice recognition without adjusting sound signal intensity, then the device complexity is reduced, but the voice recognition accuracy deteriorates
Solution Approach 1:
The system performs preliminary analysis of the first sound signal to identify gain values before processing the second sound signal. This preliminary action prepares the appropriate intensity adjustment in advance, ensuring accurate voice recognition when the voice command is received, while avoiding real-time complexity during critical recognition moments.
Solution Approach 2:
The patent replaces manual user adjustment (mechanical interaction) with automatic electronic gain adjustment. The processor automatically identifies and applies appropriate gain values based on sound signal analysis, substituting the need for users to physically adjust microphone settings or re-utter commands, thereby improving recognition accuracy without significant complexity increase.
2Measurement precision
If the electronic apparatus adjusts sound signal intensity based on gain values, then the voice recognition accuracy is improved, but the processing time increases
Solution Approach 1:
The system analyzes sound signals and identifies appropriate gain values in advance (preliminary action) before the actual voice command recognition occurs. This preparation ensures that when the second sound signal arrives, the correct intensity adjustment is already determined, reducing real-time processing delays while maintaining high recognition accuracy.
3Adaptability or versatility
If the electronic apparatus uses fixed microphone settings, then the ease of operation is maintained, but the adaptability to different voice intensities deteriorates
Solution Approach 1:
The system performs self-service by automatically analyzing incoming sound signals and identifying appropriate gain values without user intervention. The processor autonomously adjusts sound signal intensity based on the identified gain values, enabling the device to adapt to different voice intensities and distances while maintaining ease of operation, as users need not manually configure settings.
Solution Approach 2:
The patent dynamically changes the gain parameter of the sound signal processing based on analyzed characteristics. By adjusting the gain value according to the identified sound intensity and distance, the system adapts to varying user conditions (different voice intensities, distances from microphone) while maintaining simple user interaction, thus improving adaptability without compromising ease of operation.
Data Source
AI summary
An electronic apparatus is provided. The electronic apparatus includes a microphone, a communication interface including circuitry and a processor configured to, based on identifying that a trigger word is included in a first sound signal received through the microphone, enter a voice recognition mode, identify a gain value for adjusting an intensity of the first sound signal to be in a predetermined intensity range based on the intensity of the first sound, adjust an intensity of a second sound signal received through the microphone in the voice recognition mode based on the identified gain value, and control the communication interface to transmit a user command obtained based on voice recognition regarding the adjusted second sound signal, to an external apparatus.


