Display Audio Attenuation During Voice Command Recognition
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Image display apparatuses face decreased user speech recognition rates due to background noise interference from their own audio signal output, which exceeds predetermined levels, leading to malfunction and errors in command execution.
Innovation Solution
The apparatus includes a controller that activates speech recognition by reducing the audio signal output to a predetermined level when a speech recognition start command is received and restores the original volume if no input is detected within a predetermined time, improving recognition performance and preventing errors.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If the image display apparatus outputs audio signal at normal volume, then the user can enjoy audio content, but the speech recognition rate decreases due to background noise interference
Solution Approach 1:
The system proactively reduces audio output volume when speech recognition is initiated, before the recognition process completes. This preliminary action prevents background noise from interfering with the speech recognition process, ensuring accurate command detection while maintaining normal audio playback for the user.
Solution Approach 2:
The audio output volume is dynamically adjusted based on the operational state of the speech recognition system. During speech recognition operations, the volume is automatically reduced to minimize interference, and restored to normal levels when recognition is complete, creating an adaptive system that responds to real-time conditions.
2Reliability
If the audio signal output volume is reduced during speech recognition, then speech recognition performance improves, but the user cannot hear audio content simultaneously
Solution Approach 1:
The audio volume reduction is applied periodically and temporarily only during the speech recognition window. The system reduces volume when speech input is detected, maintains this reduced state only for the duration needed to process the command, then restores normal volume. This periodic adjustment ensures speech recognition accuracy while minimizing impact on overall audio content accessibility.
3Device complexity
If the speech recognition operates without volume control, then the system structure remains simple, but speech recognition errors increase due to audio interference
Solution Approach 1:
The existing audio output unit is made multi-functional by enabling it to operate in two distinct modes: normal audio playback mode and speech recognition support mode. The same hardware component (audio output unit) performs both functions, eliminating the need for separate volume control hardware while improving speech recognition accuracy through intelligent software-based volume management.
Data Source
Figure 1
Figure 2
Figure 3
AI summary
Provided are an image display apparatus (1) able to improve voice recognition performance by decreasing the volume of an audio signal output from the image display apparatus (1) to a predetermined level or less when the image display apparatus (1) recognizes user voice, and a method of controlling the same. The image display apparatus (1) enabling voice recognition includes a first voice input unit (112) to receive a user-side audio signal, an audio output unit (122) to output an audio signal processed by the image display apparatus (1), a first voice recognizer (210) to recognize the user-side audio signal received through the first voice input unit (112), and a controller (220) to decrease the volume of the audio signal output through the audio output unit (122) to a predetermined level if a voice recognition start command is received.