Display Voice Recognition with Dynamic Audio Volume Control
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Image display apparatuses face decreased voice recognition performance due to background noise interference, as they receive and process user voice commands alongside audio signals from the device, leading to reduced accuracy in command recognition.
Innovation Solution
The implementation of an image display apparatus with a voice recognition system that includes a voice inputter, audio outputter, and a controller to decrease the audio signal volume to a predetermined level when a voice recognition start command is detected, and optionally incorporates a background sound canceller to isolate user voice signals, thereby enhancing recognition accuracy.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If the image display apparatus outputs audio signals at normal volume, then the audio output quality is maintained, but the voice recognition accuracy deteriorates due to background noise interference
Solution Approach 1:
The system performs preliminary actions by detecting the voice recognition start command word before actual voice recognition begins. Upon detection, it proactively decreases the audio output volume to a predetermined level in advance, creating favorable conditions for accurate voice recognition by reducing background noise interference before the recognition process starts
Solution Approach 2:
The audio output volume is made dynamic rather than static. The controller automatically adjusts the volume level based on the recognition state: decreasing to a predetermined level when voice recognition is active and restoring to normal level when recognition is complete. This dynamic adjustment resolves the contradiction between maintaining audio quality and ensuring recognition accuracy
2Measurement precision
If the audio signal volume is decreased during voice recognition, then the voice recognition accuracy improves, but the audio output quality deteriorates
Solution Approach 1:
The audio volume adjustment operates periodically based on voice recognition cycles. The controller decreases volume when a voice recognition start command is detected and restores it after recognition completes (when no control command is received for a predetermined time or an end command word is detected). This periodic adjustment ensures high recognition accuracy during active recognition while maintaining normal audio quality during non-recognition periods
Solution Approach 2:
The system implements dynamic volume control that adapts to different operational states. Rather than maintaining a fixed low volume, the audio output dynamically switches between normal volume (when recognition is not active) and reduced volume (when recognition is active), thereby preserving audio output quality while ensuring recognition accuracy when needed
3Speed
If the image display apparatus continuously monitors for voice commands, then the responsiveness to user commands improves, but the system complexity increases due to continuous background sound processing
Solution Approach 1:
The system extracts and isolates specific voice recognition command words from the continuous audio stream for detection. Rather than processing all audio data continuously, the controller focuses on detecting predetermined command words that trigger voice recognition mode. This extraction approach improves responsiveness to commands while reducing the complexity of continuous background sound processing by targeting specific recognition triggers
Data Source
AI summary
Provided are an image display apparatus and a method of controlling the same. The image display apparatus enabling voice recognition includes: a first voice inputter which receives a user-side audio signal; an audio outputter which outputs an audio signal processed by the image display apparatus; a first voice recognizer which recognizes the user-side audio signal received through the first voice inputter; and a controller which decreases a volume of the audio signal output through the audio outputter to a predetermined level if a voice recognition start command is received.


