Image Display Voice Control With Dynamic Audio Volume Reduction
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
The issue of decreased voice recognition performance in image display apparatuses due to background noise and audio signal interference, particularly when the volume of background sound or audio output exceeds a certain level, is addressed.
Innovation Solution
The implementation of a voice recognition system that includes a first voice inputter, an audio outputter, a first voice recognizer, and a controller to decrease the audio output volume to a predetermined level upon receiving a voice recognition start command, optionally with a background sound canceller and a command word generator to enhance recognition accuracy.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Illumination intensity
If the audio output volume is increased to provide better audio experience, then the audio quality is improved, but the voice recognition accuracy deteriorates due to background noise interference
Solution Approach 1:
The audio output volume is dynamically adjusted based on the voice recognition state. When voice recognition is activated, the volume is automatically reduced to a predetermined level to minimize interference with the microphone input. This dynamic adjustment allows the system to optimize both audio playback quality and voice recognition accuracy at different operational states
Solution Approach 2:
The system performs preliminary volume reduction before voice recognition processing begins. By detecting the voice recognition start command and proactively lowering the audio output volume in advance, the system prevents background noise from interfering with the voice input signal, thereby ensuring high recognition accuracy from the outset
2Measurement precision
If the audio output volume is reduced to improve voice recognition, then the voice recognition accuracy is improved, but the audio output quality deteriorates
Solution Approach 1:
The audio output volume is dynamically adjusted based on the voice recognition state. When voice recognition is activated, the volume is automatically reduced to a predetermined level to minimize interference with the microphone input. This dynamic adjustment allows the system to optimize both audio playback quality and voice recognition accuracy at different operational states
Solution Approach 2:
The system changes the audio output volume parameter in response to voice recognition commands. By modifying this physical parameter (volume level) based on operational context, the system resolves the contradiction between maintaining high audio quality and ensuring accurate voice recognition
3Measurement precision
If background sound cancellation is applied to improve voice recognition, then the voice recognition accuracy is improved, but the device complexity increases
Solution Approach 1:
The background sound canceller extracts and removes background noise components from the audio signal while preserving the voice input signal. By separating and eliminating the harmful background sound elements, the system improves voice recognition accuracy without requiring complete system redesign
Solution Approach 2:
The background sound canceller acts as an intermediary processing unit between the microphone input and the voice recognizer. This intermediate component filters out background noise before the signal reaches the recognition engine, thereby improving accuracy while isolating the complexity to a dedicated module
Data Source
AI summary
Provided are an image display apparatus and a method of controlling the same. The image display apparatus enabling voice recognition includes: a first voice inputter which receives a user-side audio signal; an audio outputter which outputs an audio signal processed by the image display apparatus; a first voice recognizer which recognizes the user-side audio signal received through the first voice inputter; and a controller which decreases a volume of the audio signal output through the audio outputter to a predetermined level if a voice recognition start command is received.


