Display Voice Recognition with Dynamic Audio Volume Control

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Image display apparatuses face decreased voice recognition performance due to background noise interference, as they receive and process user voice commands alongside audio signals from the device, leading to reduced accuracy in command recognition.

Innovation Solution

The implementation of an image display apparatus with a voice recognition system that includes a voice inputter, audio outputter, and a controller to decrease the audio signal volume to a predetermined level when a voice recognition start command is detected, and optionally incorporates a background sound canceller to isolate user voice signals, thereby enhancing recognition accuracy.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If the image display apparatus outputs audio signals at normal volume, then the audio output quality is maintained, but the voice recognition accuracy deteriorates due to background noise interference

Engineering Contradiction:
Improvevoice recognition accuracyVSAvoidbackground noise interference
Core Design Contradiction:
Measurement precisionVSObject-generated harmful factors

Solution Approach 1:

The system performs preliminary actions by detecting the voice recognition start command word before actual voice recognition begins. Upon detection, it proactively decreases the audio output volume to a predetermined level in advance, creating favorable conditions for accurate voice recognition by reducing background noise interference before the recognition process starts

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The audio output volume is made dynamic rather than static. The controller automatically adjusts the volume level based on the recognition state: decreasing to a predetermined level when voice recognition is active and restoring to normal level when recognition is complete. This dynamic adjustment resolves the contradiction between maintaining audio quality and ensuring recognition accuracy

Inventive Principle:
Principle #15Dynamics

2Measurement precision

If the audio signal volume is decreased during voice recognition, then the voice recognition accuracy improves, but the audio output quality deteriorates

Engineering Contradiction:
Improvevoice recognition accuracyVSAvoidaudio output quality
Core Design Contradiction:
Measurement precisionVSReliability

Solution Approach 1:

The audio volume adjustment operates periodically based on voice recognition cycles. The controller decreases volume when a voice recognition start command is detected and restores it after recognition completes (when no control command is received for a predetermined time or an end command word is detected). This periodic adjustment ensures high recognition accuracy during active recognition while maintaining normal audio quality during non-recognition periods

Inventive Principle:
Principle #19Periodic action

Solution Approach 2:

The system implements dynamic volume control that adapts to different operational states. Rather than maintaining a fixed low volume, the audio output dynamically switches between normal volume (when recognition is not active) and reduced volume (when recognition is active), thereby preserving audio output quality while ensuring recognition accuracy when needed

Inventive Principle:
Principle #15Dynamics

3Speed

If the image display apparatus continuously monitors for voice commands, then the responsiveness to user commands improves, but the system complexity increases due to continuous background sound processing

Engineering Contradiction:
Improvecommand response speedVSAvoidvoice recognition system complexity
Core Design Contradiction:
SpeedVSDevice complexity

Solution Approach 1:

The system extracts and isolates specific voice recognition command words from the continuous audio stream for detection. Rather than processing all audio data continuously, the controller focuses on detecting predetermined command words that trigger voice recognition mode. This extraction approach improves responsiveness to commands while reducing the complexity of continuous background sound processing by targeting specific recognition triggers

Inventive Principle:
Principle #2Taking out (Extraction)

Data Source

PatentUS9530418B2Image display apparatus and method of controlling the same
Publication Date: 2016.12.27 SAMSUNG ELECTRONICS CO LTD
  • US9530418B2 patent drawing
  • US9530418B2 patent drawing
  • US9530418B2 patent drawing

AI summary

Provided are an image display apparatus and a method of controlling the same. The image display apparatus enabling voice recognition includes: a first voice inputter which receives a user-side audio signal; an audio outputter which outputs an audio signal processed by the image display apparatus; a first voice recognizer which recognizes the user-side audio signal received through the first voice inputter; and a controller which decreases a volume of the audio signal output through the audio outputter to a predetermined level if a voice recognition start command is received.