Voice-Controlled Display Audio Ducking for Accurate Recognition

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Image display apparatuses face decreased voice recognition rates due to background noise interference, as they receive and process user voice commands alongside audio signals from the device, leading to misinterpretation of commands.

Innovation Solution

The implementation of an image display apparatus with a voice recognition system that includes a voice inputter, audio outputter, and a controller to reduce the audio signal volume to a predetermined level when a voice recognition start command is detected, and optionally uses a background sound canceller to isolate user voice signals, ensuring accurate command recognition.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If the image display apparatus outputs audio signals at normal volume, then the user can hear the audio content clearly, but the voice recognition rate decreases due to background noise interference

Engineering Contradiction:
Improvevoice recognition rateVSAvoidbackground noise interference
Core Design Contradiction:
ReliabilityVSObject-affected harmful factors

Solution Approach 1:

The system proactively reduces audio output volume when voice recognition is detected, preventing background noise from interfering with voice command recognition before the interference occurs. This preliminary action ensures clear voice recognition by adjusting audio levels in advance.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The audio output volume is dynamically adjusted based on the operational state of the system. When voice recognition mode is activated, the volume is automatically reduced to a predetermined level, and when normal operation resumes, the volume returns to its original state, creating a flexible adaptive system.

Inventive Principle:
Principle #15Dynamics

2Measurement precision

If the audio output volume is reduced during voice recognition, then voice recognition accuracy improves, but the user cannot hear audio content during the recognition process

Engineering Contradiction:
Improvevoice recognition accuracyVSAvoidaudio content accessibility
Core Design Contradiction:
Measurement precisionVSLoss of information

Solution Approach 1:

The audio volume is dynamically adjusted only during the specific time window when voice recognition is active, maintaining high recognition accuracy. The system automatically restores normal volume levels after recognition completes, ensuring audio content remains accessible during non-recognition periods.

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The volume reduction is applied periodically and temporarily only during voice recognition events, not continuously. This periodic adjustment ensures voice recognition accuracy is maintained when needed while preserving normal audio playback quality during regular operation.

Inventive Principle:
Principle #19Periodic action

3Speed

If the system continuously monitors for voice commands, then response time to user commands improves, but false recognition increases due to ongoing audio signal processing

Engineering Contradiction:
Improveresponse time to commandsVSAvoidfalse recognition rate
Core Design Contradiction:
SpeedVSReliability

Solution Approach 1:

The system prepares for voice recognition by reducing audio output volume in advance when a voice input is detected, creating optimal conditions for accurate recognition. This preliminary preparation prevents false recognition by establishing a low-noise environment before the recognition process begins.

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS11763812B2Image display apparatus and method of controlling the same
Publication Date: 2023.09.19 SAMSUNG ELECTRONICS CO LTD
  • US11763812B2 patent drawing
  • US11763812B2 patent drawing
  • US11763812B2 patent drawing

AI summary

Provided are an image display apparatus and a method of controlling the same. The image display apparatus enabling voice recognition includes: a first voice inputter which receives a user-side audio signal; an audio outputter which outputs an audio signal processed by the image display apparatus; a first voice recognizer which recognizes the user-side audio signal received through the first voice inputter; and a controller which decreases a volume of the audio signal output through the audio outputter to a predetermined level if a voice recognition start command is received.