Television Voice Control Delay Compensation
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Voice-controlled devices face delays in executing commands due to the time required for voice recognition, leading to inaccuracies in timing-based operations, such as skipping sections during video playback.
Innovation Solution
A television receiving apparatus with a control unit that detects voice inputs, processes them through voice recognition, and generates control signals by accounting for the delay between voice input and recognition, allowing for precise timing adjustments in commands like 'skip 30 seconds' by subtracting the recognition delay from the intended action.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If voice recognition processing is performed on voice commands, then the device can be controlled by voice input, but control delay occurs between voice input and command execution
Solution Approach 1:
The system performs preliminary actions by detecting the voice period and determining the first time point (start or end of voice input) before voice recognition processing completes. This allows the system to prepare timing information in advance, so when the command is executed after recognition, the timing is already calculated and ready, reducing the perceived delay.
Solution Approach 2:
The system dynamically adjusts the timing of control signals based on the actual voice period detected. Instead of using a fixed delay, the control unit calculates the delay based on the actual duration and characteristics of the voice input, making the timing adaptive and optimized for each specific voice command scenario.
2Loss of time
If voice recognition processing time is reduced, then control delay is minimized, but voice recognition accuracy may deteriorate
Solution Approach 1:
The system extracts only the essential timing information (first time point of voice period) from the voice input signal before sending it for full recognition processing. This separation allows the timing-critical path to be fast while the recognition processing can proceed independently, potentially in parallel, without blocking the timing determination.
Solution Approach 2:
The control unit acts as an intermediary that receives the voice period information and the recognition result, then synthesizes the final control signal with appropriate timing. This intermediary processing allows the system to coordinate between fast timing requirements and thorough recognition processing without forcing one to wait for the other completely.
Data Source
AI summary
Provided are a television receiving apparatus and a voice signal processing method. The television receiving apparatus includes: a broadcast signal receiving and processing unit configured to process broadcast signals according to broadcast standards; a communication unit configured to connect with a network and communicate with one or more servers and one or more external devices; monitor configured to display an image; speaker configured to output voice; microphone configured to receive a voice input; interface unit configured to receive a command signal from outside or output a signal to an external device; control unit in connection with the interface unit, the communication unit, the monitor, the speaker, the broadcast signal receiving and processing unit and configured to generate control signal for a target controlled object based on voice input from outside and send the control signal to the target controlled object to implement a control operation corresponding to the voice input.


