Continuous Voice Recognition for Display Control
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing voice-controlled display apparatuses require users to repeatedly start and end voice recognition functions for each command, leading to inconvenient operation and unnecessary interface generation.
Innovation Solution
The display apparatus continuously recognizes and processes intermediate voice recognition results, allowing it to perform operations based on partial commands without needing repeated user inputs to initiate or terminate voice recognition, and displays real-time text conversion of user utterances.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If the voice recognition function is started and ended for each command using a control device, then the display apparatus can execute functions according to user voice, but the user has to repeatedly perform start and end operations which is inconvenient
Solution Approach 1:
The voice recognition function operates continuously without requiring repeated start and end operations. The system maintains an active voice recognition state that can process multiple commands sequentially, eliminating the need for users to repeatedly initiate and terminate the function for each command input.
Solution Approach 2:
The display apparatus automatically manages the voice recognition function state. When the apparatus determines that a voice command has been fully processed and executed, it automatically ends the voice recognition function without requiring user intervention through a control device, making the system self-managing.
2Productivity
If the voice recognition function operates continuously to recognize intermediate results, then multiple commands can be executed without repeated user input, but the system complexity increases
Solution Approach 1:
The system performs preliminary recognition of intermediate voice results during the ongoing voice input process. By analyzing and recognizing commands in real-time as the user speaks, the system can prepare to execute multiple commands before the user finishes speaking, improving efficiency without requiring complex post-processing.
Solution Approach 2:
The display apparatus provides visual feedback by displaying recognized text in real-time on the screen. This feedback mechanism allows the system to confirm intermediate recognition results to the user, enabling the system to manage complex continuous recognition operations while keeping the user informed and in control of the process.
Data Source
Figure 1
Figure 2
Figure 3
AI summary
A display apparatus controlled based on a user's uttered voice and a method of controlling a display apparatus based on a user's uttered voice are provided. A display apparatus includes a processor, a memory, and a display. The processor is configured to receive an uttered voice of a user, determine text corresponding to the uttered voice of the user as an intermediate recognition result, determine a command based on a result obtained by comparing the intermediate recognition result with a previous intermediate recognition result that is stored in the memory, and perform an operation according to the command.