Speech Recognition Display Mode Adjustment for Status Visibility
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Users find it difficult to intuitively determine the status of speech recognition processing solely by visually examining the results, as the existing techniques do not provide clear indicators of the processing status.
Innovation Solution
An information processing device and method that acquires parameters related to speech recognition processing, such as utterance volume and noise levels, to dynamically adjust the display mode of the recognition results, including text size, blurring, and added objects, allowing users to visually understand the processing status.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Loss of information
If speech recognition processing is performed on sound information, then speech recognition results are obtained, but it is difficult for users to intuitively understand the processing status
Solution Approach 1:
The patent applies color changes to the display of speech recognition results to indicate different processing statuses. When speech recognition is performed, the text color changes to visually distinguish the processing state, allowing users to intuitively understand whether recognition is occurring without additional complexity to the system.
Solution Approach 2:
The patent adds a visual dimension to speech recognition feedback by displaying text information in a different dimensional space (visual display) rather than only auditory output. This allows users to simultaneously see the recognition results and processing status, enhancing intuitive understanding without adding physical complexity.
2Loss of information
If display mode is changed based on parameters, then user understanding of processing status is improved, but device complexity increases
Solution Approach 1:
The patent changes display parameters (such as text color, size, or style) based on speech recognition parameters (such as confidence level or processing state). This allows the system to convey processing status information through simple parameter adjustments rather than complex structural changes, maintaining device simplicity while improving information delivery.
Solution Approach 2:
The patent makes the display unit serve multiple functions: it displays both the speech recognition results and the processing status information simultaneously through different visual characteristics. This multi-functionality reduces the need for separate indicators or components, thereby avoiding increased device complexity while still providing comprehensive status information.
Data Source
AI summary
There is provided an information processing device technology that enables the user to know intuitively the situation in which the speech recognition processing is performed, the information processing device including: an information acquisition unit configured to acquire a parameter related to speech recognition processing on sound information based on sound collection; and an output unit configured to output display information used to display a speech recognition processing result for the sound information on the basis of a display mode specified depending on the parameter.


