Speech Volume Feedback via Motion Object Display
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current speech recognition technologies do not effectively indicate whether a user's speech is being uttered at a volume sufficient for recognition, leading to potential misrecognition or noise interference.
Innovation Solution
An information processing device with a determination portion to assess user-uttered speech volume and a display controller that displays a moving object on a screen when the volume exceeds a recognizable threshold, allowing users to adjust their speech accordingly.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Loss of information
If speech recognition is performed without volume indication, then the system is simple, but the user cannot determine if speech is at recognizable volume
Solution Approach 1:
The patent implements a feedback mechanism by displaying a motion object that moves toward a display object when speech volume exceeds the recognizable threshold. This visual feedback provides users with real-time information about their speech volume, resolving the contradiction by adding necessary information without excessive complexity.
Solution Approach 2:
The patent uses visual changes in the display object (motion objects moving toward the display object) to indicate speech volume status. This approach provides clear volume information through visual cues rather than complex auditory or text-based feedback systems.
2Ease of operation
If no visual feedback is provided, then the display is simple, but users cannot adjust their speech volume appropriately
Solution Approach 1:
The visual feedback system displays motion objects that move toward a display object based on speech volume, enabling users to intuitively understand and adjust their speech volume. This resolves the contradiction by providing essential operational guidance through a relatively simple visual mechanism.
Solution Approach 2:
The display updates periodically or continuously based on speech input, providing dynamic visual feedback that guides users to adjust their speech volume. This periodic update mechanism delivers operational information without requiring complex continuous monitoring displays.
3Reliability
If speech volume threshold is not monitored, then processing is faster, but recognition accuracy decreases
Solution Approach 1:
The system performs preliminary volume assessment before full speech recognition processing by comparing speech volume against a predetermined threshold. This preliminary check ensures recognition accuracy by filtering out inaudible speech while maintaining processing efficiency through a simple threshold comparison rather than complex analysis.
Solution Approach 2:
The patent uses a fixed volume threshold parameter to determine whether speech is recognizable. This parameter-based approach maintains processing speed by using simple threshold comparison while ensuring reliability by establishing a clear criterion for acceptable speech volume.
Data Source
AI summary
To provide a technology capable of allowing a user to find whether speech is uttered with a volume at which speech recognition can be performed. Provided is an information processing device including: a determination portion configured to determine a user-uttered speech volume on the basis of input speech; and a display controller configured to control a display portion so that the display portion displays a display object. The display controller causes the display portion to display a first motion object moving toward the display object when the user-uttered speech volume exceeds a speech recognizable volume.


