Dynamic Text Display Control for Speech Recognition Conversation
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing speech recognition applications on smartphones are limited in displaying text conversions, hindering natural conversation due to restricted text display amounts.
Innovation Solution
An information processing device with a sound acquisition unit and a display control unit that adjusts the display of text information based on the input sound and display amounts, including features like text reduction, editing, and feedback control to facilitate natural conversation.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Loss of information
If speech recognition converts user speech to text and displays it, then communication assistance is provided, but the display amount is limited and natural conversation is hindered
Solution Approach 1:
The display control unit dynamically adjusts the display amount of text information in real-time based on the input amount of sound information and current display amount, enabling the system to adapt to varying conversation needs and maintain natural flow
Solution Approach 2:
The system implements feedback control by monitoring both the display amount of text information and the input amount of sound information, using this feedback to continuously optimize the display amount and ensure natural conversation assistance
2Loss of information
If text information is displayed for conversation assistance, then communication is supported, but excessive text display may overwhelm the user
Solution Approach 1:
The system changes the parameter of display amount based on the input amount of sound information, adjusting text display to optimal levels that prevent information overload while maintaining comprehensive communication support
3Loss of information
If the display shows all speech recognition results, then complete information is provided, but the display device capacity is exceeded
Solution Approach 1:
The display amount is dynamically controlled based on the relationship between sound input amount and current display amount, allowing the system to maximize information display within the fixed display area capacity
Solution Approach 2:
The system adjusts the display amount parameter according to the input amount of sound information, optimizing the use of display area to show comprehensive speech recognition results without exceeding device capacity
Data Source
AI summary
The present technology relates to an information processing device, an information processing method, and an information processing system that are capable of establish smooth and natural conversation with a person who has difficulty in hearing. The information processing device includes a sound acquisition unit that acquires sound information of a first user that is input to a sound input device and a display control unit that controls display of text information on a display device for a second user, the text information corresponding to the acquired sound information. The display control unit performs control related to display amount of the text information on the display device on the basis of at least one of the display amount of the text information on the display device or input amount of the sound information input through the sound input device.


