Information Processing for Attention-Based Remote Dialogue Audio
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional dialogue systems fail to ensure that addressing voice reaches the intended user, reducing the convenience of remote communication by outputting the voice at the same volume to all users, which may not be optimal for the intended recipient.
Innovation Solution
An information processing apparatus estimates the parameter of each interlocutor based on video analysis, displaying alerts for users with lower parameters, and adjusts audio volume accordingly, transmitting high volume to the focused interlocutor and low volume to others.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If the addressing voice is outputted to other users at the same volume as to the predetermined user, then the audio transmission is simple and uniform, but the addressing voice might not reach the predetermined user effectively
Solution Approach 1:
The patent applies local quality by assigning different audio volumes to different users based on their roles and positions in the dialogue. The predetermined user (e.g., teacher) receives audio at a different volume than other users (e.g., students), ensuring that the addressing voice reaches the intended recipient effectively. This is achieved through the volume setting unit that configures specific volume levels for each user based on their user type.
Solution Approach 2:
The patent changes the audio volume parameter dynamically based on the user type and dialogue context. The volume setting unit adjusts the audio volume parameter for each user individually, allowing the system to optimize voice delivery reliability for different user groups without requiring complex hardware modifications.
2Ease of operation
If the audio volume is adjusted differently for different users, then the convenience of dialogue service is improved, but the audio processing complexity increases
Solution Approach 1:
The patent applies preliminary action by pre-configuring volume settings for each user type before the dialogue begins. The volume setting unit establishes the audio volume parameters in advance based on user roles (e.g., teacher vs. student), eliminating the need for real-time complex calculations during the dialogue and simplifying the processing burden during actual use.
Solution Approach 2:
The system automatically assigns and manages different audio volumes for different users without requiring manual intervention during the dialogue. The volume setting unit self-manages the audio processing by automatically applying the appropriate volume levels based on user type, reducing the operational complexity for users while improving service convenience.
Data Source
AI summary
An information processing apparatus supports dialogue using dialogue apparatuses used by a plurality of users located remotely from each other. The information processing apparatus includes a controller configured to estimate, during execution of a remote dialogue, a parameter of each interlocutor with respect to a topic of the remote dialogue based on a video of each interlocutor, display an alert, on a first dialogue apparatus of a first user, indicating each interlocutor whose estimated parameter is less than a threshold, and when the first user gazes at and speaks to an interlocutor displayed with the alert among a plurality of interlocutors as second users displayed on a screen of the first dialogue apparatus, transmit first audio data with a high volume to a second dialogue apparatus of the interlocutor and transmit second audio data with a low volume to second dialogue apparatus of interlocutors other than the interlocutor.


