Information Processing System for Animated Message Reproduction
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current social network services (SNS) used through portable information terminals lack diversity in expression and engagement, limiting effective communication between users, as messages are primarily displayed and not effectively analyzed or reproduced for enhanced interaction.
Innovation Solution
An information processing system that allows multiple terminals to communicate through a server, featuring message acceptance and transmission units, a representation output unit for chronological display, and a reproduction output unit for audio and animated representation of messages based on content analysis, enabling enhanced user interaction.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If messages are merely displayed as text and images in chronological order, then the system maintains simple display structure, but user engagement and communication diversity are limited
Solution Approach 1:
The patent introduces a server as an intermediary that performs speech-to-text conversion and message analysis. The server receives audio messages from users, converts them to text, analyzes the content and speaker identity, and returns structured data to the terminal. This mediator approach enables complex processing (audio analysis, speaker identification) without increasing the complexity of the terminal device, while significantly enhancing communication expression diversity through animated avatars and audio playback.
Solution Approach 2:
The patent replaces traditional text-based interaction mechanisms with audio-based interaction. Instead of users typing and reading text messages, the system allows users to speak messages that are converted to text and paired with animated avatar representations. This substitution of mechanical typing/reading with audio speaking/listening enhances engagement and expression diversity while the server handles the conversion complexity.
2Ease of operation
If the system provides audio output and animated representation of messages, then user engagement is enhanced, but processing complexity and resource consumption increase
Solution Approach 1:
The server acts as an intermediary that performs speech-to-text conversion and message analysis. The server receives audio messages from users, converts them to text, analyzes the content and speaker identity, and returns structured data to the terminal. This mediator approach enables complex processing (audio analysis, speaker identification) without increasing the complexity of the terminal device, while significantly enhancing communication expression diversity through animated avatars and audio playback.
Solution Approach 2:
The patent extracts the complex processing functions (audio recognition, speaker identification, message analysis) from the terminal device and places them on the server. The terminal only needs to handle simple tasks like playing audio and displaying animated avatars based on data received from the server. This extraction of complex functions reduces terminal device complexity while maintaining enhanced user interaction capabilities.
3Measurement precision
If speech-to-text conversion and speaker identification are performed, then message analysis precision is improved, but processing time and computational resources increase
Solution Approach 1:
The system performs speech-to-text conversion and speaker identification as preliminary actions when the audio message is first received and uploaded to the server, before the message is displayed to other users. This preliminary processing ensures that when the message is retrieved and displayed, the analysis is already complete, minimizing the perceived processing time for end users while maintaining high precision in message analysis.
Solution Approach 2:
The server acts as an intermediary that performs speech-to-text conversion and message analysis. The server receives audio messages from users, converts them to text, analyzes the content and speaker identity, and returns structured data to the terminal. This mediator approach enables complex processing (audio analysis, speaker identification) without increasing the complexity of the terminal device, while significantly enhancing communication expression diversity through animated avatars and audio playback.
Data Source
AI summary
A first information processing apparatus includes a first message acceptance unit which accepts input of a first message and a first message transmission unit which transmits the accepted first message and first character information to a server. A second information processing apparatus includes a second message acceptance unit which accepts input of a second message and a second message transmission unit which transmits the accepted second message and second character information to the server. The first information processing apparatus further includes a representation output unit which has a display unit display in chronological order, the first message brought in correspondence with a first character based on the first character information and the second message brought in correspondence with a second character based on the second character information obtained through the server and a reproduction output unit which provides audio output of the first message and the second message.


