Instant Messaging Sound Change Processing for Bandwidth Reduction
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current instant messaging methods lack expressive forms, leading to misunderstandings, and existing methods like text, voice, and video messaging face limitations such as simplex expression, inconvenient operation, and high traffic usage.
Innovation Solution
An instant messaging method that involves sound change processing using libraries like Soundtouch to create personalized sounds, which are then synthesized with pre-stored animations on a terminal to form analog image data, reducing network traffic and improving communication efficiency.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Loss of information
If video chatting is used to truly present videos of both chatting parties, then the expressiveness of communication is improved, but network bandwidth occupation and traffic costs increase significantly
Solution Approach 1:
The patent segments the video communication into two parts: pre-stored animation segments (prepared in advance) and real-time sound transmission (processed and sent only when needed). This segmentation allows the system to convey expressive visual information without continuously transmitting full video data, thereby reducing network bandwidth occupation while maintaining communication expressiveness.
Solution Approach 2:
The patent applies preliminary action by pre-storing animation segments locally on the terminal before communication occurs. These animations are prepared in advance and can be quickly retrieved and combined with real-time sound during communication, eliminating the need to transmit large video files in real-time and significantly reducing traffic costs.
2Loss of information
If picture or emoticon is used to enrich expression of user emotion, then the expressiveness is improved, but operation convenience deteriorates due to needing to search through large numbers of options
Solution Approach 1:
The patent enables the system to automatically match and select appropriate animation segments based on the analyzed sound characteristics (such as pitch, volume, and tone) without requiring manual user selection. The system serves itself by autonomously choosing the most suitable animation to match the user's emotional state, thereby improving operation convenience while maintaining rich emotional expression.
Solution Approach 2:
The patent implements feedback by analyzing the sound recorded from the user and using this analysis to automatically select and match the appropriate animation segment. The system receives feedback from the sound characteristics and adjusts the animation selection accordingly, creating a closed-loop system that enhances emotional expression without requiring manual intervention from the user.
3Ease of operation
If text is used as the most widely used chatting manner, then ease of operation is improved, but expressiveness deteriorates as it is hard to express real feeling and mood of a user
Solution Approach 1:
The patent merges two different communication modalities: sound (which carries emotional information) and animation (which provides visual expression). By combining these elements, the system creates a communication form that retains the ease of operation of text-based messaging while adding the emotional expressiveness of voice and visual animation, thereby resolving the contradiction between operational simplicity and emotional conveyance.
Data Source
AI summary
The present disclosure provides an instant messaging method and system, a communication information processing method, a terminal, and a storage medium. The instant messaging method includes: receiving, by a first terminal, a sound recorded by a user, and performing sound change processing on the sound recorded by the user to provide a changed sound; sending, by the first terminal, the changed sound after the sound change processing to a second terminal for the second terminal to synthesize the received, changed sound that has gone through the sound change processing with a pre-stored animation, so as to form analog image data and to play the analog image data. The present disclosure has advantages including rich communication forms, convenient operation, and high network transmission efficiency.


