Voice Message Component Interaction System
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
In existing voice messaging systems, recipients cannot interact independently with voice components or access information about each sender within a compounded voice message, especially on terminals with limited capabilities, leading to cumbersome user experiences.
Innovation Solution
A method and device that allow users to interact with voice messages by detecting user interactions during playback, sending signals for actions related to voice components, and providing associated information, enabling independent interaction and information access without requiring end-of-message wait or lengthy vocal synthesis.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If voice components are played as a single audio sequence, then the voice message can be delivered simply, but the user cannot interact independently with each voice component or access information about each sender
Solution Approach 1:
The voice message is segmented into multiple independent voice components, each associated with metadata about senders and content. This segmentation enables the user to interact with each component independently rather than treating the message as a single unified stream, directly resolving the contradiction by maintaining simplicity while enabling sophisticated interaction.
Solution Approach 2:
An intermediary system is introduced between the voice message delivery and user interaction layers. This intermediary manages the association between voice components and sender information, facilitating independent interaction without requiring the user to manually navigate complex structures, thus improving ease of operation while managing system complexity.
2Ease of operation
If the user waits for the end of the message to interact, then the system remains simple, but the user experience becomes cumbersome and time-consuming
Solution Approach 1:
The system performs preliminary action by delivering voice components with associated metadata before the user needs to interact with them. The voice message is played with embedded information about senders and components, allowing the user to access information immediately during playback rather than waiting for the message to end, thus reducing time loss while maintaining simple interaction.
Solution Approach 2:
The useful action of information delivery continues during the voice message playback itself, rather than occurring only after playback completes. By continuously providing access to sender information and voice component details during the message delivery, the system eliminates the need for post-message interaction wait time, improving both ease of operation and time efficiency.
3Loss of information
If information is synthesized vocally and inserted into the message, then the user receives information, but the message becomes lengthy and irksome
Solution Approach 1:
The system extracts the information delivery function from the voice message content itself. Instead of embedding information within the message audio stream, the system separates information about senders and components into independent metadata associated with each voice component. This allows users to access information on demand without it being embedded in the message playback, thus avoiding message lengthening while ensuring information accessibility.
Solution Approach 2:
The system creates a copy of the voice message with embedded metadata about senders and components. This copied version contains the same audio content plus associated information structures, allowing the user to access information about each component without the information itself being vocalized or inserted into the original message, thereby avoiding unnecessary lengthening.
Data Source
AI summary
A method for consulting a voice message received and a method for providing a compounded voice message. The compounded voice message is composed of at least one first and one second voice component and is associated with a group of items of information relating to the voice components. The method for consultation includes reading at least one voice component of the voice message, detecting at least one user interaction concomitant upon the reading of the at least one voice component, sending at least one signal relating to the interaction detected, and receiving a command of an action to be performed relating to a voice component read from the voice message and to at least one item of information of the group of items of information associated with the voice message. Also provided is a device implementing the method for consulting a compounded voice message.


