Multimodal Conversation Parking and Retrieval
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional communication systems are limited in handling multimodal conversations, particularly in parking and retrieving such conversations across various modalities, which hinders efficient communication and user experience in modern unified communication systems.
Innovation Solution
The system enables subscribers to park established multimodal conversations across multiple modalities (audio, video, presentations, etc.) and notify other subscribers for retrieval, allowing seamless continuation of conversations using various notification methods, with content playback during the waiting period.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If conventional call parking is used in traditional communication systems, then audio calls can be transferred and held, but the system is restricted to single modality (audio only) and cannot handle multimodal conversations
Solution Approach 1:
The patent extends the traditional call parking function to support multiple communication modalities (audio, video, data sharing, whiteboarding, instant messaging) within a unified parking mechanism. The system allows subscribers to park conversations that include various modalities simultaneously, making the parking function universal across different communication types rather than limited to audio only.
Solution Approach 2:
The patent segments the multimodal conversation into distinct modality components (audio stream, video stream, data sharing session, whiteboarding session) that can be independently managed, parked, and retrieved. Each modality can be handled separately while maintaining the overall conversation context, allowing flexible recombination during retrieval.
2Adaptability or versatility
If traditional call parking is implemented, then single modality calls can be held, but retrieval and continuation across multiple modalities and devices is not supported
Solution Approach 1:
The patent creates a copy of the conversation state and context information when parking occurs. This includes capturing the current state of all modalities (audio, video, data sharing, whiteboarding) and storing it alongside the parked conversation. When retrieving, this copied context information is restored to enable seamless continuation of the multimodal conversation across different devices and modalities.
Solution Approach 2:
The system introduces an intermediary parking mechanism that acts as a mediator between the original conversation and the retrieval process. This intermediary stores and manages the multimodal conversation state, enabling the conversation to be transferred across different devices and modalities while preserving context. The intermediary handles the complexity of multimodal state management, allowing seamless continuation without direct peer-to-peer coordination.
3Ease of operation
If notification methods are limited in traditional systems, then simple call transfers can be achieved, but flexible notification to subscribers for conversation retrieval is not possible
Solution Approach 1:
The patent implements dynamic notification methods that can adapt to the subscriber's current state and preferences. The system can send notifications through multiple channels (instant messaging, email, voice call, push notifications) and can dynamically select the most appropriate method based on availability information and user preferences. The notification process is flexible and can be adjusted in real-time, allowing subscribers to be notified through their preferred modality and reducing waiting time for conversation retrieval.
Data Source
Figure 1
Figure 2
Figure 3
AI summary
Established multimodal conversations are enabled to be parked within an enhanced communication system such that a subscriber of the system can be notified through a variety of means and enabled to retrieve selected or all modalities for continuing the conversation. Different modalities may be parked together or separately. While waiting for the subscriber to retrieve the conversation, a participant may receive audio, video, presentation, or other forms of content as playback.