Cross-Modality Call Subsystem for Voice-Text Conversion
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Users often face challenges in effectively communicating through mobile phones due to limitations in switching between talking-listening and typing-reading modalities, leading to missed calls and unanswered text messages, especially in situations like meetings or noisy environments.
Innovation Solution
A system and method that facilitate mixed modality interactions by allowing users to respond to voice calls and text messages using either voice or text, utilizing a cross-modality call subsystem and a telephony server to convert voice data to text and vice versa, enabling seamless communication across different modalities.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If users communicate using a single modality (voice or text), then the communication method is simple and straightforward, but communication effectiveness deteriorates when users are in environments where that modality is impractical (e.g., noisy places, meetings, driving)
Solution Approach 1:
The communication system is designed to support multiple modalities (voice and text) within a single communication framework. The server can handle both voice calls and text messages, and can convert between modalities, allowing the system to adapt to different user needs and environmental conditions while maintaining a unified communication interface
Solution Approach 2:
The communication server acts as an intermediary that receives communications in one modality and can convert them to another modality. When a user receives a voice call but prefers to respond via text (or vice versa), the server facilitates the modality conversion, enabling effective communication regardless of the user's current environmental constraints
2Adaptability or versatility
If users switch between talking-listening and typing-reading modalities, then adaptability to different environments improves, but communication efficiency deteriorates due to the complexity of managing multiple modalities and the time required to switch
Solution Approach 1:
The server automates the modality switching process by detecting user preferences and environmental factors, then automatically converting communications between voice and text modalities. This eliminates the manual effort users would otherwise spend switching between modalities while maintaining adaptability to different environments
Solution Approach 2:
The system monitors user behavior patterns and environmental context to automatically determine the most appropriate modality for communication, reducing the cognitive load and time users spend deciding which modality to use. The system serves itself by making intelligent modality selections based on accumulated data
3Ease of operation
If users cannot agree on the communication modality, then each user maintains their preferred method, but communication reliability deteriorates resulting in missed calls and unanswered messages
Solution Approach 1:
The server mediates between users with different modality preferences by automatically converting communications to the recipient's preferred modality. This allows each user to maintain their autonomy in choosing their preferred communication method while ensuring reliable delivery by adapting to the other user's preferences
Solution Approach 2:
The system dynamically changes the communication modality parameter based on user preferences and context. When User A calls User B, the server can convert the voice call to a text message if User B prefers text, or convert a text message to voice if User A prefers voice, ensuring communication delivery without requiring users to compromise their preferences
Data Source
AI summary
Methods and systems to facilitate communications between users via different modalities. A method includes identifying, by a first user device, a voice call originating from a second user device, and presenting a user interface to a user of the first user device, where the user interface provides an option to respond to the voice call by voice and an option to respond to the voice call in a text form. The method further includes detecting that the user of the first user device has selected the option to respond to the voice call in the text form, and causing a user response to the voice call to be converted into voice data for the second user device.


