Voicemail Interaction via Audio Feature Recognition
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Traditional IVR voicemail systems require users to constantly move their device between their ear and visual space, leading to cumbersome interactions and reliance on hearing audio prompts, which can be frustrating and error-prone, especially in noisy environments.
Innovation Solution
Facilitating voicemail interactions by receiving audio segments from IVR systems and displaying visual information on user devices, such as prompts, notifications, and confirmations, using speech recognition, human-imperceptible affixed information, and embedded information to reduce the need for auditory focus and prevent errors.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If users interact with traditional IVR voicemail systems using only audio prompts, then the system remains simple and compatible with basic telephony, but user experience deteriorates due to the need to constantly move the device between ear and visual space
Solution Approach 1:
The patent introduces visual display dimension to the traditionally audio-only IVR voicemail system. The user device displays visual representations of audio prompts, notifications, and confirmations, allowing users to view information visually while keeping the device in a single position, thereby eliminating the need to constantly move the device between ear and visual space.
Solution Approach 2:
The patent introduces an intermediary processing layer that converts audio segment features into visual display information. The user device receives audio segments from the IVR system, performs feature recognition, and generates corresponding visual displays, acting as a mediator between the audio-based voicemail system and the user's visual perception.
2Reliability
If users rely on auditory cues in noisy environments, then the system maintains audio-only simplicity, but reliability deteriorates due to difficulty in hearing prompts accurately
Solution Approach 1:
The patent implements visual feedback by displaying recognized features and system responses on the user device screen. Users can visually confirm what the system has heard and what actions are available, providing a feedback loop that compensates for audio masking in noisy environments and reduces interaction errors.
Solution Approach 2:
The patent changes the information presentation parameter from audio-only to visual display. By converting audio prompt information into visual form, the system alters the perception channel, making interaction reliable regardless of audio environment quality.
3Ease of operation
If the system processes and displays visual information from audio segments, then user experience improves through continuous visual feedback, but device complexity increases due to speech recognition and feature analysis capabilities
Solution Approach 1:
The patent leverages the user device's existing multi-functionality, particularly its audio processing and display capabilities, to perform feature recognition and visual presentation. Rather than adding dedicated specialized hardware, the system utilizes the device's universal processing abilities to handle both audio reception and visual display functions.
Solution Approach 2:
The user device performs self-service by autonomously analyzing audio segments, recognizing features, and generating appropriate visual displays without requiring external assistance. The device uses its own processing power and existing components to provide the enhanced visual interface functionality.
Data Source
AI summary
Example methods and apparatus to facilitate voicemail interaction are disclosed. A disclosed example method involves, during a call session with a voicemail system, receiving an audio segment from the voicemail system. The example method also involves performing feature recognition on the audio segment and outputting a display element to a user interface based on a recognized feature in the audio segment.


