Voice Font Association for Text-to-Speech Caller Identification
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current portable electronic devices provide limited information through audible outputs, such as ring tones, while lacking caller or email originator identification, and text-to-speech conversion increases bandwidth and costs.
Innovation Solution
The device associates a voice font with contact records, enabling text-to-speech conversion to announce communications in the voice of the originator, reducing data transmission and providing identification through voice output.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Loss of information
If text-to-speech conversion is used to provide audible output, then information accessibility is improved, but data transmission bandwidth and costs increase
Solution Approach 1:
The system pre-loads voice font data into the device's memory during idle periods or when connected to a network, so that text-to-speech conversion can be performed locally without requiring real-time data transmission. This preliminary action ensures information accessibility is improved while avoiding increased bandwidth requirements during actual use.
Solution Approach 2:
The system creates local copies of voice font data and stores them in the device's memory or cache. These copied voice data can be reused multiple times for different text-to-speech conversions without requiring additional data transmission, thus improving information accessibility while maintaining constant bandwidth usage.
2Device complexity
If audible notifications are provided through simple ring tones, then device complexity is reduced, but information content is limited
Solution Approach 1:
The notification system is designed to handle multiple types of information delivery through a single unified framework. It can output simple ring tones for basic notifications, text-to-speech conversions for detailed information, and caller identification for contextual awareness. This multi-functional approach increases information content while keeping the overall system architecture relatively simple through standardized processing pathways.
Solution Approach 2:
The notification system dynamically adjusts its output based on the type of communication received. For simple notifications, it uses basic ring tones; for emails and messages, it provides text-to-speech conversion with caller identification. This dynamic adaptation allows the system to provide appropriate information content for each situation without requiring permanently complex functionality for all cases.
3Measurement precision
If caller identification information is transmitted separately, then identification accuracy is improved, but data transmission time increases
Solution Approach 1:
The system merges caller identification data with the main communication data stream, processing and displaying identification information simultaneously with the receipt of the actual message or call data. This integration allows accurate identification to be provided without requiring separate transmission time, as the identification data is extracted and processed in parallel with the main data flow.
Solution Approach 2:
The system performs preliminary processing of incoming data streams to extract and identify caller information before the main communication content is fully received or processed. This preliminary identification action ensures accurate caller recognition is achieved without adding to the overall transmission time, as the identification extraction occurs during the initial phase of data reception.
Data Source
AI summary
A method of associating a voice font with a contact for text-to-speech conversion at an electronic device includes obtaining, at the electronic device, the voice font for the contact, and storing the voice font in association with a contact data record stored in a contacts database at the electronic device. The contact data record includes contact data for the contact.


