Voice Profile Server Adaptor for Cross-Device Compatibility
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing voice recognition systems require large storage capacity and powerful CPUs for voice training, with profiles being device-specific and incompatible with other systems, and devices with small screens face difficulties in displaying training text effectively.
Innovation Solution
A system that creates and manages user voice profiles in a common repository, using a training server and adaptor to convert audio and textual information into formats compatible with multiple voice servers, allowing for remote storage and use across various devices, with features like automatic text formatting and notification of profile updates.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If voice training is performed locally on the device, then the voice profile is device-specific and compatible with that device, but the device requires large storage capacity and powerful CPU
Solution Approach 1:
The patent extracts the voice profile storage and processing functions from the client device to a remote server. The client device only needs to communicate with the server, while the server handles the storage and conversion of voice profiles. This eliminates the need for large local storage and powerful CPU on the client device.
Solution Approach 2:
The patent introduces a training server as an intermediary between the client device and the voice recognition system. The server acts as a mediator that receives voice profiles from clients, converts them to appropriate formats, and stores them. This intermediary approach allows simple client devices to work with multiple voice recognition systems without requiring complex local processing.
2Adaptability or versatility
If voice profile is stored locally on the device, then it is compatible with that specific device, but it cannot be used with other devices or voice servers
Solution Approach 1:
The patent creates a universal voice profile storage system where a single voice profile stored on the server can serve multiple client devices and different voice recognition systems. The training server adaptor converts the stored profile into formats compatible with different voice servers, enabling one profile to work across multiple platforms and devices.
Solution Approach 2:
The patent changes the storage location parameter from local device storage to remote server storage. This parameter change enables the voice profile to be accessed by multiple devices and converted to different formats as needed, thereby improving compatibility across different systems while maintaining the reliability of device-specific functionality through format conversion.
3Reliability
If voice profile is created for a specific voice server, then it works with that server, but a new profile must be created when the server changes
Solution Approach 1:
The patent performs preliminary action by pre-converting and pre-storing voice profiles on the server in a standardized format. When a client needs to use a voice profile with a specific voice server, the system quickly converts the pre-stored profile to the required format. This eliminates the need to create new profiles from scratch when servers change, significantly reducing the time loss.
4Ease of operation
If training text is displayed on small screen devices, then the device can be used for voice training, but the user must constantly scroll to read the text
Solution Approach 1:
The patent addresses the small screen display problem by implementing adaptive text formatting that adjusts the presentation of training text based on the device's display characteristics. The system can format text in a way that optimizes readability on small screens, potentially using techniques like progressive disclosure, collapsible sections, or audio assistance to reduce the time users spend scrolling and reading.
Data Source
AI summary
A system and method for creating user voice profiles enables a user to create a single user voice profile that is compatible with one or more voice servers. Such a system includes a training server that receives audio information from a client associated with a user and stores the audio information and corresponding textual information. The system further includes a training server adaptor. The training server adaptor is configured to receive a voice profile format and a communication protocol corresponding to one of the plurality of voice servers, convert the audio information and corresponding textual information into a format compatible with the voice profile format and communication protocol corresponding to the one of the plurality of voice servers, and provide the converted audio information and corresponding textual information to the one of the plurality of voice servers.


