Voice Profile Server Adaptor for Cross-Device Compatibility

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing voice recognition systems require large storage capacity and powerful CPUs for voice training, with profiles being device-specific and incompatible with other systems, and devices with small screens face difficulties in displaying training text effectively.

Innovation Solution

A system that creates and manages user voice profiles in a common repository, using a training server and adaptor to convert audio and textual information into formats compatible with multiple voice servers, allowing for remote storage and use across various devices, with features like automatic text formatting and notification of profile updates.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If voice training is performed locally on the device, then the voice profile is device-specific and compatible with that device, but the device requires large storage capacity and powerful CPU

Engineering Contradiction:
Improvevoice profile compatibilityVSAvoidstorage capacity and CPU power
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

The patent extracts the voice profile storage and processing functions from the client device to a remote server. The client device only needs to communicate with the server, while the server handles the storage and conversion of voice profiles. This eliminates the need for large local storage and powerful CPU on the client device.

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The patent introduces a training server as an intermediary between the client device and the voice recognition system. The server acts as a mediator that receives voice profiles from clients, converts them to appropriate formats, and stores them. This intermediary approach allows simple client devices to work with multiple voice recognition systems without requiring complex local processing.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Adaptability or versatility

If voice profile is stored locally on the device, then it is compatible with that specific device, but it cannot be used with other devices or voice servers

Engineering Contradiction:
Improvevoice profile compatibility across devicesVSAvoiddevice-specific functionality
Core Design Contradiction:
Adaptability or versatilityVSReliability

Solution Approach 1:

The patent creates a universal voice profile storage system where a single voice profile stored on the server can serve multiple client devices and different voice recognition systems. The training server adaptor converts the stored profile into formats compatible with different voice servers, enabling one profile to work across multiple platforms and devices.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The patent changes the storage location parameter from local device storage to remote server storage. This parameter change enables the voice profile to be accessed by multiple devices and converted to different formats as needed, thereby improving compatibility across different systems while maintaining the reliability of device-specific functionality through format conversion.

Inventive Principle:
Principle #35Parameter changes

3Reliability

If voice profile is created for a specific voice server, then it works with that server, but a new profile must be created when the server changes

Engineering Contradiction:
Improvevoice server compatibilityVSAvoidtime to create new voice profile
Core Design Contradiction:
ReliabilityVSLoss of time

Solution Approach 1:

The patent performs preliminary action by pre-converting and pre-storing voice profiles on the server in a standardized format. When a client needs to use a voice profile with a specific voice server, the system quickly converts the pre-stored profile to the required format. This eliminates the need to create new profiles from scratch when servers change, significantly reducing the time loss.

Inventive Principle:
Principle #10Preliminary action

4Ease of operation

If training text is displayed on small screen devices, then the device can be used for voice training, but the user must constantly scroll to read the text

Engineering Contradiction:
Improveaccessibility of voice trainingVSAvoidtime to read training text
Core Design Contradiction:
Ease of operationVSLoss of time

Solution Approach 1:

The patent addresses the small screen display problem by implementing adaptive text formatting that adjusts the presentation of training text based on the device's display characteristics. The system can format text in a way that optimizes readability on small screens, potentially using techniques like progressive disclosure, collapsible sections, or audio assistance to reduce the time users spend scrolling and reading.

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

Data Source

PatentUS7440894B2Method and system for creation of voice training profiles with multiple methods with uniform server mechanism using heterogeneous devices
Publication Date: 2008.10.21 MICROSOFT TECHNOLOGY LICENSING LLC
  • US7440894B2 patent drawing
  • US7440894B2 patent drawing
  • US7440894B2 patent drawing

AI summary

A system and method for creating user voice profiles enables a user to create a single user voice profile that is compatible with one or more voice servers. Such a system includes a training server that receives audio information from a client associated with a user and stores the audio information and corresponding textual information. The system further includes a training server adaptor. The training server adaptor is configured to receive a voice profile format and a communication protocol corresponding to one of the plurality of voice servers, convert the audio information and corresponding textual information into a format compatible with the voice profile format and communication protocol corresponding to the one of the plurality of voice servers, and provide the converted audio information and corresponding textual information to the one of the plurality of voice servers.