Voice Dialer Text-to-Speech Contact Setup

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current voice dialing systems in cellular telephones are cumbersome to set up, especially for users with large contact lists, as they require manual selection and recording of contact names for speech recognition.

Innovation Solution

A system that utilizes a text-to-speech engine to generate audio files from contact information and a voice dialer for speech recognition, allowing for easier setup and operation by converting contact lists into speech and enabling voice dialing through a text-to-speech engine integrated within a telephony device or network server.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If manual selection and recording of contact names is used for voice dialing setup, then speech recognition accuracy is improved, but setup time and user convenience deteriorate

Engineering Contradiction:
Improvespeech recognition accuracyVSAvoidsetup time
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

The system performs preliminary action by automatically generating audio files for all contacts in the address book before the user needs to use voice dialing. The text-to-speech engine converts contact names to audio files and stores them for later recognition, eliminating the need for users to manually record each contact's name during setup.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system performs self-service by automatically processing the entire contact list and generating audio files without requiring user intervention. The text-to-speech engine independently converts all contact names to audio format and stores them for voice recognition, freeing the user from manual recording tasks.

Inventive Principle:
Principle #25Self-service

2Measurement precision

If manual selection and recording of contact names is used for voice dialing setup, then speech recognition accuracy is improved, but ease of operation deteriorates

Engineering Contradiction:
Improvespeech recognition accuracyVSAvoiduser convenience
Core Design Contradiction:
Measurement precisionVSEase of operation

Solution Approach 1:

The system performs preliminary action by automatically generating audio files for all contacts in the address book before the user needs to use voice dialing. The text-to-speech engine converts contact names to audio files and stores them for later recognition, eliminating the need for users to manually record each contact's name during setup.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system performs self-service by automatically processing the entire contact list and generating audio files without requiring user intervention. The text-to-speech engine independently converts all contact names to audio format and stores them for voice recognition, freeing the user from manual recording tasks.

Inventive Principle:
Principle #25Self-service

3Ease of operation

If automated text-to-speech conversion is used, then ease of operation is improved, but device complexity increases

Engineering Contradiction:
Improvesetup convenienceVSAvoidsystem complexity
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The system achieves multi-functionality by integrating the text-to-speech engine directly into the voice dialing apparatus. This single integrated component performs multiple functions: converting contact names to audio files, storing them for recognition, and enabling voice dialing operations, thereby reducing overall system complexity despite the added capability.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The text-to-speech engine acts as an intermediary component that bridges the gap between text contact information and audio recognition. By introducing this specialized module, the system simplifies the overall architecture by centralizing the conversion function rather than requiring multiple separate components for contact processing and audio generation.

Inventive Principle:
Principle #24Intermediary (Mediator)

4Productivity

If automated text-to-speech conversion is used, then productivity is improved, but device complexity increases

Engineering Contradiction:
Improvevoice dialing setup efficiencyVSAvoidsystem complexity
Core Design Contradiction:
ProductivityVSDevice complexity

Solution Approach 1:

The system achieves multi-functionality by integrating the text-to-speech engine directly into the voice dialing apparatus. This single integrated component performs multiple functions: converting contact names to audio files, storing them for recognition, and enabling voice dialing operations, thereby reducing overall system complexity despite the added capability.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The text-to-speech engine acts as an intermediary component that bridges the gap between text contact information and audio recognition. By introducing this specialized module, the system simplifies the overall architecture by centralizing the conversion function rather than requiring multiple separate components for contact processing and audio generation.

Inventive Principle:
Principle #24Intermediary (Mediator)

Data Source

PatentUS7636426B2Method and apparatus for automated voice dialing setup
Publication Date: 2009.12.22 UNIFY BETEILIGUNGSVERWALTUNG GMBH & CO KG
  • US7636426B2 patent drawing
  • US7636426B2 patent drawing
  • US7636426B2 patent drawing

AI summary

A telecommunications device includes a voice dialer and a text-to-speech engine. The text-to-speech engine is configured to convert at least a portion of a user contact list information to speech and the voice dialer is configured to receive an audio input and perform a voice recognition, comparing said audio input to converted user contact list information.