Silent Voice Communication via Pre-recorded Snippet Menus

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

There is a need for a method to enable two-way voice communication between a user and a remote party where the user cannot speak aloud, such as in meetings or libraries, without disrupting others, as existing systems like Text-to-Speech are too time-consuming for mobile devices.

Innovation Solution

A system and process using a communication device with a user interface and display to select pre-recorded voice snippets from hierarchical menus to respond to a remote party, allowing users to communicate without speaking, with options for manual activation, automatic initiation, and the ability to record personal voice snippets.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If Text-to-Speech systems are used to enable silent communication, then users can communicate without speaking, but the system becomes too time-consuming for mobile devices

Engineering Contradiction:
Improveease of communication without speakingVSAvoidtime consumption
Core Design Contradiction:
Ease of operationVSLoss of time

Solution Approach 1:

The system pre-records multiple voice snippets corresponding to different menu selections before actual use. When a user needs to communicate silently, they simply select from pre-prepared menus rather than generating speech in real-time, thus eliminating the time-consuming text-to-speech conversion process while enabling immediate communication responses

Inventive Principle:
Principle #10Preliminary action

2Ease of operation

If a user answers a call by moving to another location or telling the caller they will call back, then the user can speak privately, but people around are disturbed by the action

Engineering Contradiction:
Improveability to answer call privatelyVSAvoiddisturbance to surrounding people
Core Design Contradiction:
Ease of operationVSObject-affected harmful factors

Solution Approach 1:

The system introduces an intermediary mechanism (pre-recorded voice snippets selected via menu) between the user and the caller. This allows the user to remain in place without moving or making audible responses, while still conveying information to the caller through the playback of pre-recorded messages, thus eliminating disturbance to surrounding people

Inventive Principle:
Principle #24Intermediary (Mediator)

3Speed

If pre-recorded voice snippets are stored in the communication device memory, then response transmission is fast, but device memory capacity is consumed

Engineering Contradiction:
Improveresponse transmission speedVSAvoidmemory capacity
Core Design Contradiction:
SpeedVSQuantity of substance

Solution Approach 1:

The system segments the voice storage into two parts: frequently used voice snippets are stored locally in the communication device memory for fast access, while less frequently used snippets can be stored remotely on a server or cloud storage. The system intelligently loads only necessary snippets into local memory, thus achieving fast response transmission while minimizing local memory consumption

Inventive Principle:
Principle #1Segmentation

Data Source

PatentUS7443962B2System and process for speaking in a two-way voice communication without talking using a set of speech selection menus
Publication Date: 2008.10.28 MICROSOFT TECHNOLOGY LICENSING LLC
  • US7443962B2 patent drawing
  • US7443962B2 patent drawing
  • US7443962B2 patent drawing

AI summary

A system and process for enabling a communication device having computing capability, a user interface and display, to conduct two-way voice communications between a user and a remote party over a communication link in such a manner that the remote party speaks but the user does not, is presented. In general, a series of menus listing potential responses is displayed on the display of the communication device. In addition, there are a plurality of backchanneling responses provided that the user can select. These responses are employed by the user to communicate with the remote party, rather than speaking. This is accomplished by the user selecting one of the available responses. Once a selection has been made, a pre-recorded voice snippet corresponding to the selected response is accessed. The accessed voice snippet is then played back and transmitted to the remote party over the communication link.