Asynchronous Audio Messaging System with Speech Recognition

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Traditional text messaging systems are cumbersome and voicemail communications are time-consuming and impractical, lacking contextual information, making them inefficient for real-time communication and message exchange.

Innovation Solution

An asynchronous audio messaging system that allows users to record and send voice messages with recipient and action information, using speech recognition to convert audio into text and transmit both formats, with optional enhancements like expiration reminders and contextual information.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Productivity

If traditional text messaging systems are used, then messages can be transmitted, but the process requires multiple steps and takes time to compose

Engineering Contradiction:
Improvemessage transmission speedVSAvoidmessage composition complexity
Core Design Contradiction:
ProductivityVSEase of operation

Solution Approach 1:

The patent replaces the mechanical keyboard input system with an audio-based input system. Users speak their messages which are captured by a microphone and processed through automatic speech recognition, eliminating the need for physical or virtual keyboard interaction and significantly reducing message composition time

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Solution Approach 2:

The system performs automatic speech recognition and message formatting without requiring user intervention. The audio message is automatically transcribed to text, formatted with metadata, and prepared for transmission without the user needing to manually compose or edit the message

Inventive Principle:
Principle #25Self-service

2Loss of time

If conventional voicemail communications are used, then messages can be stored and retrieved, but users spend unnecessarily long time to obtain voicemail messages

Engineering Contradiction:
Improvevoicemail retrieval timeVSAvoidvoicemail access complexity
Core Design Contradiction:
Loss of timeVSEase of operation

Solution Approach 1:

The system performs preliminary actions by automatically transcribing the audio message to text and formatting it with metadata before the user needs to retrieve it. This pre-processing eliminates the need for users to listen to lengthy voicemail instructions and commands, allowing them to quickly access and understand message content

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent extracts the essential message content from the audio recording and presents it in text format, separating the core information from the redundant voicemail interface elements such as navigation instructions and system announcements, allowing users to directly access the meaningful content

Inventive Principle:
Principle #2Taking out (Extraction)

3Loss of information

If voicemail communications are used, then messages can be transmitted, but no additional contextual information is provided to help users understand message context

Engineering Contradiction:
Improvemessage contextual informationVSAvoidmessage understanding efficiency
Core Design Contradiction:
Loss of informationVSProductivity

Solution Approach 1:

The patent segments the message information into distinct components: the transcribed text content, metadata (sender, recipient, timestamp), and optional contextual tags. This structured segmentation allows users to quickly scan and understand different aspects of the message without having to process a monolithic audio stream

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The system introduces text transcription as an intermediary representation between the audio message and the user. This intermediate text format provides contextual information and allows users to quickly assess message content before deciding whether to listen to the full audio, improving message understanding efficiency

Inventive Principle:
Principle #24Intermediary (Mediator)

Applied Scientific Principles

This section explains which scientific principles are used to turn an abstract innovation direction into a practical engineering solution.

Function Achieved in This Case

Enables quick and efficient message creation and transmission, providing contextual information and reducing the time required for message retrieval, enhancing user convenience and accessibility.

Implementation Method 1

the speech processing module may perform speech recognition on the audio to determine a command from the audio, such as a command to 'send a message,' and a recipient, such as 'Bob'

Methodology Applied
Scientific EffectAutomatic Speech Recognition:

Data Source

PatentUS10002611B1Asynchronous audio messaging
Publication Date: 2018.06.19 AMAZON TECH INC
  • US10002611B1 patent drawing
  • US10002611B1 patent drawing
  • US10002611B1 patent drawing

AI summary

Systems, devices, and techniques may provide asynchronous audio messaging. Asynchronous audio messaging may enable a user to quickly and easily create and transmit a message to a recipient. The user may simply record a message for a recipient. The message may include an indication of the recipient of the message, an action (e.g., to send a message, etc.) and/or other types of information. A messaging module may modify the message to create a modified version of the message and then generate an additional version of the modified message in a different media type. The modified message and the addition version of the modified message may be transmitted to the recipient. In some embodiments, the messaging module may transmit other information such as location information, an expiration, or other information derived from the message to enhance the message.