Audio Call Analysis with Automatic Pattern Recognition

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

During audio calls, users often face difficulties in capturing and recording information such as phone numbers due to being engaged in activities like driving or running, and sending text messages may incur additional costs or not be supported in all locations.

Innovation Solution

A device with a communication interface and input interface that generates an audio recording of an audio call, performs speech-to-text conversion, and identifies patterns like phone numbers, allowing users to select actions such as updating contact information directly from the audio call without needing to write down the information.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If users manually write down information during audio calls, then information accuracy is improved, but user convenience deteriorates and safety is compromised due to distractions

Engineering Contradiction:
Improveinformation accuracyVSAvoiduser convenience
Core Design Contradiction:
Measurement precisionVSEase of operation

Solution Approach 1:

The system performs automatic speech-to-text conversion and pattern recognition without requiring user intervention. The device captures audio, converts it to text, identifies patterns like phone numbers automatically, and presents actions for the user to review and confirm, making the system serve itself rather than requiring manual note-taking

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The patent replaces the mechanical action of manually writing down information with an automated computational system. Speech-to-text conversion algorithms and pattern recognition software substitute for the physical act of writing, automatically extracting and identifying information patterns from the audio call

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Loss of information

If users send text messages to share information, then information exchange is improved, but additional costs are incurred and service availability is reduced

Engineering Contradiction:
Improveinformation exchangeVSAvoidservice availability
Core Design Contradiction:
Loss of informationVSAdaptability or versatility

Solution Approach 1:

The system automatically captures and processes information shared during audio calls through speech-to-text conversion and pattern recognition. It identifies patterns like phone numbers, emails, and addresses, then presents actionable options to the user, eliminating the need for manual text message follow-ups and associated costs

Inventive Principle:
Principle #25Self-service

3Productivity

If automated speech-to-text conversion is performed, then information capture speed is improved, but device complexity increases

Engineering Contradiction:
Improveinformation capture speedVSAvoidprocessing complexity
Core Design Contradiction:
ProductivityVSDevice complexity

Solution Approach 1:

The device leverages existing multi-functional capabilities already present in modern mobile devices, including audio recording, speech-to-text conversion, and pattern recognition. By utilizing these universal functions that serve multiple purposes, the patent avoids adding dedicated specialized components, thereby managing complexity while maintaining high information capture speed

Inventive Principle:
Principle #6Universality (Multi-functionality)

Data Source

PatentUS10701198B2Audio call analysis
Publication Date: 2020.06.30 QUALCOMM INC
  • US10701198B2 patent drawing
  • US10701198B2 patent drawing
  • US10701198B2 patent drawing

AI summary

A device includes a communication interface, an input interface, and a processor. The communication interface is configured to receive an audio signal associated with an audio call. The input interface is configured to receive user input during the audio call. The processor is configured to generate an audio recording of the audio signal in response to receiving the user input. The processor is also configured to generate text by performing speech-to-text conversion of the audio recording. The processor is further configured to perform a comparison of the text to a pattern. The processor is also configured to identify, based on the comparison, a portion of the text that matches the pattern. The processor is further configured to provide the portion of the text and an option to a display. The option is selectable to initiate performance of an action corresponding to the pattern.