Audio Call Analysis with Automatic Pattern Recognition
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
During audio calls, users often face difficulties in capturing and recording information such as phone numbers due to being engaged in activities like driving or running, and sending text messages may incur additional costs or not be supported in all locations.
Innovation Solution
A device with a communication interface and input interface that generates an audio recording of an audio call, performs speech-to-text conversion, and identifies patterns like phone numbers, allowing users to select actions such as updating contact information directly from the audio call without needing to write down the information.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If users manually write down information during audio calls, then information accuracy is improved, but user convenience deteriorates and safety is compromised due to distractions
Solution Approach 1:
The system performs automatic speech-to-text conversion and pattern recognition without requiring user intervention. The device captures audio, converts it to text, identifies patterns like phone numbers automatically, and presents actions for the user to review and confirm, making the system serve itself rather than requiring manual note-taking
Solution Approach 2:
The patent replaces the mechanical action of manually writing down information with an automated computational system. Speech-to-text conversion algorithms and pattern recognition software substitute for the physical act of writing, automatically extracting and identifying information patterns from the audio call
2Loss of information
If users send text messages to share information, then information exchange is improved, but additional costs are incurred and service availability is reduced
Solution Approach 1:
The system automatically captures and processes information shared during audio calls through speech-to-text conversion and pattern recognition. It identifies patterns like phone numbers, emails, and addresses, then presents actionable options to the user, eliminating the need for manual text message follow-ups and associated costs
3Productivity
If automated speech-to-text conversion is performed, then information capture speed is improved, but device complexity increases
Solution Approach 1:
The device leverages existing multi-functional capabilities already present in modern mobile devices, including audio recording, speech-to-text conversion, and pattern recognition. By utilizing these universal functions that serve multiple purposes, the patent avoids adding dedicated specialized components, thereby managing complexity while maintaining high information capture speed
Data Source
AI summary
A device includes a communication interface, an input interface, and a processor. The communication interface is configured to receive an audio signal associated with an audio call. The input interface is configured to receive user input during the audio call. The processor is configured to generate an audio recording of the audio signal in response to receiving the user input. The processor is also configured to generate text by performing speech-to-text conversion of the audio recording. The processor is further configured to perform a comparison of the text to a pattern. The processor is also configured to identify, based on the comparison, a portion of the text that matches the pattern. The processor is further configured to provide the portion of the text and an option to a display. The option is selectable to initiate performance of an action corresponding to the pattern.


