Voice Recording Interface with Real-Time Text Search

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Users face inconvenience in checking the content of previously recorded audio without having to terminate the current audio recording and restart it, as existing methods require stopping and replaying audio files to access previous recordings.

Innovation Solution

An electronic device with a memory to store audio signals, a display unit to show the progressive state and STT-based text, and a controller to allow selection and reproduction of specific audio portions without stopping the recording, using speech-to-text technology to convert and display audio signals in real-time.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If the user terminates audio recording to check previous content, then the user can review recorded audio, but the user must restart recording which increases operation time and reduces convenience

Engineering Contradiction:
Improveconvenience of checking recorded contentVSAvoidtime to check and replay audio
Core Design Contradiction:
Ease of operationVSLoss of time

Solution Approach 1:

The system performs preliminary actions by continuously storing audio signals in memory during the recording process and converting them to text in real-time. This allows the user to access and review previously recorded content without needing to terminate and restart the recording, thus resolving the contradiction between ease of operation and time loss.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system creates a textual copy of the audio signal using speech-to-text conversion, which is stored in memory alongside the original audio. This copy allows the user to review content by selecting text portions rather than replaying the entire audio file, reducing the time and effort required to check recorded content.

Inventive Principle:
Principle #26Copying

2Productivity

If the user replays audio file to check content, then the user can review recorded audio, but the existing method requires stopping and restarting which reduces productivity

Engineering Contradiction:
Improveefficiency in managing audio recordingsVSAvoidtime to access recorded content
Core Design Contradiction:
ProductivityVSLoss of time

Solution Approach 1:

The system prepares the audio signal for quick access by continuously converting it to text format and storing both the audio and text in memory during recording. This preliminary processing enables the user to instantly access specific portions of recorded content without time-consuming playback operations, significantly improving productivity.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system extracts the essential information from the audio signal by converting it to text through speech-to-text technology. This extracted text can be selectively searched and reviewed, allowing users to quickly access specific content without replaying the entire audio file, thus reducing time loss and improving efficiency.

Inventive Principle:
Principle #2Taking out (Extraction)

3Ease of operation

If the system stores audio signal continuously, then the user can access any portion without interrupting recording, but the system complexity increases

Engineering Contradiction:
Improveability to access recorded portions during recordingVSAvoidsystem structure for real-time storage and conversion
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The memory component serves multiple functions: it stores the original audio signal, stores the converted text, and enables random access to any portion of recorded content. This multi-functionality allows the system to provide real-time access to recorded portions without requiring separate systems for storage and conversion, thus managing complexity while improving ease of operation.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The speech-to-text conversion process acts as an intermediary between the audio signal and the user interface. By converting audio to text in real-time and storing both formats in memory, the system creates a bridge that allows users to access content through text selection without interrupting the audio recording process, balancing functionality with manageable complexity.

Inventive Principle:
Principle #24Intermediary (Mediator)

Applied Scientific Principles

This section explains which scientific principles are used to turn an abstract innovation direction into a practical engineering solution.

Function Achieved in This Case

Enables users to check and reproduce specific portions of recorded audio without interrupting the current recording process, enhancing user convenience and efficiency in managing audio recordings.

Implementation Method 1

the controller may convert the audio signal being input into text, and the memory may store the converted text

Methodology Applied
Scientific EffectSpeech-to-text conversion:

Data Source

PatentUS9514749B2Method and electronic device for easy search during voice record
Publication Date: 2016.12.06 LG ELECTRONICS INC
  • US9514749B2 patent drawing
  • US9514749B2 patent drawing
  • US9514749B2 patent drawing

AI summary

An electronic device for allowing the user not to terminate audio recording while at the same time checking the content corresponding to a recorded portion in real time during the audio recording, and a method of reproducing the audio signal. An electronic device according to an embodiment disclosed in the present disclosure may include a memory configured to store an audio signal being input; a display unit configured to display at least one of an item indicating a progressive state in which the audio signal is stored therein and an STT-based text for the audio signal; a user interface unit configured to receive the selection of a predetermined portion of the item indicating the progressive state or the selection of a partial character string of the text from the user; and a controller configured to reproduce an audio signal corresponding to the selected portion or the selected character string.