Voice Recording Interface with Real-Time Text Search
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Users face inconvenience in checking the content of previously recorded audio without having to terminate the current audio recording and restart it, as existing methods require stopping and replaying audio files to access previous recordings.
Innovation Solution
An electronic device with a memory to store audio signals, a display unit to show the progressive state and STT-based text, and a controller to allow selection and reproduction of specific audio portions without stopping the recording, using speech-to-text technology to convert and display audio signals in real-time.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If the user terminates audio recording to check previous content, then the user can review recorded audio, but the user must restart recording which increases operation time and reduces convenience
Solution Approach 1:
The system performs preliminary actions by continuously storing audio signals in memory during the recording process and converting them to text in real-time. This allows the user to access and review previously recorded content without needing to terminate and restart the recording, thus resolving the contradiction between ease of operation and time loss.
Solution Approach 2:
The system creates a textual copy of the audio signal using speech-to-text conversion, which is stored in memory alongside the original audio. This copy allows the user to review content by selecting text portions rather than replaying the entire audio file, reducing the time and effort required to check recorded content.
2Productivity
If the user replays audio file to check content, then the user can review recorded audio, but the existing method requires stopping and restarting which reduces productivity
Solution Approach 1:
The system prepares the audio signal for quick access by continuously converting it to text format and storing both the audio and text in memory during recording. This preliminary processing enables the user to instantly access specific portions of recorded content without time-consuming playback operations, significantly improving productivity.
Solution Approach 2:
The system extracts the essential information from the audio signal by converting it to text through speech-to-text technology. This extracted text can be selectively searched and reviewed, allowing users to quickly access specific content without replaying the entire audio file, thus reducing time loss and improving efficiency.
3Ease of operation
If the system stores audio signal continuously, then the user can access any portion without interrupting recording, but the system complexity increases
Solution Approach 1:
The memory component serves multiple functions: it stores the original audio signal, stores the converted text, and enables random access to any portion of recorded content. This multi-functionality allows the system to provide real-time access to recorded portions without requiring separate systems for storage and conversion, thus managing complexity while improving ease of operation.
Solution Approach 2:
The speech-to-text conversion process acts as an intermediary between the audio signal and the user interface. By converting audio to text in real-time and storing both formats in memory, the system creates a bridge that allows users to access content through text selection without interrupting the audio recording process, balancing functionality with manageable complexity.
Applied Scientific Principles
This section explains which scientific principles are used to turn an abstract innovation direction into a practical engineering solution.
Function Achieved in This Case
Enables users to check and reproduce specific portions of recorded audio without interrupting the current recording process, enhancing user convenience and efficiency in managing audio recordings.
Implementation Method 1
the controller may convert the audio signal being input into text, and the memory may store the converted text
Data Source
AI summary
An electronic device for allowing the user not to terminate audio recording while at the same time checking the content corresponding to a recorded portion in real time during the audio recording, and a method of reproducing the audio signal. An electronic device according to an embodiment disclosed in the present disclosure may include a memory configured to store an audio signal being input; a display unit configured to display at least one of an item indicating a progressive state in which the audio signal is stored therein and an STT-based text for the audio signal; a user interface unit configured to receive the selection of a predetermined portion of the item indicating the progressive state or the selection of a partial character string of the text from the user; and a controller configured to reproduce an audio signal corresponding to the selected portion or the selected character string.


