Text-Based Audio Editing for Precise Segment Replacement
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional audio processing methods require manual adjustment of progress bars for editing, leading to cumbersome operations and low efficiency.
Innovation Solution
An audio processing method that involves obtaining text information with playback periods, receiving user inputs to select and modify fields, and modifying audio segments based on these inputs without manual progress bar adjustment.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If manual progress bar adjustment is used for audio editing, then the user can locate and modify audio segments, but the operation process becomes cumbersome and processing efficiency decreases
Solution Approach 1:
The patent introduces text information as an intermediary between the user and the audio content. Instead of directly manipulating the audio progress bar, users interact with text representations of audio segments, which then map back to the corresponding audio portions for editing. This intermediary layer simplifies the operation process while maintaining editing capability.
Solution Approach 2:
The patent replaces the mechanical interaction of manually dragging and adjusting progress bars with a text-based selection and modification system. Users select text segments and input modification instructions, which are then automatically translated into the corresponding audio segment edits, eliminating the need for precise manual progress bar manipulation.
2Measurement precision
If the user repeatedly adjusts the progress bar to accurately locate the playback period, then the desired audio segment can be found, but the operation time increases and efficiency decreases
Solution Approach 1:
The patent performs preliminary action by pre-processing the audio into text information with associated playback period data before the user needs to edit. This text representation is prepared in advance, allowing users to directly select and locate segments without repeated progress bar adjustments, thus saving time while maintaining precision.
Solution Approach 2:
The patent creates a text copy or representation of the audio content that mirrors the temporal structure of the original audio. This text copy serves as a simplified interface where users can accurately locate segments through text selection rather than time-consuming progress bar manipulation, preserving precision while reducing time loss.
Data Source
Figure 1
Figure 2-1
Figure 2-2~2-3
AI summary
Provided in the embodiments of the present invention are an audio processing method and an electronic device. The method comprises: first acquiring text information corresponding to an audio to be processed, wherein the text information comprises text to be processed and playing periods corresponding to fields in said text; then receiving a first input with regard to said text, and in response to the first input, determining a field, indicated by the first input, in said text as a field to be processed; next, receiving a second input with regard to the field to be processed, and in response to the second input, acquiring a target audio segment; and finally, according to the target audio segment, modifying an audio segment at the playing period corresponding to the field to be processed, in order to obtain a target audio.