Text-Mapped Audio Processing for Precise Location Editing
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing audio processing methods are inefficient and labor-intensive due to the difficulty in accurately determining the to-be-processed location, requiring repeated listening to audio, which lowers processing efficiency and effectiveness.
Innovation Solution
An audio processing method that displays target audio and corresponding text information, allowing for the selection of a to-be-processed text location, which is then mapped to a corresponding audio location for precise processing, thereby improving efficiency and accuracy.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If a user determines a to-be-processed location in audio by listening to the audio, then the user can identify the location, but the processing efficiency is low due to repeated listening
Solution Approach 1:
The patent introduces text information as an intermediary between the audio and the user. The text information corresponds to the audio and is displayed on the interface, allowing users to locate processing positions through text rather than repeatedly listening to audio. This intermediary enables accurate location determination while significantly improving processing efficiency.
2Productivity
If text information is displayed corresponding to audio with mapping relationship, then processing location can be determined quickly, but the device complexity increases
Solution Approach 1:
The text information serves multiple functions: it represents the audio content, enables location determination, and provides a interface for user interaction. By making the text information multi-functional, the patent achieves improved processing efficiency without proportionally increasing system complexity, as the same text element performs multiple roles.
Data Source
AI summary
This application relates to an audio processing method, an electronic device, and a storage medium. The method includes: displaying a target audio clip and corresponding target text information having a mapping relationship between a location of an audio segment in the target audio clip and a location of text information in the corresponding target text information; receiving, a selection of a location in the corresponding target text information as a to-be-processed text location; matching a to-be-processed audio location of an audio segment that has the mapping relationship with the to-be-processed text location; and processing the target audio at the to-be-processed audio location to generate an updated target audio clip, and updating the corresponding target text information at the to-be-processed text location to generate updated target text information; and displaying the updated target audio clip and the updated target text information.


