Music Segment Recognition for Automatic Video Clip Extraction
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing video editing methods require manual interception of video clips, which is time-consuming and inefficient.
Innovation Solution
A method for automatically recognizing music segments in video data by performing music recognition on audio frames, determining music segments, and extracting video clips with the same playback period as the music segments, thereby enabling automatic video editing.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Productivity
If manual interception of video clips is used, then video editing can be performed, but the process is time-consuming and inefficient
Solution Approach 1:
The system performs automatic music recognition and video clip extraction without requiring manual operation. The computer device autonomously identifies music segments in audio data and extracts corresponding video clips, making the editing process self-service and eliminating time-consuming manual interception
Solution Approach 2:
The patent replaces the mechanical manual operation of video clip interception with an automated music recognition system. By using audio analysis algorithms to identify music segments and automatically extract corresponding video portions, the system substitutes human manual work with automated computational processes
2Productivity
If automatic music recognition is performed on audio frames, then video editing efficiency is improved, but the device complexity increases
Solution Approach 1:
The audio data is divided into discrete audio frames, and music recognition is performed on each frame independently. This segmentation approach breaks down the complex task of analyzing entire audio tracks into manageable frame-level operations, making the system more tractable while maintaining automation
Solution Approach 2:
The system performs music recognition on audio frames at a granular level, potentially analyzing more frames than strictly necessary. This partial action approach ensures thorough music segment identification by examining individual frames, which simplifies the overall logic compared to attempting to recognize music patterns across entire audio tracks at once
Data Source
AI summary
A video editing method is performed by a computer device. The method includes: performing music recognition on audio data in first video data to obtain a recognition result of each of audio frames in the audio data, the recognition result indicating whether the audio frame belongs to a music audio frame; determining a music segment in the audio data based on the recognition results of the audio frames, the music segment comprising a plurality of music audio frames; and extracting, from the first video data, a video clip with a same playback period as the music segment as second video data comprising the music segment.


