Music Segment Recognition for Automatic Video Clip Extraction

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing video editing methods require manual interception of video clips, which is time-consuming and inefficient.

Innovation Solution

A method for automatically recognizing music segments in video data by performing music recognition on audio frames, determining music segments, and extracting video clips with the same playback period as the music segments, thereby enabling automatic video editing.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Productivity

If manual interception of video clips is used, then video editing can be performed, but the process is time-consuming and inefficient

Engineering Contradiction:
Improvevideo editing efficiencyVSAvoidtime for manual interception
Core Design Contradiction:
ProductivityVSLoss of time

Solution Approach 1:

The system performs automatic music recognition and video clip extraction without requiring manual operation. The computer device autonomously identifies music segments in audio data and extracts corresponding video clips, making the editing process self-service and eliminating time-consuming manual interception

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The patent replaces the mechanical manual operation of video clip interception with an automated music recognition system. By using audio analysis algorithms to identify music segments and automatically extract corresponding video portions, the system substitutes human manual work with automated computational processes

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Productivity

If automatic music recognition is performed on audio frames, then video editing efficiency is improved, but the device complexity increases

Engineering Contradiction:
Improvevideo editing efficiencyVSAvoidcomplexity of music recognition system
Core Design Contradiction:
ProductivityVSDevice complexity

Solution Approach 1:

The audio data is divided into discrete audio frames, and music recognition is performed on each frame independently. This segmentation approach breaks down the complex task of analyzing entire audio tracks into manageable frame-level operations, making the system more tractable while maintaining automation

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The system performs music recognition on audio frames at a granular level, potentially analyzing more frames than strictly necessary. This partial action approach ensures thorough music segment identification by examining individual frames, which simplifies the overall logic compared to attempting to recognize music patterns across entire audio tracks at once

Inventive Principle:
Principle #16Partial or excessive action

Data Source

PatentUS20260065941A1Video editing method and apparatus, computer device, and storage medium
Publication Date: 2026.03.05 TENCENT TECHNOLOGY (SHENZHEN) CO LTD
  • US20260065941A1 patent drawing
  • US20260065941A1 patent drawing
  • US20260065941A1 patent drawing

AI summary

A video editing method is performed by a computer device. The method includes: performing music recognition on audio data in first video data to obtain a recognition result of each of audio frames in the audio data, the recognition result indicating whether the audio frame belongs to a music audio frame; determining a music segment in the audio data based on the recognition results of the audio frames, the music segment comprising a plurality of music audio frames; and extracting, from the first video data, a video clip with a same playback period as the music segment as second video data comprising the music segment.