Video Generation Method Using Audio Segmentation
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional video editing methods require significant user effort and time, as they involve processing multiple materials, making the editing process cumbersome and time-consuming.
Innovation Solution
A method and apparatus for generating a video by obtaining an image material and an audio material, determining music points to divide the audio into segments, generating video segments of corresponding lengths, and splicing them together with the audio as an audio track, thereby simplifying the editing process.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If conventional video editing methods are used to process multiple materials, then video editing functionality is achieved, but user time and energy consumption increase significantly
Solution Approach 1:
The audio material is segmented into multiple audio segments based on detected music points, and corresponding video segments are generated for each audio segment. This segmentation allows automated processing of different parts of the video with appropriate image materials, significantly reducing manual editing time while maintaining editing quality
Solution Approach 2:
The system automatically detects music points in the audio material, divides the audio into segments, selects appropriate image materials, and generates video segments without requiring continuous user intervention. The automated workflow enables the system to serve itself in the editing process, minimizing user time investment
2Productivity
If automated video generation is implemented to reduce user effort, then editing efficiency improves, but video content diversity may be compromised
Solution Approach 1:
Different image materials are selected and applied to different video segments based on their local characteristics and corresponding audio segments. The system allows for localized customization where users can specify different image materials for different music points, ensuring each segment has appropriate and diverse content while maintaining automated generation efficiency
Solution Approach 2:
The system dynamically adjusts the video generation process by detecting music points and adapting the segmentation and image material selection to the specific characteristics of each audio material. This dynamic approach ensures high productivity while maintaining content diversity through automated adaptation to different input materials
Data Source
AI summary
Disclosed are a video generation method and apparatus, an electronic device, and a computer readable medium. A specific embodiment of the method comprises: obtaining a video footage and an audio footage, the video footage comprising a picture footage; determining a music point of the audio footage, the music point being used for dividing the audio footage into a plurality of audio clips; using the video footage to generate a video clip for each music clip in the audio footage to obtain a plurality of video clips, corresponding music clips and video clips having the same duration; and splicing the plurality of video clips according to the time when music clips respectively corresponding to the plurality of video clips appear in the audio footage, and adding the audio footage as a video audio signal to obtain a composite video.


