Dual-Source Audio Synthesis via Energy Suppression
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current technologies lack an efficient method to automatically synthesize music from two songs with the same accompaniment but different vocals, requiring manual selection and processing, which is cumbersome and resource-intensive.
Innovation Solution
A dual sound source audio data processing method that identifies and filters songs with the same accompaniment using lyrics and accompaniment filtering, decodes audio data into mono channels, combines them, and performs energy suppression to create a seamless alternation effect, allowing for automatic synthesis of music with the same accompaniment sung differently.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If manual selection and processing of songs is performed, then synthesis accuracy is improved, but operation complexity and time consumption increase
Solution Approach 1:
The system automatically identifies same-source song pairs by extracting and comparing lyrics and accompaniment features without requiring manual selection. The processing module autonomously performs synthesis operations based on automatic identification results, eliminating the need for user intervention in song selection and pairing.
Solution Approach 2:
The system pre-processes song data by extracting lyrics, identifying accompaniment portions, and storing these features in advance. This preliminary extraction and organization of features enables rapid automatic identification and synthesis execution when needed, improving both accuracy and efficiency.
2Manufacturing precision
If manual processing of audio data is performed, then synthesis quality is improved, but resource consumption increases
Solution Approach 1:
The system extracts and separates lyrics portions from accompaniment portions in songs, storing these extracted features for efficient reuse. By pre-extracting and caching these essential components, the system avoids redundant processing during synthesis operations, reducing computational resource consumption while maintaining synthesis quality.
3Productivity
If automatic synthesis is implemented, then productivity is improved, but system complexity increases
Solution Approach 1:
The system divides the synthesis process into distinct functional modules: a feature extraction module that preprocesses songs, an identification module that automatically matches same-source pairs, and a synthesis execution module that performs the actual mixing. This modular segmentation enables automatic high-efficiency synthesis while organizing system complexity into manageable, independent components.
Data Source
Figure 1~2
Figure 3
Figure 4~5(b)
AI summary
Embodiments of the present invention provide a dual sound source audio data processing method and apparatus. The method includes: obtaining audio data of a same-source song pair, the same-source song pair including two songs having a same accompaniment but sung differently; decoding the audio data of the same-source song pair, to obtain two pieces of mono audio data; combining the two pieces of mono audio data to one piece of two-channel audio data; and dividing play time corresponding to a two-channel audio to multiple play periods, and performing the energy suppression on a left audio channel or a right audio channel of the two-channel audio in different play periods.