Audio Processing Method for Multi-Track Harmony Synthesis
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional audio processing methods result in poor auditory effects due to users' lack of professional training, leading to suboptimal control over vocal and resonance aspects during singing, which affects the quality of dry audio recordings.
Innovation Solution
A method and apparatus for audio processing that involves determining the beginning and ending times of lyric words, detecting pitch and fundamental frequency, tuning up the audio by specific key intervals to create harmonies, and synthesizing these harmonies with the target dry audio to enhance its auditory effect.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If users sing without professional training and control, then the recording process is simple and quick, but the auditory effect of the dry audio is poor
Solution Approach 1:
The system automatically performs pitch detection, harmony generation, and audio processing without requiring user intervention or professional singing skills. The user simply records dry audio, and the system self-service completes the enhancement process through automated pitch correction and harmony synthesis.
Solution Approach 2:
The system changes audio parameters by detecting pitch and generating harmonies at different key intervals. It transforms the dry audio into enhanced audio with improved auditory effect through parameter adjustments in pitch, frequency, and harmonic layers.
2Manufacturing precision
If multiple harmonies are synthesized to improve auditory effect, then the audio quality is enhanced, but the processing complexity increases
Solution Approach 1:
The processing is segmented into distinct modules: pitch detection module, harmony generation module (with first and second harmonies at different key intervals), and mixing module. Each module handles a specific task, making the complex processing manageable and systematic.
Solution Approach 2:
The system generates multiple harmonies (first harmony with positive integer key interval, second harmony with larger key interval) which may be more than strictly necessary, but this excessive action ensures optimal auditory effect by providing rich harmonic layers for selection and mixing.
Data Source
AI summary
A method and apparatus for audio processing, an electronic device, and a computer-readable storage medium are provided in the present disclosure. The method includes: obtaining a target dry audio, and determining a beginning and ending time of each lyric word in the target dry audio; detecting a pitch of the target dry audio and a fundamental frequency during the beginning and ending time, and determining a current pitch name of the lyric word based on the fundamental frequency and the pitch; tuning up the lyric word by a first key interval to obtain a first harmony, and tuning up the lyric word by different second key intervals respectively to obtain different second harmonies; synthesizing the first harmony and the second harmonies to form a multi-track harmony; and mixing the multi-track harmony with the target dry audio to obtain a synthesized dry audio.


