Audio Processing Method for Multi-Track Harmony Synthesis

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional audio processing methods result in poor auditory effects due to users' lack of professional training, leading to suboptimal control over vocal and resonance aspects during singing, which affects the quality of dry audio recordings.

Innovation Solution

A method and apparatus for audio processing that involves determining the beginning and ending times of lyric words, detecting pitch and fundamental frequency, tuning up the audio by specific key intervals to create harmonies, and synthesizing these harmonies with the target dry audio to enhance its auditory effect.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If users sing without professional training and control, then the recording process is simple and quick, but the auditory effect of the dry audio is poor

Engineering Contradiction:
Improverecording process simplicityVSAvoidauditory effect quality
Core Design Contradiction:
Ease of operationVSManufacturing precision

Solution Approach 1:

The system automatically performs pitch detection, harmony generation, and audio processing without requiring user intervention or professional singing skills. The user simply records dry audio, and the system self-service completes the enhancement process through automated pitch correction and harmony synthesis.

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The system changes audio parameters by detecting pitch and generating harmonies at different key intervals. It transforms the dry audio into enhanced audio with improved auditory effect through parameter adjustments in pitch, frequency, and harmonic layers.

Inventive Principle:
Principle #35Parameter changes

2Manufacturing precision

If multiple harmonies are synthesized to improve auditory effect, then the audio quality is enhanced, but the processing complexity increases

Engineering Contradiction:
Improveauditory effect qualityVSAvoidprocessing complexity
Core Design Contradiction:
Manufacturing precisionVSDevice complexity

Solution Approach 1:

The processing is segmented into distinct modules: pitch detection module, harmony generation module (with first and second harmonies at different key intervals), and mixing module. Each module handles a specific task, making the complex processing manageable and systematic.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The system generates multiple harmonies (first harmony with positive integer key interval, second harmony with larger key interval) which may be more than strictly necessary, but this excessive action ensures optimal auditory effect by providing rich harmonic layers for selection and mixing.

Inventive Principle:
Principle #16Partial or excessive action

Data Source

PatentUS20230402047A1Audio processing method and apparatus, electronic device, and computer-readable storage medium
Publication Date: 2023.12.14 TENCENT MUSIC ENTERTAINMENT TECH (SHENZHEN) CO LTD
  • US20230402047A1 patent drawing
  • US20230402047A1 patent drawing
  • US20230402047A1 patent drawing

AI summary

A method and apparatus for audio processing, an electronic device, and a computer-readable storage medium are provided in the present disclosure. The method includes: obtaining a target dry audio, and determining a beginning and ending time of each lyric word in the target dry audio; detecting a pitch of the target dry audio and a fundamental frequency during the beginning and ending time, and determining a current pitch name of the lyric word based on the fundamental frequency and the pitch; tuning up the lyric word by a first key interval to obtain a first harmony, and tuning up the lyric word by different second key intervals respectively to obtain different second harmonies; synthesizing the first harmony and the second harmonies to form a multi-track harmony; and mixing the multi-track harmony with the target dry audio to obtain a synthesized dry audio.