The invention discloses a no-reference sound
correction method based on a music
score big
language model, and relates to a mismatch song
correction method in the technical field of intelligent audio
processing. The objective of the invention is to solve the problems of poor
pitch and
rhythm collaborative correction effect, low note
feature extraction precision and easy loss of emotional expression in a non-reference scene in the existing sound correction technology. The method comprises the following steps: S1, acquiring and preprocessing off-tune singing audio data; s2, carrying out note
accurate segmentation on the preprocessed audio; s3, note level features are extracted, and Octuple-
MIDI symbolization input is generated; s4, realizing non-reference
pitch-
rhythm joint error correction through a music
score big
language model, and outputting a standard
MIDI sequence; s5, based on a standard
MIDI sequence, completing cooperative accurate correction of the sound
pitch and the
rhythm of the song; and S6, outputting the natural corrected singing sound through a high-fidelity synthesis technology. The method is used in the field of off-tune singing
processing such as karaoke entertainment, AI music creation, singing sound correction and the like.