The invention relates to the technical field of voice
processing, can be applied to business scenes of financial science and technology,
medical health and the like, and discloses an
audio signal authenticity
verification method, device, equipment and medium, and the method comprises the steps: constructing an original audio text
data set, and generating an adversarial sample set, inputting the original audio text
data set and the adversarial sample set into an audio detection model for joint training to obtain an audio detection model subjected to adversarial training; the method comprises the steps of obtaining a to-be-detected
audio signal and extracting an acoustic feature of the to-be-detected
audio signal, obtaining a non-acoustic feature associated with the to-be-detected audio
signal, constructing a multi-dimensional
feature vector according to the acoustic feature and the non-acoustic feature, inputting the multi-dimensional
feature vector into an audio detection model to generate an abnormal index, and executing a hierarchical response operation based on the abnormal index. According to the method, the robustness of the model is enhanced by introducing adversarial sample training, and the multi-dimensional
feature vector is constructed by fusing the multi-
modal features, so that accurate recognition and hierarchical response to the voice
cloning attack are realized.