The embodiment of the invention provides a sound
copying method and device, equipment and a storage medium, which can be applied to scenes such as cloud technology,
artificial intelligence, intelligent traffic, auxiliary driving, audio and video, in the method,
feature extraction is performed on reference audio in advance, and reference
timbre features and reference
rhythm features are obtained. And obtaining and storing a sound feature file of the reference audio based on the reference
timbre feature and the reference
rhythm feature, so that when the reference audio is used as input for multiple times of
audio synthesis, only the sound feature file of the reference audio needs to be read in each time of
audio synthesis, and based on the first text feature of the text to be synthesized and the sound feature file, the sound feature file of the reference audio is read. According to the method and the device, the synthetic audio corresponding to the to-be-synthesized text is generated without repeatedly reading the reference audio and repeatedly calculating the reference audio, so that the time consumption and
resource consumption of sound
copying are effectively reduced, and the
waiting time of synthesizing the audio by using a sound
copying model is also effectively reduced, thereby improving the use experience and enhancing the
controllability of the
system.