The invention relates to the technical field of voice synthesis, in particular to a multi-dimensional AI platform intelligent voice
response system using the voice synthesis technology, which screens out matched historical user question voices according to text
semantic similarity corresponding to
current user question voices and historical user question voices, and sends the matched historical user question voices to a user terminal. Obtaining a
voice training set and the weight of each element in the
voice training set according to the voice feature similarity between the question voice of the
current user and the question voice of the matched historical user in combination with the user
score value of the manual reply voice, training an
acoustic model, obtaining a reply text corresponding to the question voice of the
current user, and obtaining the question voice of the current user; and inputting into the trained
acoustic model, generating a reply voice
signal, and then outputting to the current user. According to the method, the
voice training set is screened and constructed from historical manual reply voices, so that an
acoustic model can learn a more natural and smooth voice synthesis mode, and the reply voice
signal contains rich voice features and expression
modes.