Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

5 results about "Speaker adaptation" patented technology

IMPROVING THE NATURALNESS OF SPEAKER-ADAPTED SPEECH SYNTHESIS

Techniques for improving the naturalness of synthetic speech are revealed. A speaker-adaptation model for a speech synthesis pipeline is presented, achieved by training an acoustic model, along with a post-processing model for modifying speech features output by the acoustic model. Additional reference training examples based on simulated input-output datasets with limited resources for reference speakers are provided for this training.
Owner:FRAUNHOFER GESELLSCHAFT ZUR FORDERUNG DER ANGEWANDTEN FORSCHUNG EV

Speaker-adaptive speech end detection for conversational AI applications

In various instances, an end of speech (EOS) of an audio signal is determined based at least in part on a speech rate of a speaker, and for a segment of the audio signal, an EOS is indicated based at least in part on an EOS threshold determined based at least in part on the speech rate of the speaker.
Owner:NVIDIA CORP

Naturalness of speaker-adapted speech synthesis

Techniques of improving the naturalness of synthetic speech are disclosed. Speaker-adaption of a speech synthesis pipeline by training of an acoustic model and a post-processing model for modifying speech features output by the acoustic model are disclosed. For this training, additional reference training samples are obtained based on simulated low-resource input-output datasets for reference speakers.
Owner:FRAUNHOFER GESELLSCHAFT ZUR FORDERUNG DER ANGEWANDTEN FORSCHUNG EV

Residual adapters for few-shot text-to-speech speaker adaptation

A method for residual adapters for few-shot text-to-speech speaker adaptation includes obtaining a text-to-speech (TTS) model configured to convert text into representations of synthetic speech, the TTS model pre-trained on an initial training data set. The method further includes augmenting the TTS model with a stack of residual adapters. The method includes receiving an adaption training data set including one or more spoken utterances spoken by a target speaker, each spoken utterance in the adaptation training data set paired with corresponding input text associated with a transcription of the spoken utterance. The method also includes adapting, using the adaption training data set, the TTS model augmented with the stack of residual adapters to learn how to synthesize speech in a voice of the target speaker by optimizing the stack of residual adapters while parameters of the TTS model are frozen.
Owner:GOOGLE LLC

Residual adapters for few-shot text-to-speech speaker adaptation

A method for residual adapters for few-shot text-to-speech speaker adaptation includes obtaining a text-to-speech (TTS) model configured to convert text into representations of synthetic speech, the TTS model pre-trained on an initial training data set. The method further includes augmenting the TTS model with a stack of residual adapters. The method includes receiving an adaption training data set including one or more spoken utterances spoken by a target speaker, each spoken utterance in the adaptation training data set paired with corresponding input text associated with a transcription of the spoken utterance. The method also includes adapting, using the adaption training data set, the TTS model augmented with the stack of residual adapters to learn how to synthesize speech in a voice of the target speaker by optimizing the stack of residual adapters while parameters of the TTS model are frozen.
Owner:GOOGLE LLC