Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

3 results about "Speaker recognition system" patented technology

A sentiment asymmetric speaker recognition system based on a generative adversarial network

ActiveCN116543774BSolving the problem where robustness will be greatly reducedRich emotional categoriesFast Fourier transformSpeaker recognition system
The application provides a sentiment asymmetry speaker recognition system based on a generative adversarial network, and a use process of the system comprises the following steps: firstly, converting the voice corresponding to a plurality of sentiment texts into pre-learned multi-sentiment spectrograms through pre-emphasis, framing and windowing and fast Fourier transform; secondly, training the generative adversarial network using the pre-learned multi-sentiment spectrograms, and supervising the conversion between the spectrograms of neutrality and other sentiments in the training process of the generative adversarial network through overall loss; then, converting the spectrogram of neutrality of a registered user into other sentiment spectrograms using the generative adversarial network, and obtaining the user registration multi-sentiment spectrogram by combining the spectrogram of neutrality and other sentiment of the registered user; finally, training a speaker recognition network using the user registration multi-sentiment spectrogram, calculating the speaker classification probability of the voice to be detected, and obtaining the final voiceprint recognition result. The system solves the problem of the decline of the performance of speaker recognition caused by the inconsistency between the registration and the voice sentiment in the actual application scene.
Owner:EAST CHINA UNIV OF SCI & TECH

A defense training method against adversarial samples of a speaker recognition system

PendingCN122337210ASpeaker recognition systemEngineering
This invention relates to the field of speech adversarial example defense, specifically a novel feature fusion defense training method based on speaker facial features, comprising: (1) constructing an audio-video dataset; (2) concatenating audio-video vectors to build a feature fusion module with a multi-head attention mechanism; (3) training the speaker recognition system by retaining only the audio feature portion of the cross-modal fusion features; and (4) generating adversarial examples using FGSM, PGD, CW, FakeBob, SirenAttack, and Kenansville for defense performance testing. This invention introduces speaker facial feature identity consistency information during speaker identity registration to train the speaker recognition system. In the case of unknown attack methods, cross-modal information enhances the robustness and accuracy of the speaker recognition system, achieving defense against adversarial examples of unknown attack methods.
Owner:BEIJING UNIV OF POSTS & TELECOMM +1