The invention discloses a
digital human explanation and display control method based on voice recognition driving, which relates to the technical field of voice recognition and comprises the following steps of: receiving an original voice
signal input by a user, preprocessing the original voice
signal, and sending the preprocessed voice
signal into a voice recognition path and a style analysis path in parallel; generating a user
question text and a real style embedding vector; integrating the structured knowledge data, the prospective explanation theme and the matching style parameter set through a multi-
modal fusion mechanism, generating an explanation text sequence, and synthesizing a voice audio
stream; and inputting the voice audio
stream and the matching style parameter set into a digital person for driving, generating a non-linguistic behavior sequence, rendering the non-linguistic behavior sequence in real time through a content synchronization mechanism and a multi-terminal linkage protocol, and outputting a digital person explanation video
stream through display trigger logic. According to the method, voice recognition and style analysis are processed in parallel, and a
dynamic feature interaction gating mechanism is adopted, so that deep optimization of semantic content and style features in voice signals is realized.