The invention relates to the technical field of
digital human interaction, in particular to a
digital human intelligent question and answer
interaction method for a multi-
modal visual platform, which comprises the following steps of: receiving multi-
modal input of a user, performing
privacy protection processing, and performing sensitive field identification, desensitization and hierarchical storage; then semantic analysis and interaction intention recognition are conducted on the processed input, a retrieval request is generated,
relevant information is searched in a professional
knowledge base based on the retrieval request, meanwhile, source identifiers of entries are reserved, a
retrieval result and a user intention are input into a question and answer generation model, and an interaction answer is generated; and adding a corresponding source identifier to the answer to realize content
traceability, and finally performing multi-
modal output, including voice broadcast,
expression action and visual display, through a
digital human image, so as to realize safe, credible and intuitive interactive experience.