A dialogue summary and key intention extraction method, system and device based on a large model and multi-dimensional acoustic features

CN122157662APending Publication Date: 2026-06-05杭州智慧沟通智能科技有限公司

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
杭州智慧沟通智能科技有限公司
Filing Date
2026-05-08
Publication Date
2026-06-05

Smart Images

  • Figure CN122157662A_ABST
    Figure CN122157662A_ABST
Patent Text Reader

Abstract

The application discloses a dialogue summary and key intention extraction method, system and equipment based on a large model and multi-dimensional acoustic characteristics, and belongs to the technical field of artificial intelligence. The method comprises the following steps: obtaining dialogue audio and identifying to obtain a text Token sequence with a timestamp; a multi-dimensional acoustic feature vector is extracted, and a multi-modal feature matrix is constructed; the matrix is input into a large language model, the acoustic features are mapped into Key and Value matrices, the attention bias term is calculated and superimposed into the model self-attention weight matrix, and the dynamic adjustment of the attention of the text Token sequence is realized; and the dialogue summary and the key intention list are decoded and output based on the adjusted weight. The application proposes an attention bias injection mechanism, realizes the underlying fusion of acoustic features and text semantics through the attention layer, effectively solves the feature loss problem in the voice dialogue, and significantly improves the accuracy and robustness of intention recognition.
Need to check novelty before this filing date? Find Prior Art