A multi-modal sentiment analysis method and system fusing image caption and BERT
CN117115516BActive Publication Date: 2026-08-28NANJING UNIV OF POSTS & TELECOMM
View PDF 0 Cites 0 Cited by
Patent Information
- Application Number
- CN202310989477.6
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- Filing Date
- 2023-08-08
- Publication Date
- 2026-08-28
- Estimated Expiration
- 2043-08-08
AI Technical Summary
Technical Problem
[0005]因此,本发明解决的技术问题是:现有的多模态情感分析存在单模态具有局限性,精准度低,稳定性低,以及结合社交媒体上的图片、视频等富文本信息来分析用户的情感倾向的问题
✦ Generated by Eureka AI based on patent content.
Smart Images

Figure CN117115516B_ABST
Abstract
The application discloses a kind of multi-modal sentiment analysis method and system of fusion image caption and BERT, it is related to multi-modal sentiment analysis technical field including extracting the image generation description image of multi-modal data set;Image feature is obtained by ResNet, the hidden layer representation of target is calculated using BERT encoder, and the final visual representation is obtained based on target image matching layer;Image description and text of multi-modal data are input BERT encoder to calculate the feature representation of image description and the feature representation of text, and the final text feature representation is obtained;Calculate multi-modal hidden layer representation, and the final sentiment classification is obtained by pooling layer fully connected layer and Softmax.This application extracts multi-modal data set image, realizes cross-modal sentiment analysis by combining more than two modalities, increases the understanding and recognition ability to image content, effectively solves the limitation of single mode, converts the information of image into visual representation with more expressive ability and semantic information, and improves the stability of multi-modal system.
Need to check novelty before this filing date? Find Prior Art