Method for solving video question and answer problem by utilizing graph theory-based multi-interaction network mechanism
A video and question technology, applied in the field of video question answering and answer generation, can solve problems such as lack of temporal dynamic information modeling
- Summary
- Abstract
- Description
- Claims
- Application Information
AI Technical Summary
Problems solved by technology
Method used
Image
Examples
Embodiment
[0170] The present invention is verified experimentally on the well-known data sets TGIF-QA, MSVD-QA and MSRVTT-QA. Table 1-Table 3 are the results of training and testing on the three data sets in this embodiment.
[0171] Table 1: Statistics of samples in the TGIF-QA dataset
[0172]
[0173] Table 2: Statistics of samples in the MSVD-QA dataset
[0174]
[0175] Table 3: Statistics of samples in the MSRVTT-QA dataset
[0176]
[0177] In order to objectively evaluate the performance of the algorithm of the present invention, the present invention adopts different evaluation mechanisms for different types of problems. For state transitions, repeated behaviors, single-frame image question answering, classification accuracy (ACC) is used to measure accuracy; for repeated counts, the mean squared error (MSE) between the correct answer and the predicted answer is used.
[0178] The final experimental results are shown in Table 4-Table 6:
[0179] Table 4: Comparison ...
PUM
Abstract
Description
Claims
Application Information
- R&D Engineer
- R&D Manager
- IP Professional
- Industry Leading Data Capabilities
- Powerful AI technology
- Patent DNA Extraction
Browse by: Latest US Patents, China's latest patents, Technical Efficacy Thesaurus, Application Domain, Technology Topic.
© 2024 PatSnap. All rights reserved.Legal|Privacy policy|Modern Slavery Act Transparency Statement|Sitemap