The invention relates to the technical field of AI information, in particular to a multi-
modal information comparison method based on semantic guidance, which comprises the following steps: S1,
data acquisition and preprocessing; s2, constructing a model; s3, model training; s4, evaluating and optimizing the model; and S5, carrying out online prediction and feedback. Through the designed and constructed
algorithm architecture, heterogeneous image information comparison of specific information is realized through
image processing by means of the text semantic guide model. Specific
information searching and comparison of the heterogeneous images are realized, and the complexity of searching and comparison of the specific information in the heterogeneous images is greatly reduced. According to the method, the same
semantic information of different images is searched and compared by analyzing the
semantic information of the characters, so that the
usability is improved. The
layout distribution, the information format and the like of the to-be-compared image are allowed to be inconsistent, and the application scene of the model is widened. The influence of human factors is reduced, the production efficiency is improved, and the probability of information error comparison is reduced.