The invention discloses a multi-agent
perception and
cognition generation and API (Application Program Interface) calling multi-
modal evaluation method. The method comprises the following steps of: 1, downloading and deploying a standardized
data set to the local; 2, evaluating the
perception ability of the multi-agent, including perceiving ten sub-tasks, and calculating scores by using the single-graph single-question accuracy and the single-graph double-question full-
match rate as
perception indexes; 3, calculating scores by adopting the single-graph single-question accuracy rate and the single-graph double-question full-
match rate as cognitive indexes, and outputting MME basic scores in combination with the perception scores; 4, generating and evaluating a multi-agent picture, extracting a target keyword, detecting a to-be-evaluated image to obtain an actual object category set, and calculating an image element matching
score in combination with matching of the target keyword and an object category; 5, executing multi-agent API calling validity evaluation, and outputting an API calling matching degree
score; and 6, multi-dimensional scores are fused, a comprehensive
evaluation result is output, and evaluation of the multi-
modal multi-agent
large model is realized.