The application discloses a multi-agent multi-
modal behavior cognitive screening and
risk assessment system and method, comprising: collecting and preprocessing
gait, facial video and voice signals of a subject; respectively performing feature analysis on the preprocessed
gait, facial video and voice signals to generate corresponding
gait, facial and audio structured evidence units; calculating cross-
modal evidence consistency scores and evidence conflict degrees according to a multi-
modal structured evidence set, performing cross-modal semantic fusion and multi-domain risk reasoning, generating a comprehensive behavior risk result and sub-domain risk results in the emotion domain, the cognitive domain and the social domain; when the
risk assessment result reaches a preset trigger condition, introducing human-in-the-loop review to confirm or correct the
risk assessment result; and generating the fusion reasoning result and the human-in-the-loop review result as a standardized assessment report. The application improves the stability, safety and explainability of risk assessment in the case of missing modalities, local anomalies or inconsistent multi-modal conclusions.