The invention discloses a multi-camera collaborative single-target tracking method and device, belongs to the technical field of
computer vision target tracking, and is suitable for scenes such as intelligent monitoring and public safety. The method comprises the following steps: constructing a camera space matrix and defining a view field range; the recognition robustness is improved by fusing the target face and the body features; tracking and constructing a target trajectory based on the fusion features; a missing frame is reconstructed or a future frame is predicted through a forward
diffusion-reverse generation model, and trajectory coherence is optimized in combination with space-time constraints; probability graph loss optimization of a
fixation point area is realized by using an
encoder and a saliency prediction module, and finally a video conforming to
fixation point constraints is jointly generated. The device comprises a video frame acquisition and preprocessing module, a
feature extraction and fusion module, a target tracking module, a missing frame reconstruction and future frame prediction module, a
fixation point positioning module, a joint optimization module and a
video output module, and all the modules cooperatively realize end-to-end tracking. According to the method, the shielding scene recognition capability is enhanced through multi-
modal feature fusion, the problem of track breakage is solved by means of a
diffusion model, the monitoring requirement of a specific area is met in combination with fixation point optimization, the precision, continuity and practicability of target tracking in a multi-camera environment are effectively improved, and the method has remarkable technical advantages and application value.