The present invention provides a sending and receiving method and device for three-dimensional
point cloud-assisted video semantic communication. The sending method comprises: a transmitting end acquires a video, generates a
mask image set for the first few frames, and transmits it to a receiving end; estimates the camera
pose corresponding to the
mask image set for the first few frames, and generates a three-dimensional
point cloud. A
mask image set is generated for the remaining video frames, and the projection of the corresponding three-dimensional
point cloud at the corresponding camera
pose is calculated. This projection, the mask image set for the remaining frames, and the semantic priority are input into an
image compression model to obtain a residual
semantic vector; and the residual
semantic vector and other information are transmitted to a receiving end. The receiving method comprises: receiving the mask image set for the first few frames to generate a three-dimensional point cloud. The residual compressed
semantic vector and other information are received, and the projection of the three-dimensional point cloud corresponding to the remaining frames at the corresponding camera
pose is calculated. The projection, the residual compressed semantic vector, and the semantic priority are input into an
image decompression model to restore the mask image set for the remaining frames. All mask image sets are synthesized to obtain a complete video. The present invention achieves efficient and high-quality transmission of video data through three-dimensional point cloud-assisted video semantic communication and joint coding of
communication source and channel.