The invention relates to the technical field of sign language intelligent communication, in particular to a deaf-mute sign
language translation pronunciation
system which comprises a
data acquisition and preprocessing module, a sign
language model recognition module, a dynamic gesture model module and a voice conversion module. According to the sign
language translation pronunciation
system for the deaf-mute, wide-angle camera intelligent glasses are adopted for visual collection, wearable equipment is not needed, video streams and
thermodynamic diagrams are combined, visual information and key point probability distribution are complementary, the influence of single-mode
noise is reduced, video space-time features are extracted through a 3D CNN, long-range dependence is captured in combination with a self-attention mechanism of Transform, and therefore the sign
language translation pronunciation
system for the deaf-mute is obtained. The sign
language recognition precision is improved, an internal reference matrix and a
distortion coefficient are calculated through a calibration board, wide-angle lens
distortion is corrected, it is ensured that hand key points are accurately positioned, YOLOv8 is adopted to segment an interference object,
background noise is prevented from affecting recognition, 21 key point
thermodynamic diagrams are generated through MediaPipe,
fault tolerance of low-confidence-coefficient key points is enhanced through
Gaussian kernel
diffusion, and the recognition accuracy is improved. And the model identification precision is improved.