The invention discloses a video translation method and
system based on
artificial intelligence. The method relates to the technical field of video translation and comprises the following steps of original sound track extraction, target AI
speaker adaptation, AI dubbing generation and
mouth shape synchronization and video synthesis. According to the method, independent audio and video streams are obtained by adopting an audio and video
separation technology, and multiple original sound tracks are extracted through a voice separation model; matching or generating an adaptive target AI speaker module in a preset tone
library; converting the original language voice into a text, translating the text into a target language text, and synthesizing an AI dubbing audio track in combination with a target AI speaker module; and finally, the independent video
stream and the multi-AI dubbing audio track are input into the
mouth shape synchronization model to output a translated video, so that the
timbre fitting degree, the voice quality and the voice consistency of the same speaker of AI dubbing are improved, and meanwhile, the
resource utilization rate of video translation and the
processing efficiency under batch tasks are improved. The problem that in the prior art, video translation is low in quality and efficiency is solved.