The invention provides a voice recognition method and
system based on an AI
large model, and belongs to the technical field of voice recognizing.The method comprises the steps that firstly,
noise reduction and
frequency band enhancement in a
noise environment are achieved through a pre-trained anti-
noise suppression network, and a basis is provided for follow-up
processing in combination with multi-dimensional spectrum quality scores; based on
similarity matching of element feature vectors and a pre-constructed dialect thermodynamic diagram
library and matching
weight adjustment based on
frequency spectrum quality scores, accurate modeling of specific dialect pronunciation deviation is achieved, moreover, through an acoustic
adaptation matrix and a
language model adaptation matrix generated through a super network, the adaptability of the model to different dialects is improved, and in addition, the accuracy of the model is improved. According to the method, a fusion thermodynamic diagram and an acoustic
adaptation matrix are jointly injected into a pre-trained
acoustic model, the recognition accuracy of dialect phonemes is improved through multi-level attention correction, and finally, accurate
speech recognition is achieved by adopting a thermodynamic diagram guided cluster
search algorithm and combining
verification of an adversarial discrimination network.