This application proposes a speech decoding
processing method, apparatus, electronic device, and storage medium. When a blank decoding candidate exists among all decoding candidates corresponding to the current audio acoustic output of an
audio recognition system, the method determines whether the current audio acoustic output is a blank output. If it cannot be determined that the current audio acoustic output is blank, this application does not employ a
language model where each node has a self-
jumping arc for self-
jumping. Instead, it creates a new activation arc for self-
jumping on the active node with blank decoding candidates in a pre-defined language decoding network, and adds blank decoding candidate information as a
label to the activation arc. Based on the activation arc
label, the path
label of the active node is determined, thereby reducing the memory usage of the self-jumping arc portion. Based on this, this application reduces the memory and computing power requirements of
speech recognition, mitigating the limitations of
speech recognition technology in practical applications.