The invention relates to a high-performance voice
processing method, and the method comprises the steps: obtaining
voice data transmitted through a network, generating corresponding scene
fingerprint information, determining a frame parameter of a lightweight multi-scene
adaptation frame, and a corresponding
processing parameter, and if the duration of a lost segment is smaller than a preset threshold value, determining that the segment is lost. If yes, generating a corresponding first compensation result according to the acoustic compensation parameter; if the duration is larger than or equal to a preset threshold value, whether the acoustic compensation process and the semantic
processing process are performed in parallel or not is judged according to the frame parameters, if not, a corresponding second compensation result is generated, and if yes, a corresponding second compensation result is generated according to the acoustic compensation process and the semantic processing process; and outputting the corresponding target voice transmission data. The definition and the stability of the voice
signal in a complex environment can be effectively improved, different compensation
modes are flexibly switched or executed in parallel according to the
packet loss duration and the resource state, and the effectiveness and the stability of a
recovery mechanism are improved.