A speech recognition method and device for the power dispatching field, an electronic device, and a storage medium

By combining multi-stage iterative training with a speech recognition model generated from power dispatch terminology and scene noise, the problems of low accuracy and poor anti-interference ability in the field of power dispatching have been solved, achieving higher recognition accuracy and robustness.

CN122417015APending Publication Date: 2026-07-17POWER DISPATCHING CONTROL CENT OF GUANGDONG POWER GRID CO LTD
View PDF 0 Cites 0 Cited by

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
POWER DISPATCHING CONTROL CENT OF GUANGDONG POWER GRID CO LTD
Filing Date
2026-04-30
Publication Date
2026-07-17

AI Technical Summary

Technical Problem

Existing speech recognition technologies face problems such as low accuracy in recognizing technical terms and poor anti-interference capabilities in the field of power dispatching. This is mainly due to the severe lack of coverage of technical terms and complex noise distribution in the training data, which makes the model prone to omissions or misrecognitions in real complex scenarios.

Method used

By acquiring historical real dispatch speech and scene-enhanced synthesized speech as training samples, and combining them with power dispatch terminology, multi-stage iterative training is carried out, including parameter optimization of encoder, predictor and decoder, to generate target speech recognition model and enhance its adaptability to professional terminology and complex noise.

Benefits of technology

It significantly improves the recognition accuracy and anti-interference robustness of the target speech recognition model in real business scenarios, solves the problems of omission and misrecognition of professional terms, and improves the recognition accuracy of scheduling instructions.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN122417015A_ABST
    Figure CN122417015A_ABST
Patent Text Reader

Abstract

本发明公开了一种面向电力调度领域的语音识别方法、装置、电子设备及存储介质,属于语音识别技术领域,所述方法包括:获取待识别电力调度语音并输入目标语音识别模型,输出调度指令文本;所述模型通过历史真实调度语音及其转写文本,以及基于专业术语生成的场景增强合成语音及对应文本构建训练样本;基于上述样本对初始模型进行多阶段迭代训练,在每次迭代中提取声学特征并生成预测文本序列,计算预测文本与标签之间的差异得到训练损失,并据此更新模型参数,直至满足预设条件,获得目标语音识别模型。通过实施本发明,能够解决现有技术存在的语音识别模型由于专业术语以及复杂噪声训练数据匮乏,导致对电力调度专业术语容易出现错识的问题。
Need to check novelty before this filing date? Find Prior Art