一种用于调度电话的声纹识别处理方法

By decoupling voiceprint features using deep neural networks and orthogonal projection operators of channel subspace basis vector groups, the nonlinear distortion problem introduced by low bit rate encoding and decoding in dispatch telephones is solved, achieving stable and efficient identity recognition under multi-protocol channels.

CN122177124BActive Publication Date: 2026-07-17INFORMATION & COMM CO OF STATE GRID SHAANXI ELECTRIC POWER CO LTD +1

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
INFORMATION & COMM CO OF STATE GRID SHAANXI ELECTRIC POWER CO LTD
Filing Date
2026-05-11
Publication Date
2026-07-17

AI Technical Summary

Technical Problem

Existing technologies struggle to effectively decouple the nonlinear distortions introduced by low bit-rate encoding and decoding in dispatching telephones, leading to irreversible collapse of voiceprint features in the manifold space and resulting in decreased recognition performance. In particular, it is difficult to maintain feature stability and purity under multi-protocol channels and short voice conditions.

Method used

A deep neural network model is used in conjunction with spectral entropy and channel subspace basis vectors. Orthogonal projection operators are used to decouple voiceprint features, eliminate channel distortion, and endpoint detection logic is used to filter out non-human voice segments. A dual orthogonal feature cleaning mechanism is constructed to ensure feature purity.

Benefits of technology

Maintaining the stability of voiceprint features under multi-protocol and extreme channel conditions ensures that recognition performance does not degrade, achieving compactness and accuracy in cross-channel identity recognition, and adapting to full-spectrum interference in complex scheduling scenarios.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN122177124B_ABST
    Figure CN122177124B_ABST
Patent Text Reader

Abstract

本发明涉及语音信号处理技术领域,公开了一种用于调度电话的声纹识别处理方法,包括:采集调度通话的数字语音信号并转换为频域声学特征序列,利用深度神经网络模型统计聚合生成混合特征向量;计算各频带谱熵值并构建特征性权重矩阵,利用权重矩阵对混合特征向量执行加权乘法,得到加权声纹向量;利用预置信道子空间基向量组构建正交投影算子;将加权声纹向量与正交投影算子执行矩阵乘法运算,在滤除信道投影分量同时利用权重分布阻断高谱熵频带噪声扩散,得到解耦声纹向量;计算相似度并输出指令,本发明通过谱熵物理门控与代数几何投影的协同,解决非线性量化畸变导致的特征流形缩问题。
Need to check novelty before this filing date? Find Prior Art