Audio data processing method and device, electronic equipment and storage medium
By acquiring the Mel spectrum of audio data and using a residual denoising diffusion model to predict high-frequency features, the problem of audio quality degradation in existing technologies is solved, achieving higher audio resolution and better listening experience in audio data processing.
Patent Information
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- BEIJING ZITIAO NETWORK TECH CO LTD
- Filing Date
- 2024-10-31
- Publication Date
- 2026-05-01
AI Technical Summary
Existing technologies, in scenarios such as voice calls, video conferencing, and live video streaming, compress audio data to reduce the overhead of device and network resources, resulting in a decline in audio quality and affecting the user's listening experience.
By acquiring the Mel spectrum data of the initial audio data, the high-frequency features are predicted using the residual denoising diffusion model in the band extension module to generate enhanced Mel spectrum data. This enhanced Mel spectrum data is then processed by the audio restoration module to generate optimized audio data with higher audio resolution.
It improves the audio resolution of audio data, enhances sound details and texture, and improves the user's listening experience.
Smart Images

Figure CN121963753A_ABST