Audio encoding and decoding method and device

By selecting a target virtual speaker from a virtual speaker set to generate a virtual speaker signal and a residual signal for encoding, the problems of large data volume and high bandwidth occupancy in multi-channel encoding and decoding methods are solved, and a more efficient encoding and decoding process is achieved.

CN114582357BActive Publication Date: 2025-09-12HUAWEI TECH CO LTD +1
2 Cites 0 Cited by

Patent Information

Application Number
CN202011377433.0
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2020-11-30
Publication Date
2025-09-12
Estimated Expiration
2040-11-30

AI Technical Summary

Technical Problem

Existing multi-channel encoding and decoding methods require adapting the codec according to the number of channels of the original scene audio signal, resulting in large data volume and high bandwidth occupancy, making it difficult to effectively reduce the amount of encoding and decoding data.

Method used

By selecting the target virtual speaker from the virtual speaker set, generating a virtual speaker signal and a residual signal, and encoding them, the direct encoding of the original scene audio signal is reduced. The encoding data volume of the virtual speaker signal and the residual signal is independent of the number of channels, thereby improving the encoding efficiency.

Benefits of technology

It effectively reduces the amount of encoded data and improves encoding and decoding efficiency, while ensuring the encoding quality and the audio signal reconstruction effect at the decoding end.

✦ Generated by Eureka AI based on patent content.
Patent Text Reader

Abstract

The present application discloses an audio encoding and decoding method and apparatus for reducing the amount of encoding and decoding data to improve encoding and decoding efficiency. The present application provides an audio encoding method, comprising: selecting a first target virtual speaker from a preset virtual speaker set based on a first scene audio signal; generating a first virtual speaker signal based on the first scene audio signal and attribute information of the first target virtual speaker; obtaining a second scene audio signal using the attribute information of the first target virtual speaker and the first virtual speaker signal; generating a residual signal based on the first scene audio signal and the second scene audio signal; encoding the first virtual speaker signal and the residual signal, and writing them into a bitstream.
Need to check novelty before this filing date? Find Prior Art