AI-based panoramic audio generation method, system, and storage medium
By extracting multimodal features and modeling spherical harmonic functions, combined with adaptive rendering algorithms, high-precision panoramic sound audio is automatically generated, solving the problems of low efficiency and inaccurate positioning in traditional panoramic sound generation, and achieving efficient adaptation and immersive experience in complex acoustic environments.
Patent Information
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- SHANGHAI RUIHEFENG ELECTRONIC TECHNOLOGY CO LTD
- Filing Date
- 2025-12-17
- Publication Date
- 2026-05-26
AI Technical Summary
Traditional panoramic sound audio generation relies on manual intervention, resulting in long production cycles, high costs, and unstable spatial positioning accuracy. Existing AI technology lacks accurate modeling of three-dimensional spatial characteristics and device adaptability, making it unable to effectively reproduce the sound field in complex acoustic environments.
By using multimodal feature extraction, spherical harmonic function modeling, and adaptive rendering algorithms, the system achieves automated fusion and conversion of the original audio signal and scene parameters, constructs a high-precision 3D spatial sound field model, and adaptively renders multi-channel panoramic audio based on the target device parameters.
It achieves efficient and high-precision adaptive generation of panoramic audio, solving the problems of low efficiency, high cost and insufficient spatial positioning accuracy in traditional methods, and providing an immersive auditory experience.
Smart Images

Figure CN121547723B_ABST