Wide-Angle Video Audio Suppression for Clear Main Speaker Sound
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current technologies lack effective means to improve sound reception quality for conference devices that capture wide-angle images and corresponding audio signals.
Innovation Solution
A video content providing method and device that obtain a wide viewing angle image stream and corresponding audio content, determine regions of interest, select candidate regions, suppress audio components not corresponding to designated regions, and integrate the processed audio and video into specific video content for improved sound quality.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Area of stationary object
If a wide-angle lens is used to capture wide viewing angle images, then the imaging range is improved, but the sound reception quality deteriorates
Solution Approach 1:
The patent segments the audio signal into multiple components corresponding to different sound source directions, and segments the video frame into multiple regions of interest. By processing audio and video segments separately and integrating them, the system achieves directional audio precision while maintaining wide-angle video coverage.
Solution Approach 2:
The patent applies different quality standards to different regions: high audio quality is applied to sound sources within regions of interest, while audio from other directions is suppressed. This local quality approach allows the system to maintain wide-angle imaging while providing focused high-quality audio for specific areas.
2Loss of information
If audio from all directions is captured in wide-angle video, then the completeness of audio coverage is improved, but the clarity of main speaker audio deteriorates
Solution Approach 1:
The patent extracts audio components from specific sound source directions and extracts regions of interest from the video frame. By taking out only the relevant audio-video pairs and integrating them, the system maintains complete audio coverage while ensuring clear audio for main speakers in the designated regions.
Solution Approach 2:
The system dynamically identifies regions of interest and their corresponding sound source directions in real-time, adjusting which audio components are enhanced and which are suppressed based on the current video content and speaker positions.
Data Source
AI summary
A video content providing method and a video content providing device are provided. The method includes the following. A wide viewing angle image stream and a corresponding first audio content are obtained. A plurality of regions of interest in the wide viewing angle image stream are determined, and candidate regions in the regions of interest are integrated into a first frame. A designated region is selected from the candidate regions, and a corresponding first audio component are found from the first audio content. Each first audio component is suppressed to adjust the first audio content into a second audio content. The first frame and the second audio content are integrated into a specific video content, and the specific video content is provided.


