Cascade Conference Audio Orientation Alignment
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
In cascade video conferences, the orientation of audio data often does not correspond with the image orientation at different conference sites, leading to a suboptimal user experience due to inconsistent voice and image alignments across screens.
Innovation Solution
A method and apparatus that receive and process audio streams from cascade and non-cascade conference sites, selecting and adjusting audio data to ensure one-to-one correspondence between image and voice orientations by sending audio streams through different audio channels or cascade channels, allowing individual orientation adjustments for each site.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Loss of energy
If audio data from cascade conference sites is mixed and transmitted through a single cascade channel, then bandwidth consumption is reduced, but the orientation of audio data cannot correspond with image orientation at different conference sites
Solution Approach 1:
The patent segments the mixed audio data into separate audio streams for each conference site, transmitting them through different audio channels. This allows each conference site to receive and process audio data independently, enabling proper orientation correspondence between audio and video while maintaining efficient bandwidth utilization through selective transmission.
Solution Approach 2:
The patent applies local quality by allowing different conference sites to have different audio orientation configurations. Each conference site can independently adjust audio orientation parameters according to its specific display layout and spatial requirements, rather than using a uniform orientation for all sites.
2Device complexity
If audio data orientation is adjusted globally for all conference sites, then processing complexity is reduced, but individual conference sites cannot achieve accurate voice-image alignment
Solution Approach 1:
The patent implements dynamic audio orientation adjustment where each conference site can independently modify its audio orientation parameters based on its specific requirements. The system dynamically adapts audio orientation for each site rather than using a static global configuration, enabling accurate voice-image alignment while maintaining manageable processing complexity through automated parameter distribution.
3Ease of operation
If separate audio channels are used for each conference site, then accurate orientation adjustment is achieved, but bandwidth consumption increases
Solution Approach 1:
The patent applies partial action by transmitting audio data for only the most relevant conference sites through separate channels, rather than transmitting all audio data to all sites. The system selectively transmits audio streams based on spatial relationships and conference site configurations, achieving accurate orientation adjustment where needed while conserving bandwidth for less critical transmissions.
Data Source
Figure 1~2
Figure 3
Figure 4
AI summary
Embodiments of the present invention discloses a method for processing cascade conference sites in a cascade conference, which is used to implement one-to-one correspondence between an image orientation and a voice orientation of each conference site in the cascade conference, and improve the user experience of a participant. The method of the embodiments of the present invention includes: receiving an audio code stream sent by a cascade conference site, where the audio code stream sent by the cascade conference site is sent based on that different conference sites occupy different audio sound channels or audio cascade channels; receiving an audio code stream sent by a non-cascade conference site; selecting audio data satisfying a preconfigured condition from audio data to be selected,