Selective Audio Channel Mixing for Teleconference Processing
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Multi-party teleconferencing systems face challenges such as audio signal saturation, distortion, and increased processing requirements due to the complexity of mixing audio signals from multiple participants, leading to poor quality and resource-intensive processing.
Innovation Solution
An apparatus and method that estimate channel quality parameters for encoded audio signals without decoding them, select the best signals based on these parameters, and combine only the selected signals to reduce processing power and improve audio quality.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If all audio signals from all participants are decoded and mixed, then complete audio coverage is achieved, but processing resources increase proportionally to the number of users
Solution Approach 1:
The system decodes and mixes only a subset of audio signals from participants rather than all participants. The processor selects active channels based on activity thresholds and mixes only those selected channels, reducing processing resources while maintaining adequate audio coverage for the teleconference.
2Adaptability or versatility
If all audio signals from all participants are summed, then complete participant coverage is achieved, but audio signal saturation and distortion occur
Solution Approach 1:
The system selectively mixes only active audio channels below a saturation threshold. The processor monitors audio signal levels and excludes channels that would cause saturation when summed, preventing distortion while maintaining coverage of active participants.
Solution Approach 2:
The system extracts and removes inactive or silent channels from the mixing process. By identifying and excluding channels with insufficient activity or background noise only, the system prevents saturation while maintaining audio quality.
3Adaptability or versatility
If audio signals from silent participants are included, then complete participant inclusion is achieved, but background noise level increases
Solution Approach 1:
The system includes only participants whose audio signals exceed a minimum activity threshold. By setting this threshold, the system excludes silent participants who would contribute only background noise, while still including active participants to maintain conversation quality.
Data Source
Figure 1
Figure 2
Figure 3
AI summary
An apparatus comprising an ingress port configured to receive a signal comprising a plurality of encoded audio signals corresponding to a plurality of sources; and a processor coupled to the ingress port and configured to calculate a parameter for each of the plurality of encoded audio signals, wherein each parameter is calculated without decoding any of the encoded audio signals, select some, but not all, of the plurality of encoded audio signals according to the parameter for each of the encoded audio signals, decode the selected signals to generate a plurality of decoded audio signals, and combine the plurality of decoded audio signals into a first audio signal.