H.248 Media Stream Grouping for Conference Bandwidth Optimization
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current multimedia conference systems face challenges in dynamically managing media streams, particularly in selecting the appropriate video streams for conference participants based on voice activity, leading to inefficient bandwidth usage and transcoding efforts.
Innovation Solution
The system employs voice activity detection and policy-based management of media streams using the H.248 protocol, where media streams from the same sender are grouped and processed within the same termination, and policies are applied to decide on forwarding or discarding streams based on voice activity, optimizing stream selection and destination for each participant.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If all media streams from a sender are forwarded to all participants, then complete media distribution is achieved, but bandwidth consumption and processing load increase significantly
Solution Approach 1:
The patent applies local quality by differentiating stream handling based on voice activity status. When a participant is actively speaking (voice activity detected), their media streams are forwarded to all participants with high priority. When not speaking, stream forwarding is restricted or discarded. This selective quality adjustment optimizes bandwidth usage while maintaining essential media distribution.
Solution Approach 2:
The system dynamically adjusts media stream forwarding decisions based on real-time voice activity detection results. The policy applied to media streams changes dynamically according to whether the sender is currently speaking, enabling adaptive bandwidth management that responds to actual conference conditions rather than using static forwarding rules.
2Adaptability or versatility
If multiple video streams are sent to each participant, then video quality and options are improved, but device processing load and transcoding efforts increase
Solution Approach 1:
The patent applies preliminary action by pre-grouping media streams from the same sender into identified subgroups before distribution. The system prepares and organizes stream subgroups in advance, so that when participants receive streams, the grouping information is already available. This reduces real-time processing complexity and transcoding efforts at the receiving end.
Solution Approach 2:
The patent segments media streams into distinct subgroups based on their origin (same sender). By dividing the overall media distribution into identifiable subgroups, the system enables more efficient processing and selection at the recipient side, reducing the complexity of handling all streams uniformly while maintaining the ability to provide multiple video stream options.
3Ease of operation
If media streams are processed individually, then precise control over each stream is achieved, but control signaling overhead and processing complexity increase
Solution Approach 1:
The patent merges control by applying policies to subgroups of media streams from the same sender rather than handling each stream completely independently. This grouping approach reduces control signaling overhead by enabling batch or collective control decisions for related streams while still maintaining the ability to apply precise control where needed through the policy framework.
Data Source
Figure 1~2
Figure 3~5
Figure 6
AI summary
It is provided a method, comprising detecting if a first signaling indicating the desire to send plural first media streams including anaudio stream is received from a sender; informing the resource function processor that at least a subgroup of the first media streams including the audio stream originates from the sender if the first signaling is received from the sender; instructing a resource function processor to perform voice activity detection on the audio stream if the first signaling is received from the sender; instructing the resource function processor to apply a policy on the subgroup of the first media streams, wherein the policy includes passing or discarding at least some of the first media streams of the subgroup and/or selecting destinations for the media streams depending on a result of the voice activity detection on the audio stream.