Media Content Manager for Conference Visual Composition

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

In multimedia conference systems, it is challenging to effectively display all participants simultaneously due to limited display space, leading to confusion and difficulty in identifying the active speaker, especially as the number of participants increases.

Innovation Solution

A media content manager component that selectively displays GUI views of actively speaking participants by using a video decoder module, active speaker detector, media stream manager, and media selection module to prioritize and replace participants based on speech activity, ensuring that only active speakers are prominently displayed.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Loss of information

If all participants are displayed simultaneously in the virtual meeting environment, then complete information about all participants is provided, but visual clutter increases and it becomes difficult to identify the active speaker

Engineering Contradiction:
Improveinformation completenessVSAvoidvisual clarity
Core Design Contradiction:
Loss of informationVSEase of operation

Solution Approach 1:

The system extracts and displays only the most relevant information (active speaker) prominently in the main view, while other participant information is made available through alternative means such as the participant roster or thumbnail views. This selective extraction resolves the contradiction by preventing visual clutter while maintaining information accessibility.

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

Different regions of the display interface have different qualities and purposes: the main view provides high-quality, large-scale display of the active speaker for clear identification, while the participant roster provides comprehensive but condensed information about all participants. This spatial differentiation of information quality resolves the contradiction between completeness and clarity.

Inventive Principle:
Principle #3Local quality

2Loss of information

If the number of displayed participants is increased to show all participants, then information completeness is improved, but the complexity of the display system increases

Engineering Contradiction:
Improveparticipant visibilityVSAvoiddisplay complexity
Core Design Contradiction:
Loss of informationVSDevice complexity

Solution Approach 1:

The display system is segmented into multiple functional components: a main view for the active speaker, a participant roster for comprehensive participant information, and potentially thumbnail views for other active participants. This segmentation allows the system to manage complexity by organizing information into distinct, manageable sections rather than attempting to display all participants in a single complex view.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The system transitions from a two-dimensional grid layout to a hierarchical organization where participant information is arranged in a roster format with different levels of detail. This dimensional change allows comprehensive participant visibility without increasing the visual complexity of the main display area, as the roster provides structured, scrollable access to all participants.

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

3Loss of information

If multiple participants are displayed in the virtual meeting environment, then information completeness is improved, but the difficulty of identifying the active speaker increases

Engineering Contradiction:
Improveparticipant informationVSAvoidactive speaker identification
Core Design Contradiction:
Loss of informationVSDifficulty of detecting and measuring

Solution Approach 1:

The system uses visual differentiation through color coding, borders, or highlighting to distinguish the active speaker from other participants in the display. The active speaker's video feed or name is emphasized with distinct visual markers, making them easily identifiable among multiple displayed participants without requiring complex visual search.

Inventive Principle:
Principle #32Color changes

Solution Approach 2:

The system proactively identifies and prepares the active speaker's information for display before it becomes necessary to show it. By continuously monitoring speech activity and pre-positioning the active speaker's video feed or name in the prominent main view, the system eliminates the need for users to search through multiple participant displays to identify who is speaking.

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS8316089B2Techniques to manage media content for a multimedia conference event
Publication Date: 2012.11.20 MICROSOFT TECHNOLOGY LICENSING LLC
  • US8316089B2 patent drawing
  • US8316089B2 patent drawing
  • US8316089B2 patent drawing

AI summary

Techniques to manage media content for a multimedia conference event are described. An apparatus may comprise a media content manager component operative to generate a visual composition of decoded media streams for a multimedia conference event. If it is determined that the total number of decoded media streams is greater than the total number of available display frames in a visual composition then an active group of decoded media streams may be selected from among the total number of decoded media streams for mapping to the available display frames based on speech activity. Other embodiments are described and claimed.