Participant Importance Word Cloud in Video Sessions

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

In multi-device video communication sessions, there is a need to represent the assessments of participant importance effectively, as existing methods lack a systematic way to weight and visualize the contributions of participants.

Innovation Solution

A method and system that processes recorded content to detect vocal expressions, generates text elements, and creates a word cloud based on these elements, incorporating participant ratings as a weighting factor to reflect the importance of participants, thereby visualizing the session content.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If participant ratings are incorporated into word cloud generation, then the representation of participant importance is improved, but the system complexity increases

Engineering Contradiction:
Improveparticipant importance assessmentVSAvoidsystem complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The system segments the word cloud generation process by creating separate processing streams: one for detecting vocal expressions and generating text elements, and another for processing participant ratings. These segmented processes are then integrated through a weighting mechanism that combines text element frequency with participant importance ratings, resolving the contradiction by organizing complexity into manageable segments while achieving precise importance assessment.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent introduces an intermediary weighting mechanism that mediates between the raw text elements from vocal expressions and the final word cloud representation. This intermediary layer processes both the frequency of text elements and the participant ratings, combining them through a weighting algorithm that produces a nuanced representation of participant importance without requiring direct complex interaction between all system components.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Measurement precision

If vocal expression detection is performed on recorded content, then the accuracy of content analysis is improved, but the processing time increases

Engineering Contradiction:
Improvecontent analysis accuracyVSAvoidprocessing time
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

The system performs preliminary action by detecting vocal expressions and generating text elements from recorded content before the word cloud generation process. This preliminary processing organizes the raw audio data into structured text elements that can be more efficiently weighted and displayed, reducing the computational burden during the final word cloud generation phase while maintaining high analysis accuracy.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent applies partial action by selectively processing only the vocal expressions that meet certain criteria for inclusion in the word cloud. Rather than analyzing every single vocal expression with equal depth, the system identifies and processes the most significant expressions based on their relevance to participant importance, achieving accurate content analysis while reducing overall processing time through selective attention to key elements.

Inventive Principle:
Principle #16Partial or excessive action

Data Source

PatentUS8654942B1Multi-device video communication session
Publication Date: 2014.02.18 GOOGLE LLC
  • US8654942B1 patent drawing
  • US8654942B1 patent drawing
  • US8654942B1 patent drawing

AI summary

A method of multi-device video communication. A server receives recorded content from a multi-device video communication session and processes the recorded content to detect vocal expressions from a plurality of participants. The server generates a plurality of text elements each corresponding to one or more of the vocal expressions. The server receives at least one rating for at least one participant of the plurality of participants and generates a word cloud based on the plurality of text elements and at least in part on the at least one rating for the at least one participant.