Virtual Presentation Stage for Video Conference Eye Contact

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional video conference platforms hinder effective communication by limiting the speaker's ability to present multiple content items simultaneously and maintain eye contact with the audience, leading to inefficiencies and increased latency due to manual content switching and lack of non-verbal cues.

Innovation Solution

A method to create a combined video stream that overlays a speaker's video stream, content items, and teleprompter notes on a background image, allowing the speaker to present multiple content items seamlessly while maintaining eye contact, with teleprompter notes only visible to the speaker, and providing a unified interface for both the speaker and other participants.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Productivity

If a speaker uses conventional video conference platforms to present content, then the speaker can share content items, but the speaker must manually switch between content items which increases latency and breaks eye contact

Engineering Contradiction:
Improvecontent presentation efficiencyVSAvoidlatency during content switching
Core Design Contradiction:
ProductivityVSLoss of time

Solution Approach 1:

The system merges multiple content items and the speaker's video feed into a single composite video stream. This allows the speaker to present multiple content items simultaneously without manual switching, eliminating latency and maintaining continuous eye contact with the audience.

Inventive Principle:
Principle #5Merging (Combining)

Solution Approach 2:

The system performs preliminary actions by pre-positioning multiple content items and their associated teleprompter notes in the composite video stream before the presentation begins. This allows the speaker to access all content items instantly without interruption, as everything is already prepared and integrated into the unified stream.

Inventive Principle:
Principle #10Preliminary action

2Adaptability or versatility

If a speaker manually switches between content items, then the speaker can present different content, but the speaker loses eye contact with the audience

Engineering Contradiction:
Improvecontent presentation flexibilityVSAvoideye contact maintenance
Core Design Contradiction:
Adaptability or versatilityVSEase of operation

Solution Approach 1:

The system combines the speaker's video feed with multiple content items and teleprompter notes into a single composite stream. The speaker can reference teleprompter notes displayed in the stream while maintaining eye contact with the camera, eliminating the need to look away at separate content windows during manual switching.

Inventive Principle:
Principle #5Merging (Combining)

Solution Approach 2:

The composite video stream acts as an intermediary that integrates teleprompter notes and content items directly into the speaker's field of view. This allows the speaker to access all necessary information without breaking eye contact, as the intermediary stream presents everything in one unified location.

Inventive Principle:
Principle #24Intermediary (Mediator)

3Loss of information

If teleprompter notes are made visible to all participants, then transparency is improved, but the speaker's unique access to notes is compromised

Engineering Contradiction:
Improveinformation transparencyVSAvoidvisibility control
Core Design Contradiction:
Loss of informationVSLength of moving object

Solution Approach 1:

The system applies local quality by making teleprompter notes visible only to the speaker's client device while keeping them hidden from other participants. This selective visibility allows the speaker to access notes privately without other participants seeing them, maintaining both information access and confidentiality.

Inventive Principle:
Principle #3Local quality

Solution Approach 2:

The system segments the video stream into different visibility layers: content items are visible to all participants, while teleprompter notes are segmented as a separate layer visible only to the speaker. This segmentation allows differential visibility control within the same composite stream.

Inventive Principle:
Principle #1Segmentation

4Adaptability or versatility

If multiple content items are presented simultaneously, then presentation versatility is improved, but system complexity increases

Engineering Contradiction:
Improvecontent presentation capabilityVSAvoidvideo stream processing complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The system merges multiple content items, teleprompter notes, and the speaker's video feed into a single composite video stream using video composition techniques. This unified approach simplifies the overall system architecture compared to managing multiple separate streams, as everything is integrated into one stream that can be transmitted and processed as a single unit.

Inventive Principle:
Principle #5Merging (Combining)

Data Source

PatentUS20250097375A1Generating a virtual presentation stage for presentation in a user interface of a video conference
Publication Date: 2025.03.20 GOOGLE LLC
  • US20250097375A1 patent drawing
  • US20250097375A1 patent drawing
  • US20250097375A1 patent drawing

AI summary

Systems and methods for generating a virtual presentation stage for presentation in a user interface of a video conference are provided. A first participant video stream representing a first participant of a plurality of participants of a video conference is received from a camera of a first client device of the first participant. A combined video stream is created comprising a background image, one or more images representing one or more content items presentable by the first participant during the video conference, the first participant video stream, and one or more teleprompter notes associated with at least one of the one or more content items. A user interface (UI) is provided for display on the first client device of the first participant, wherein the UI comprises a visual item corresponding to the combined video stream while the first participant is presenting at least one of the one or more content items to one or more other participants of the video conference.