Multi-User Streaming Video Interfaces with Hand Gesture Mapping

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Collaborative editing environments lack the capability to use gesture recognition as input, leading to inadequate user experience and increased computing resource and network bandwidth consumption.

Innovation Solution

A computing device uses a glass board application that tracks hand motions and gestures via camera, recognizing drawing and gesture commands to enhance collaborative editing, reducing the need for separate interfaces and optimizing resource consumption.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If gesture recognition capability is added to collaborative editing environment, then user experience and interaction quality are improved, but device complexity and processing requirements increase

Engineering Contradiction:
Improveuser experienceVSAvoidprocessing requirements
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The patent replaces traditional tactile input devices (mouse, keyboard) with gesture-based optical input system. The camera captures hand gestures and the system processes these visual inputs to control collaborative editing functions, eliminating the need for physical interaction devices and improving user experience through more natural interaction.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Solution Approach 2:

The patent introduces an intermediary processing layer between the camera and the collaborative editing application. The gesture recognition system acts as a mediator that translates physical hand gestures into digital commands, managing the complexity of gesture interpretation while providing simple user interactions.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Adaptability or versatility

If gesture recognition is implemented in collaborative editing, then separate applications are reduced, but computing resources and network bandwidth consumption increase

Engineering Contradiction:
Improveapplication integrationVSAvoidcomputing resources
Core Design Contradiction:
Adaptability or versatilityVSUse of energy by moving object

Solution Approach 1:

The patent merges the collaborative editing application with gesture recognition capabilities into a single integrated system. Instead of requiring separate applications for video conferencing and collaborative editing, the system combines both functions with unified gesture-based control, reducing the need for multiple applications while managing resource consumption through shared processing infrastructure.

Inventive Principle:
Principle #5Merging (Combining)

Solution Approach 2:

The patent implements a universal gesture recognition system that can control multiple functions within the collaborative editing environment. A single gesture input mechanism serves multiple purposes including drawing, selecting, and editing operations, reducing the need for specialized input methods and minimizing overall computing resource requirements compared to multiple dedicated systems.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Data Source

PatentUS20250284345A1Multi-user collaborative interfaces for streaming video
Publication Date: 2025.09.11 PENCIL LEARNING TECH INC
  • US20250284345A1 patent drawing
  • US20250284345A1 patent drawing
  • US20250284345A1 patent drawing

AI summary

The present disclosure provides systems and methods of providing multi-user collaborative interfaces. A computing system may receive a video stream. The video stream may have an overlay interface defined along a first plane. The computing system may identify, in at least one frame, a region defined on a second plane on which a user's hand is oriented. The computing system may determine that a set of features associated with the user's hand within the region corresponds to a draw command. The computing system may identify, from the region, a first coordinate for the draw command based on at least one of the set of features. The computing system may translate the first coordinate to a second coordinate defined on the first plane to which to apply the draw command. The computing system may render a visual element at the second coordinate in the overlay interface.