Video Processing With Interactive Display Regions for Group Calls
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing video conferencing and live streaming systems limit interaction between participants by restricting videos to specific display regions, hindering detailed discussions and causing inconvenience during group calls.
Innovation Solution
A system and method that allows portions of a user's video to extend and interact with other users' display regions by defining interactive regions on the screen, enabling users to control the extension of their video portions to other regions through hand movements, and using image and object recognition to enhance interaction.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If videos are restricted to specific display regions, then display organization is maintained, but interaction between users is limited
Solution Approach 1:
The display screen is divided into multiple independent display regions, each assigned to a different user video. This segmentation allows videos to be organized in specific regions while enabling selective extension of individual video portions to other regions, thus maintaining display organization while enhancing interaction capabilities.
Solution Approach 2:
The patent extends the traditional two-dimensional display region boundaries by allowing video content to extend beyond its assigned region into adjacent regions. This dimensional extension enables users to point at and interact with other users' videos without requiring additional display devices or complex interface operations.
2Adaptability or versatility
If users can point at other users' videos, then interaction is enhanced, but determining the pointed region becomes more difficult
Solution Approach 1:
The patent introduces an interactive region as an intermediary zone between display regions. When a user's video portion extends to another display region, the interactive region detects hand gestures or pointing actions within this intermediate zone, making it easier to determine which region the user is pointing at without directly analyzing complex gesture data across region boundaries.
Solution Approach 2:
The patent replaces complex mechanical gesture recognition systems with image recognition and object detection algorithms. By using computer vision techniques to analyze video frames and detect hand positions or pointing gestures, the system can accurately determine the pointed region without requiring complex mechanical sensors or physical interaction mechanisms.
3Ease of operation
If video portions extend to other regions, then user interaction is improved, but video display stability is reduced
Solution Approach 1:
The patent implements dynamic video display where video portions can extend to adjacent display regions based on user interaction needs. The extension is controlled and reversible - videos maintain their stable, fixed-region display mode during normal operation, but can dynamically extend when users need to point at or interact with other users' videos, thus balancing stability with interaction flexibility.
Data Source
AI summary
The present disclosure relates to a system, a method and a computer-readable medium for video processing. The method includes displaying a live video of a first user in a first region on a user terminal and displaying a video of a second user in a second region on the user terminal. A portion of the live video of the first user extends to the second region on the user terminal. The present disclosure can improve interaction during a conference call or a group call.


