3D Avatar Input Mode Switching for Immersive Conferencing

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current online video conferencing systems fail to provide a fully immersive experience for participants, as they typically involve static video images or basic audio, lacking engagement and environmental interaction.

Innovation Solution

A method and system for a three-dimensional (3D) environment that allows multiple input/output modes, enabling users to animate avatars based on input signals, and change input modes in response to location changes, providing a more immersive and interactive experience.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If static video images and basic audio are used for online conferencing, then device complexity is reduced, but user engagement and immersion deteriorate

Engineering Contradiction:
Improveuser engagementVSAvoidsystem complexity
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The patent creates virtual copies (avatars) of users that replicate their physical actions, emotions, and presence in a 3D environment. These avatars serve as digital representations that convey user intent and emotional state without requiring the actual physical presence of users, thereby enhancing engagement while managing system complexity through standardized avatar models

Inventive Principle:
Principle #26Copying

Solution Approach 2:

The patent transitions from 2D video screens to 3D immersive environments, adding spatial depth and environmental context to conferencing. This dimensional expansion allows users to interact with virtual spaces and objects, creating more engaging experiences while distributing system complexity across multiple standardized 3D assets

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

2Adaptability or versatility

If multiple input/output modes are implemented based on user location, then adaptability improves, but device complexity increases

Engineering Contradiction:
Improveinput mode adaptabilityVSAvoidsystem complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent implements dynamic input mode switching that automatically adapts to user location and environmental context. The system transitions between different input modes (e.g., camera-based avatar control, audio-only, text input) based on real-time location data and environmental conditions, providing adaptability while managing complexity through automated mode selection algorithms

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The patent creates a universal input system that can operate across multiple modes and environments using the same core avatar infrastructure. The same avatar model serves multiple functions across different input modes (camera control, audio control, text control), reducing overall system complexity while maintaining high adaptability to different user contexts

Inventive Principle:
Principle #6Universality (Multi-functionality)

3Ease of operation

If avatars are animated based on emotion information from input signals, then user engagement improves, but processing requirements increase

Engineering Contradiction:
Improveuser engagementVSAvoidprocessing power
Core Design Contradiction:
Ease of operationVSPower

Solution Approach 1:

The patent animates avatars by changing their emotional state parameters based on detected user emotions. Instead of complex real-time physics-based animation, the system adjusts predefined emotional parameters (happy, sad, angry, neutral) that trigger corresponding avatar expressions and body language, reducing processing requirements while maintaining high user engagement through emotionally responsive avatars

Inventive Principle:
Principle #35Parameter changes

Data Source

PatentUS20250124637A1Integrated input/output (i/o) for a three-dimensional (3D) environment
Publication Date: 2025.04.17 ROBLOX CORP
  • US20250124637A1 patent drawing
  • US20250124637A1 patent drawing
  • US20250124637A1 patent drawing

AI summary

Various input modes and output modes may be used for a three-dimensional (3D) environment. A user may use a particular input mode (e.g., text, audio, video, etc.) for animating a 3D avatar of the user in the 3D environment. The user may use a particular output mode (e.g., text, audio, 3D animation, etc.) in the presentation of the 3D environment. The input/output modes may change based on conditions such as a location of the user.