Virtual Space Voice Transmission Range Detection and Display Control

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

In existing virtual communication systems, users face difficulties in designating call targets within virtual spaces, leading to time-consuming operations and uncertainty about voice transmission recipients.

Innovation Solution

An information processing device with detection, voice control, and display control means that detects user voices, outputs them to relevant avatars based on predetermined conditions, and changes the display mode of listening avatars to indicate transmission, allowing users to recognize the voice transmission range.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If automatic voice transmission without target designation is implemented, then ease of operation is improved, but information reliability deteriorates because users cannot know which users will receive the voice

Engineering Contradiction:
Improveease of operationVSAvoidinformation reliability
Core Design Contradiction:
Ease of operationVSLoss of information

Solution Approach 1:

The system provides visual feedback by changing the display mode of listening avatars to indicate which users will receive the voice. This feedback mechanism allows users to confirm the transmission range before speaking, resolving the information reliability issue while maintaining automatic transmission

Inventive Principle:
Principle #23Feedback

Solution Approach 2:

The system performs preliminary display control to show the transmission range before the actual voice transmission occurs. By changing the avatar display mode in advance, users can verify who will receive the voice and adjust their speaking direction or volume accordingly

Inventive Principle:
Principle #10Preliminary action

2Loss of information

If target designation is required for all calls, then information reliability is improved, but ease of operation deteriorates due to time-consuming operations

Engineering Contradiction:
Improveinformation reliabilityVSAvoidease of operation
Core Design Contradiction:
Loss of informationVSEase of operation

Solution Approach 1:

The system applies partial action by requiring target designation only when necessary (when multiple users are in the transmission range). When only one user is present or the transmission range is clearly defined, automatic transmission is used, reducing operational burden while maintaining reliability when needed

Inventive Principle:
Principle #16Partial or excessive action

3Ease of operation

If voice is transmitted to all users in the virtual space, then ease of operation is improved, but loss of information worsens because users cannot recognize the transmission range

Engineering Contradiction:
Improveease of operationVSAvoidinformation reliability
Core Design Contradiction:
Ease of operationVSLoss of information

Solution Approach 1:

The system changes the display mode (color/appearance) of listening avatars to visually indicate which users are within the voice transmission range. This visual differentiation allows users to easily recognize who will receive the voice without requiring manual target designation

Inventive Principle:
Principle #32Color changes

Data Source

PatentUS20240362871A1Information processing device, information processing method, and computer-readable storage medium
Publication Date: 2024.10.31 NEC CORP
  • US20240362871A1 patent drawing
  • US20240362871A1 patent drawing
  • US20240362871A1 patent drawing

AI summary

One aim of the present invention is to provide an information processing device whereby a user can be made aware of the audio transmission range in a situation where a virtual space is used and users communicate with each other. This information processing device includes: a detection unit that detects audio generated by a user operating an avatar inside a virtual space; an audio control unit that outputs the audio to a user of an avatar that fulfills prescribed conditions in a relationship with a speaking avatar, being an avatar operated by the user that provided the audio; and a display control unit that changes the display mode for a listener avatar, being an avatar that fulfills the prescribed conditions.