Avatar Voice Volume Adjustment in Noisy Virtual Spaces

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

In noisy virtual environments, existing voice chat systems struggle to facilitate clear communication between avatars due to limitations in sound volume control and potential disturbances from other avatars, leading to difficulty in hearing desired conversations.

Innovation Solution

An information processing system that includes multiple user terminals and a server, which analyze and adjust voice data volume based on association and noise levels to ensure clear communication between avatars.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If sound volume parameter is increased to make voice audible in noisy environment, then voice can be heard by desired avatar, but other nearby avatars are disturbed by loud voice

Engineering Contradiction:
Improvevoice audibilityVSAvoidnuisance to other spectators
Core Design Contradiction:
ReliabilityVSObject-affected harmful factors

Solution Approach 1:

The patent applies local quality by making sound volume adjustment specific to target avatars rather than global. The sound collection object's avatar receives enhanced volume for the sound source object's voice, while other avatars experience normal or reduced volume, creating localized audio quality differences based on avatar associations.

Inventive Principle:
Principle #3Local quality

Solution Approach 2:

The system implements feedback by continuously monitoring the noisy environment and dynamically adjusting sound volume parameters. The sound collection object's terminal receives feedback about ambient noise levels and automatically adjusts the volume of the sound source object's voice to ensure audibility while minimizing disturbance to others.

Inventive Principle:
Principle #23Feedback

2Ease of operation

If sound volume parameter is controlled only on sound source object side, then transmission can be managed, but sound collection object cannot adjust for personal hearing needs

Engineering Contradiction:
Improvetransmission controlVSAvoidpersonal hearing adjustment
Core Design Contradiction:
Ease of operationVSAdaptability or versatility

Solution Approach 1:

The patent inverts the traditional one-way control model by enabling bidirectional sound parameter control. Instead of only the sound source object controlling volume, the sound collection object can also adjust sound parameters for received voices, allowing personal hearing needs to be accommodated while maintaining transmission management.

Inventive Principle:
Principle #13The other way round (Inversion)

3Measurement precision

If user must manually determine whether to change sound volume setting, then control precision is maintained, but responsiveness to changing noise conditions decreases

Engineering Contradiction:
Improvevolume control precisionVSAvoidresponse to noise changes
Core Design Contradiction:
Measurement precisionVSSpeed

Solution Approach 1:

The system applies self-service by automatically monitoring noise levels and adjusting sound volume parameters without requiring manual user intervention. The sound collection object's terminal autonomously detects changes in ambient noise and adjusts the volume of target avatars' voices accordingly, maintaining both precision and responsiveness.

Inventive Principle:
Principle #25Self-service

Data Source

PatentUS20250111858A1Information processing system that allows user to establish conversation with desired person through avatar even in noisy environment in virtual space, edge device, server, control method, and storage medium
Publication Date: 2025.04.03 CANON KK
  • US20250111858A1 patent drawing
  • US20250111858A1 patent drawing
  • US20250111858A1 patent drawing

AI summary

An information processing system that allows a user to establish a conversation with a desired person through an avatar even in a noisy environment in a virtual space is provided. The information processing system that comprises multiple user terminals, each including a voice input unit and a voice output unit, and provides a virtual space including avatars linked to the multiple user terminals, includes a unit to generate voice data whose sound source is each avatar linked to each user terminal, and one or more processors and/or circuitry configured to execute a voice data transmitting processing, execute an association determination processing, execute a voice data analysis processing, and execute a sound volume adjustment processing that, when first voice data is associated with a second avatar and second voice data disturbs the first voice data, adjusts a sound volume of voice data to be transmitted to a second user terminal.