3D Voice Spatialisation to Avoid Overlapping Virtual Sounds

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Replicating a 3-dimensional auditory experience for speech in virtual environments, such as in-game or augmented reality environments, is challenging due to overlapping sound locations, leading to a high cognitive burden on users and a decrease in immersive experience.

Innovation Solution

An audio processing unit intelligently assigns a 3-dimensional location to speech in a virtual environment based on user preferences, game environment audio, avatar location, and other factors, adjusting the location based on detected events, and processes the speech to simulate its origin at the adjusted location.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If speech is assigned a fixed 3D location in the virtual environment, then the immersive experience is improved, but user experience deteriorates when speech location overlaps with other sounds

Engineering Contradiction:
Improveimmersive experienceVSAvoiduser experience
Core Design Contradiction:
ReliabilityVSEase of operation

Solution Approach 1:

The patent applies dynamics by making the speech 3D location adjustable and adaptive rather than fixed. The system dynamically repositions speech in the 3D audio space based on detected overlapping sounds, allowing the location to change in response to environmental conditions while maintaining immersion

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The patent changes the spatial parameters of speech by adjusting its 3D location coordinates. When overlap is detected, the system modifies the position parameters (x, y, z coordinates) of the speech source to separate it from overlapping sounds, thereby resolving the contradiction between maintaining fixed location and avoiding overlap

Inventive Principle:
Principle #35Parameter changes

2Ease of operation

If speech location is adjusted dynamically to avoid overlap, then user experience is improved, but the immersive experience deteriorates due to arbitrary placement

Engineering Contradiction:
Improveuser experienceVSAvoidimmersive experience
Core Design Contradiction:
Ease of operationVSReliability

Solution Approach 1:

The patent applies local quality by making adjustments to speech location localized and context-specific rather than arbitrary. The system identifies specific overlapping sound sources and repositions speech to adjacent non-overlapping spatial regions, maintaining logical spatial relationships while avoiding conflict

Inventive Principle:
Principle #3Local quality

Solution Approach 2:

The system uses feedback by continuously monitoring the 3D audio environment for overlapping sounds and using this information to adjust speech location. The adjustment is based on detected acoustic conditions rather than arbitrary placement, maintaining immersion through responsive, context-aware positioning

Inventive Principle:
Principle #23Feedback

3Reliability

If multiple sounds originate from the same location, then the virtual environment realism is improved, but speech intelligibility deteriorates due to cognitive burden

Engineering Contradiction:
Improvevirtual environment realismVSAvoidspeech intelligibility
Core Design Contradiction:
ReliabilityVSLoss of information

Solution Approach 1:

The patent applies segmentation by spatially separating speech from other sounds in the 3D audio space. Instead of having all sounds from the same physical location, the system segments speech into a distinct spatial region, allowing simultaneous presence of multiple sound sources without overlap and reducing cognitive burden

Inventive Principle:
Principle #1Segmentation

Data Source

PatentUS12370448B23D spatialisation of voice chat
Publication Date: 2025.07.29 SONY COMP ENTERTAINMENT EURO LTD
  • US12370448B2 patent drawing
  • US12370448B2 patent drawing
  • US12370448B2 patent drawing

AI summary

The invention provides techniques for intelligently positioning speech in a virtual environment. Factors such as user preferences, the location of virtual environment audio and/or visual events, avatar location, and others can be taken into account when selecting a suitable location for the speech. The virtual environment can be a game environment, a meeting environment, an augmented reality environment, a virtual reality environment, and the like. The invention can be implemented by an audio processing unit which may be part of a game console.