3D Voice Spatialisation to Avoid Overlapping Virtual Sounds
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Replicating a 3-dimensional auditory experience for speech in virtual environments, such as in-game or augmented reality environments, is challenging due to overlapping sound locations, leading to a high cognitive burden on users and a decrease in immersive experience.
Innovation Solution
An audio processing unit intelligently assigns a 3-dimensional location to speech in a virtual environment based on user preferences, game environment audio, avatar location, and other factors, adjusting the location based on detected events, and processes the speech to simulate its origin at the adjusted location.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If speech is assigned a fixed 3D location in the virtual environment, then the immersive experience is improved, but user experience deteriorates when speech location overlaps with other sounds
Solution Approach 1:
The patent applies dynamics by making the speech 3D location adjustable and adaptive rather than fixed. The system dynamically repositions speech in the 3D audio space based on detected overlapping sounds, allowing the location to change in response to environmental conditions while maintaining immersion
Solution Approach 2:
The patent changes the spatial parameters of speech by adjusting its 3D location coordinates. When overlap is detected, the system modifies the position parameters (x, y, z coordinates) of the speech source to separate it from overlapping sounds, thereby resolving the contradiction between maintaining fixed location and avoiding overlap
2Ease of operation
If speech location is adjusted dynamically to avoid overlap, then user experience is improved, but the immersive experience deteriorates due to arbitrary placement
Solution Approach 1:
The patent applies local quality by making adjustments to speech location localized and context-specific rather than arbitrary. The system identifies specific overlapping sound sources and repositions speech to adjacent non-overlapping spatial regions, maintaining logical spatial relationships while avoiding conflict
Solution Approach 2:
The system uses feedback by continuously monitoring the 3D audio environment for overlapping sounds and using this information to adjust speech location. The adjustment is based on detected acoustic conditions rather than arbitrary placement, maintaining immersion through responsive, context-aware positioning
3Reliability
If multiple sounds originate from the same location, then the virtual environment realism is improved, but speech intelligibility deteriorates due to cognitive burden
Solution Approach 1:
The patent applies segmentation by spatially separating speech from other sounds in the 3D audio space. Instead of having all sounds from the same physical location, the system segments speech into a distinct spatial region, allowing simultaneous presence of multiple sound sources without overlap and reducing cognitive burden
Data Source
AI summary
The invention provides techniques for intelligently positioning speech in a virtual environment. Factors such as user preferences, the location of virtual environment audio and/or visual events, avatar location, and others can be taken into account when selecting a suitable location for the speech. The virtual environment can be a game environment, a meeting environment, an augmented reality environment, a virtual reality environment, and the like. The invention can be implemented by an audio processing unit which may be part of a game console.


