Variable-Volume Spatial Audio for Concurrent Conversation Discovery

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current teleconferencing systems prevent actual conversations between participants by equally amplifying audio, making it difficult for multiple speakers to speak simultaneously and limiting interactions to text-based messages, thus reducing the unique aspects of conversation to a broadcast.

Innovation Solution

Implementing variable-volume audio based on user-controlled or system-controlled location within a virtual room, allowing participants to move closer to hear others, with supplemental content and speech-to-text analysis to enhance interaction and conversation discovery.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If audio is equally amplified for all participants, then all participants can hear everyone clearly, but actual conversations between participants become impossible and interactions are limited to text-based messages

Engineering Contradiction:
Improveaudio clarityVSAvoidconversation capability
Core Design Contradiction:
ReliabilityVSAdaptability or versatility

Solution Approach 1:

The patent applies local quality by differentiating audio processing based on spatial location. Audio from participants in close proximity is amplified more than audio from distant participants, creating localized audio zones that enable natural conversations while maintaining overall system reliability

Inventive Principle:
Principle #3Local quality

Solution Approach 2:

The patent implements dynamics by making audio amplification adjustable and adaptive rather than static. The system dynamically adjusts audio levels based on participant positions, allowing the audio distribution to change in real-time as participants move or engage in different conversation scenarios

Inventive Principle:
Principle #15Dynamics

2Adaptability or versatility

If variable-volume audio based on spatial location is implemented, then actual conversations become possible in virtual settings, but system complexity increases with location control mechanisms

Engineering Contradiction:
Improveconversation capabilityVSAvoidlocation control system
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent applies self-service by enabling participants to automatically control their own audio experience through spatial positioning. Participants simply move their interface elements to desired locations, and the system automatically adjusts audio volumes without requiring manual audio controls or complex configuration

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The patent uses spatial position as an intermediary between participant intent and audio delivery. Rather than directly controlling audio volumes, the system uses location as a mediator that automatically translates spatial relationships into appropriate audio mixing, simplifying the control interface

Inventive Principle:
Principle #24Intermediary (Mediator)

3Loss of information

If supplemental content and speech-to-text analysis are added to enhance interaction, then conversation discovery is improved, but processing time and computational resources increase

Engineering Contradiction:
Improveconversation insightsVSAvoidprocessing time
Core Design Contradiction:
Loss of informationVSLoss of time

Solution Approach 1:

The patent applies partial action by selectively processing audio content rather than analyzing all audio equally. The system focuses computational resources on identifying key conversation elements and insights, providing sufficient information enhancement without requiring complete analysis of all audio data

Inventive Principle:
Principle #16Partial or excessive action

Solution Approach 2:

The patent implements preliminary action by pre-processing audio streams into text format and extracting key insights before presenting them to participants. This preparation work is done in advance to enable quick conversation discovery and joining without real-time processing delays

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS12437765B2Spatial audio conversational analysis for enhanced conversation discovery
Publication Date: 2025.10.07 MICROSOFT TECHNOLOGY LICENSING LLC
  • US12437765B2 patent drawing
  • US12437765B2 patent drawing
  • US12437765B2 patent drawing

AI summary

Systems and methods for providing enhanced teleconferencing. An example method includes receiving audio streams from a plurality of client devices of participants of a teleconference; converting the audio streams for a first conversation within the teleconference into first text; converting the audio streams for a second conversation within the teleconference into a second text; analyzing the first text to identify one or more topics being discussed in the first conversation; analyzing the second text to identify one or more topics being discussed in the second conversation; and presenting, in a teleconference user interface, at least one of the one or more topics being discussed in the first conversation or the one or more topics being discussed in the second conversation.