Spatial Audio Transmission for Virtual Conference Privacy

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional videoconferencing technologies lack the social interaction and privacy features of in-person meetings, leading to difficulties in private conversations and ineffective communication among multiple participants, as well as security vulnerabilities and inefficient bandwidth usage.

Innovation Solution

A computer-implemented method and system that creates a three-dimensional virtual space where avatars represent users, determining sound volumes based on their positions relative to the virtual camera, allowing private communication by transmitting audio streams only to users who can hear the speaker, and preventing transmission to those who cannot.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If audio streams are transmitted to all participants in a virtual conference, then all users can hear the speaker, but bandwidth is wasted and security vulnerabilities increase

Engineering Contradiction:
Improveaudio transmission securityVSAvoidbandwidth usage
Core Design Contradiction:
ReliabilityVSLoss of energy

Solution Approach 1:

The patent implements spatial audio transmission by determining the position of each user's avatar relative to the speaker in a three-dimensional virtual space. Audio streams are selectively transmitted only to users whose avatars are within hearing distance of the speaker, rather than broadcasting to all participants. This localizes the audio transmission quality based on spatial position, reducing bandwidth waste and improving security by limiting access to authorized listeners only.

Inventive Principle:
Principle #3Local quality

2Ease of operation

If conventional videoconferencing mixes audio streams equally from multiple speakers, then all participants receive the same audio mix, but private conversations become impossible and social interaction is reduced

Engineering Contradiction:
Improvesocial interaction capabilityVSAvoidprivate conversation capability
Core Design Contradiction:
Ease of operationVSLoss of information

Solution Approach 1:

The patent segments the virtual conference space into multiple hearing zones based on the positions of avatars in three-dimensional space. Instead of creating a single mixed audio stream for all participants, the system divides audio transmission into separate channels corresponding to different spatial regions. Users in the same hearing zone can engage in private conversations while users in other zones remain excluded, enabling simultaneous private and public interactions within the same conference.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent introduces a spatial dimension to audio transmission by implementing a three-dimensional virtual space where avatar positions determine hearing relationships. This dimensional approach allows users to physically position their avatars to control who can hear them, enabling private conversations by simply moving away from other participants. The spatial dimension transforms the flat, all-or-nothing audio mixing of conventional conferencing into a nuanced, position-dependent audio environment.

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

3Reliability

If a speaker's vantage point is bound by the virtual camera view, then the speaker can see what they are presenting, but they cannot sense the presence of users outside their field of view

Engineering Contradiction:
Improveperipheral awarenessVSAvoidvirtual camera positioning
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

The patent implements a universal three-dimensional virtual space that serves multiple functions simultaneously: it provides the visual presentation view through the virtual camera while also enabling spatial audio determination and peripheral awareness. The same three-dimensional coordinate system used for rendering the visual scene is also used to calculate avatar positions and hearing relationships. This multi-functional approach allows speakers to maintain their camera-focused view while the system independently tracks all avatar positions to determine who can hear the speaker, eliminating the need for separate positioning systems.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Data Source

PatentUS11184362B1Securing private audio in a virtual conference, and applications thereof
Publication Date: 2021.11.23 KATMAI TECH INC
  • US11184362B1 patent drawing
  • US11184362B1 patent drawing
  • US11184362B1 patent drawing

AI summary

Disclosed herein is a computer-implemented method, system, device, and computer program product for securing private audio in a virtual conference. For each of the users in the virtual conference, a device of the speaking user or a server determines whether a respective user is able to hear the speaking user based on whether a respective sound volume at which the respective user is able to hear the speaking user exceeds a threshold amount. The speaking user's device or the server prevents the transmission of the audio stream to devices of the users determined not to be able to hear the speaking user.