Multi-party Call Audio Processing Spatial Filtering

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

In multi-party calls, audio signals from participants in the same physical space are often delayed and played back, causing confusion and inefficiency, as existing methods fail to account for the positional relationship between electronic devices, leading to unnecessary audio signal playback and increased processing burdens.

Innovation Solution

An audio processing method and device that determine preset attribute values for audio signals based on the positional relationship between electronic devices, prohibiting the output of audio signals from devices in the same physical space by comparing these values with target attribute values, thereby preventing unnecessary playback and improving call efficiency.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If audio signals from all participants are collected and played back through each electronic device, then all participants can hear all speakers, but participants in the same physical space experience confusion and inefficiency due to delayed playback

Engineering Contradiction:
Improveaudio signal playback coverageVSAvoidcall efficiency
Core Design Contradiction:
Adaptability or versatilityVSEase of operation

Solution Approach 1:

The patent applies local quality by differentiating audio processing based on the spatial location of electronic devices. Devices are classified into different groups (same physical space vs. different physical space) and receive different audio signal treatments. Participants in the same physical space do not receive delayed playback of locally captured audio, while participants in different locations do receive the full audio mix, thus adapting the system behavior to local spatial conditions.

Inventive Principle:
Principle #3Local quality

2Reliability

If audio signals are transmitted and played back for all participants, then complete audio coverage is achieved, but unnecessary audio playback occurs for devices in the same physical space, causing sound interference

Engineering Contradiction:
Improveaudio signal transmissionVSAvoidsound interference
Core Design Contradiction:
ReliabilityVSObject-affected harmful factors

Solution Approach 1:

The patent extracts and removes unnecessary audio signals from the playback stream. Specifically, when an electronic device is determined to be in the same physical space as the audio source, the corresponding audio signal is extracted and excluded from the playback output. This eliminates the harmful sound interference caused by playing back audio that is already present in the local environment.

Inventive Principle:
Principle #2Taking out (Extraction)

3Reliability

If audio signals are collected and processed for all participants, then comprehensive audio coverage is achieved, but processing burden on electronic devices increases

Engineering Contradiction:
Improveaudio signal collectionVSAvoidprocessing burden
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

The patent implements partial action by processing audio signals selectively rather than universally. Instead of collecting and processing audio signals for all possible participants equally, the system performs partial processing based on spatial relationships. Audio signals are only processed and transmitted to devices that need them (devices in different physical spaces), reducing the overall processing burden while maintaining necessary audio coverage.

Inventive Principle:
Principle #16Partial or excessive action

Data Source

PatentUS11477326B2Audio processing method, device, and apparatus for multi-party call
Publication Date: 2022.10.18 LENOVO SOFTWARE
  • US11477326B2 patent drawing
  • US11477326B2 patent drawing
  • US11477326B2 patent drawing

AI summary

An audio processing method for a multi-party call. The method includes obtaining a preset attribute value of each of at least one audio signal currently obtained by a first electronic device participating in a multi-party call, determining a target attribute value configured for the audio signal collected by the first electronic device, and detecting a first audio signal from the at least one audio signal currently obtained by the first electronic device and prohibiting the first electronic device from outputting the first audio signal. The preset attribute value is configured for the audio signal collected by each of a plurality of electronic devices participating in the multi-party call according to an attribute configuration rule determined according to a positional relationship between the plurality of electronic devices. A comparison result between a corresponding preset attribute value of the first audio signal and the target attribute value satisfies a first condition.