Directional Audio Separation for Meeting Recordings

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

In meetings with multiple participants, existing electronic devices struggle to accurately identify and separate individual speakers' voices, making it difficult to distinguish between speakers as the number of participants increases.

Innovation Solution

An electronic device equipped with multiple microphones and a directional recognition algorithm that classifies sound directions into sectors, allowing for the generation and reproduction of audio files with directional information, enabling selective playback of specific speakers' voices.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Quantity of substance

If multiple microphones are used to record meeting contents, then the ability to capture multiple speakers' voices is improved, but the difficulty of identifying and separating individual speakers increases

Engineering Contradiction:
Improvenumber of microphonesVSAvoiddifficulty of identifying individual speakers
Core Design Contradiction:
Quantity of substanceVSDifficulty of detecting and measuring

Solution Approach 1:

The patent segments the audio recording by associating each microphone's recording with its specific spatial direction. The controller divides the meeting recording into multiple audio files, each corresponding to a specific direction from which a speaker's voice was captured. This segmentation allows individual speakers to be identified and separated based on the directional information from different microphones, resolving the difficulty of distinguishing speakers when multiple microphones are used.

Inventive Principle:
Principle #1Segmentation

2Quantity of substance

If the number of meeting participants increases, then the comprehensiveness of meeting coverage is improved, but the ability to identify which user speaks deteriorates

Engineering Contradiction:
Improvenumber of participantsVSAvoidprecision of speaker identification
Core Design Contradiction:
Quantity of substanceVSMeasurement precision

Solution Approach 1:

The patent introduces a spatial dimension to speaker identification by capturing the directional information from which each voice originates. Instead of relying solely on temporal or spectral analysis, the system uses the spatial dimension (direction) provided by multiple microphones positioned at different locations. This additional dimension enables precise identification of which participant is speaking, even when the number of participants increases, as each speaker can be distinguished by the direction from which their voice is captured.

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

Data Source

PatentUS11301201B2Method and apparatus for playing audio files
Publication Date: 2022.04.12 SAMSUNG ELECTRONICS CO LTD
  • US11301201B2 patent drawing
  • US11301201B2 patent drawing
  • US11301201B2 patent drawing

AI summary

An electronic device is provided. The electronic device includes a controller configured to execute one or more modules, an audio reproduction module configured to reproduce an audio file including reproduction sections, each of the reproduction sections comprising audio data and directional information, a display configured to display selectable objects corresponding to the directional information, and an audio control module configured to determine whether to reproduce audio data corresponding the directional information based on an input for selecting one of the selectable objects.