Semi-supervised Source Separation Using Non-negative Matrix Factorization

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current signal processing technologies face challenges in effectively separating mixed signals from multiple sources, such as audio signals containing overlapping sounds, where computers struggle to differentiate between constituent sound sources like the human auditory system can.

Innovation Solution

The implementation of semi-supervised source separation using non-negative techniques, specifically employing non-negative hidden Markov (N-HMM) and non-negative factorial hidden Markov (N-FHMM) models to model and separate signals of interest from noise or other sources, allowing for independent processing of signal components.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If traditional signal processing methods are used to separate mixed signals, then the separation process becomes computationally complex and requires extensive training data, but the separation effectiveness and signal-to-interference ratio remain insufficient

Engineering Contradiction:
Improvesignal separation effectivenessVSAvoidprocessing complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The patent segments the mixed signal into distinct source components using non-negative matrix factorization, decomposing the spectrogram into source signals and spectral profiles. This segmentation approach enables effective separation without requiring complex processing frameworks, directly resolving the contradiction between separation effectiveness and processing complexity.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent transforms the signal processing problem into the spectrogram domain, changing the parameter representation from time-domain to frequency-time domain. This parameter transformation simplifies the separation task by exploiting the non-negativity property in the spectrogram representation, improving separation effectiveness while reducing computational complexity.

Inventive Principle:
Principle #35Parameter changes

2Reliability

If supervised learning methods with extensive training data are employed, then model accuracy improves, but the requirement for large amounts of training data increases processing complexity and time

Engineering Contradiction:
Improveseparation accuracyVSAvoidtraining time
Core Design Contradiction:
ReliabilityVSLoss of time

Solution Approach 1:

The patent implements a semi-supervised approach where the system performs self-service by automatically learning source characteristics during the separation process itself. The non-negative matrix factorization algorithm adapts to the specific mixture being processed without requiring external training data, achieving high reliability while eliminating the time loss associated with extensive training.

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The patent performs preliminary spectral analysis and non-negativity constraint application before the actual separation process. By pre-processing the signal in the spectrogram domain and establishing non-negativity constraints upfront, the system prepares the data structure to enable accurate separation without requiring subsequent extensive training, thus reducing training time while maintaining accuracy.

Inventive Principle:
Principle #10Preliminary action

3Manufacturing precision

If conventional source separation algorithms are applied to mixed audio signals, then signal components can be separated, but artifacts and interference in the separated signals increase

Engineering Contradiction:
Improvesignal component separationVSAvoidartifacts and interference
Core Design Contradiction:
Manufacturing precisionVSObject-generated harmful factors

Solution Approach 1:

The patent converts the harmful interference and artifacts into beneficial separation information by applying non-negativity constraints. The algorithm exploits the fact that source signals and spectral profiles are non-negative in the spectrogram domain, transforming what would be interference into structured components that can be cleanly separated, thus reducing artifacts while achieving precise component separation.

Inventive Principle:
Principle #22Blessing in disguise (Convert harm into benefit)

Solution Approach 2:

The patent introduces the spectrogram domain as an intermediary representation between the original mixed signal and the separated sources. By transforming the signal to the frequency-time domain and performing separation in this intermediate space using non-negative matrix factorization, the system achieves clean separation with minimal artifacts, avoiding the interference problems of direct time-domain separation.

Inventive Principle:
Principle #24Intermediary (Mediator)

Data Source

PatentUS8812322B2Semi-supervised source separation using non-negative techniques
Publication Date: 2014.08.19 ADOBE INC
  • US8812322B2 patent drawing
  • US8812322B2 patent drawing
  • US8812322B2 patent drawing

AI summary

Systems and methods for semi-supervised source separation using non-negative techniques are described. In some embodiments, various techniques disclosed herein may enable the separation of signals present within a mixture, where one or more of the signals may be emitted by one or more different sources. In audio-related applications, for instance, a signal mixture may include speech (e.g., from a human speaker) and noise (e.g., background noise). In some cases, speech may be separated from noise using a speech model developed from training data. A noise model may be created, for example, during the separation process (e.g., “on-the-fly”) and in the absence of corresponding training data.