Sequential Neural Networks for Audio Signal Separation

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing signal processing devices for hearing aids and similar devices struggle to efficiently separate and enhance speech from noisy and multi-component audio signals, leading to poor intelligibility.

Innovation Solution

A signal processing device utilizing a sequential arrangement of two neural networks, where the first neural network conditions the input signal and the second neural network separates audio signals, allowing for efficient and accurate processing, including real-time separation and customization based on specific sound sources.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If traditional signal processing methods are used to separate audio signals from noisy input, then the processing can be performed with simpler devices, but the separation accuracy and speech intelligibility deteriorate

Engineering Contradiction:
Improveseparation accuracyVSAvoidprocessing system complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The patent divides the signal processing task into two distinct neural networks: a first neural network for conditioning the input signal and a second neural network for separating audio signals. This segmentation allows each network to specialize in a specific function, improving overall separation accuracy while managing system complexity through modular architecture

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The first neural network performs preliminary conditioning of the input signal before it is fed to the second neural network for separation. This preliminary action prepares the signal by reducing noise and enhancing relevant features, which improves the accuracy of the subsequent separation process

Inventive Principle:
Principle #10Preliminary action

2Speed

If real-time audio signal separation is implemented, then the processing speed improves, but the computational complexity and processing time increase

Engineering Contradiction:
Improveprocessing speedVSAvoidcomputational complexity
Core Design Contradiction:
SpeedVSDevice complexity

Solution Approach 1:

By segmenting the processing into conditioning and separation stages with dedicated neural networks, the system can optimize each stage for real-time performance while managing overall computational complexity

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The neural networks are trained to automatically adapt to different input conditions and sound sources without requiring manual intervention or complex real-time adjustments, enabling self-service processing that maintains speed while managing complexity

Inventive Principle:
Principle #25Self-service

3Measurement precision

If the signal processing device is customized for specific sound sources, then the speech enhancement quality improves, but the adaptability to different input signals decreases

Engineering Contradiction:
Improvespeech enhancement qualityVSAvoidinput signal adaptability
Core Design Contradiction:
Measurement precisionVSAdaptability or versatility

Solution Approach 1:

The neural networks are designed to dynamically adapt their processing based on the characteristics of the input signal. The system can adjust its behavior to optimize speech enhancement for different sound sources while maintaining versatility across various input types

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The processing parameters of the neural networks can be changed based on the detected sound source characteristics. This allows the system to optimize enhancement quality for specific sources while maintaining the ability to handle different input signals through parameter adjustment

Inventive Principle:
Principle #35Parameter changes

4Measurement precision

If multiple neural networks are arranged sequentially for conditioning and separation, then the processing accuracy improves, but the device complexity increases

Engineering Contradiction:
Improveseparation accuracyVSAvoidnetwork architecture complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The sequential arrangement of specialized neural networks segments the complex processing task into manageable functional blocks, improving accuracy while controlling architecture complexity through clear functional division

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The neural networks are designed with universal processing capabilities that can handle various types of audio inputs and sound sources, reducing the need for multiple specialized networks and thereby controlling overall system complexity

Inventive Principle:
Principle #6Universality (Multi-functionality)

Data Source

PatentUS11910163B2Signal processing device, system and method for processing audio signals
Publication Date: 2024.02.20 SONOVA AG
  • US11910163B2 patent drawing
  • US11910163B2 patent drawing
  • US11910163B2 patent drawing

AI summary

A signal processing device for processing audio signals is described. The signal processing device has an input interface for receiving an input signal and an output interface for outputting an output signal. Moreover, the signal processing device has at least one first neural network for conditioning the input signal and at least one second neural network for separating one or more audio signals from the input signal. The at least one first neural network and the at least one second neural network are arranged sequentially.