Audio Decoder Classification for Speech and Music
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing audio decoding technologies face challenges in accurately classifying speech and non-speech content, leading to potential degradation of audio quality due to inadequate noise suppression and processing techniques.
Innovation Solution
A device and method that decodes an encoded audio signal to generate a synthesized signal and classifies it based on parameters extracted from the encoded audio signal, using a classifier to differentiate between speech and non-speech content, thereby enabling selective noise suppression and processing adjustments.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Object-affected harmful factors
If noise suppressor processes audio signal, then background noise is reduced, but audio quality of non-speech content is degraded
Solution Approach 1:
The system dynamically adjusts noise suppression processing based on real-time classification of the audio signal. The noise suppressor is activated only when speech content is detected, and deactivated when non-speech content (music, ring tones) is detected, allowing the system to adapt its behavior to the current audio characteristics and avoid degrading music quality while still removing background noise from speech
Solution Approach 2:
The system uses a classifier to continuously monitor the audio signal and provide feedback about its content type. This feedback loop allows the noise suppressor to adjust its operation based on whether the current signal is speech or non-speech, creating a closed-loop control system that prevents harmful noise suppression of music content while maintaining effective noise removal from speech
2Adaptability or versatility
If multiple coding technologies are used, then variety of content is supported, but device complexity increases
Solution Approach 1:
The decoder is designed with multiple coding technologies (speech mode decoder and music mode decoder) that can process different types of audio content. The system uses a classifier to identify the content type and routes the signal to the appropriate decoder, allowing a single device to handle both speech and non-speech content effectively without requiring separate dedicated devices for each function
Data Source
AI summary
A device includes a decoder configured to receive an encoded audio signal at a decoder and to generate a synthesized signal based on the encoded audio signal. The device further includes a classifier configured to classify the synthesized signal based on at least one parameter determined from the encoded audio signal.


