Dynamic Voice Enhancement Gain for Clearer Multichannel Speech

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing audio signal processing technologies compromise speech clarity in pursuit of immersive surround sound, leading to unbalanced sound and amplified background noise.

Innovation Solution

A method and system for intelligent dynamic speech enhancement that performs speech detection and intelligent enhancement gain control on multi-channel audio sources, applying dynamic loudness balancing based on signal power strength ratios and system volume levels.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Illumination intensity

If multi-channel audio coding technology is used to provide immersive surround sound experience, then spatial resolution and stereo sound experience are improved, but speech clarity is compromised

Engineering Contradiction:
Improvespatial resolutionVSAvoidspeech clarity
Core Design Contradiction:
Illumination intensityVSManufacturing precision

Solution Approach 1:

The audio signal is segmented into center channel (speech) and other channels (surround sound). The system processes the center channel differently from other channels, allowing speech enhancement without compromising the surround sound effect. This segmentation enables independent optimization of speech clarity while maintaining spatial immersion.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

Different quality enhancement is applied to different parts of the audio signal. The center channel receives speech enhancement processing while other channels maintain their original surround sound characteristics. This local quality approach ensures that speech clarity is improved in specific regions (center channel) without degrading the overall spatial sound experience.

Inventive Principle:
Principle #3Local quality

2Manufacturing precision

If dialogue volume is increased to improve speech clarity, then speech audibility is improved, but background noise and other audio elements become unbalanced

Engineering Contradiction:
Improvespeech clarityVSAvoidsound balance
Core Design Contradiction:
Manufacturing precisionVSStability of the object's composition

Solution Approach 1:

The system dynamically adjusts the enhancement gain based on real-time analysis of the audio signal characteristics, including signal power strength ratios and system volume levels. This dynamic adjustment ensures that speech is enhanced only when and where needed, maintaining the overall balance of the audio mix while improving speech clarity in varying listening conditions.

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The system uses feedback from speech detection and signal analysis to control the enhancement gain. By continuously monitoring the audio signal and adjusting the enhancement level based on detected speech characteristics, the system maintains sound balance while improving speech audibility. The feedback mechanism prevents over-enhancement that would disrupt the overall audio composition.

Inventive Principle:
Principle #23Feedback

3Manufacturing precision

If speech enhancement gain is increased to improve speech audibility, then speech clarity is improved, but audio distortion and perceptual imbalances increase

Engineering Contradiction:
Improvespeech audibilityVSAvoidaudio distortion
Core Design Contradiction:
Manufacturing precisionVSObject-affected harmful factors

Solution Approach 1:

The system changes multiple parameters simultaneously to optimize speech enhancement: signal power strength ratio, system volume level, and enhancement gain. By coordinating changes in these parameters, the system achieves improved speech audibility while preventing audio distortion. The parameter changes are interlinked to ensure that enhancement remains within acceptable distortion thresholds.

Inventive Principle:
Principle #35Parameter changes

Solution Approach 2:

The system applies partial enhancement only to the center channel speech signal rather than excessively enhancing all audio channels. This selective partial action improves speech audibility without causing the perceptual imbalances and distortion that would result from uniform enhancement across all channels. The enhancement is applied judiciously to achieve the minimum necessary improvement.

Inventive Principle:
Principle #16Partial or excessive action

Data Source

PatentUS20250131939A1Method and System of Intelligent Dynamic Voice Enhancement
Publication Date: 2025.04.24 HARMAN INT IND INC
  • US20250131939A1 patent drawing
  • US20250131939A1 patent drawing
  • US20250131939A1 patent drawing

AI summary

A method and a system for intelligent dynamic speech enhancement for an audio source, comprising performing speech detection and intelligent enhancement gain control on a multi-channel audio source input to determine speech enhancement gain, and further comprising applying the speech enhancement gain in dynamic loudness balancing performed on the multi-channel audio source input, wherein the intelligent enhancement gain control comprises setting the speech enhancement gain based on a signal power strength ratio of a signal of a center channel to a sum of signals of other channels, and setting the speech enhancement gain based on a system volume level