Voice Processing Device Speaker Transition Noise
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional voice processing devices struggle to seamlessly transition sound enhancement between multiple speakers in a meeting setting, leading to unclear audio when the sound source location switches, as they only enhance sounds from the newly identified source location after correct identification, causing noise from previous locations to be amplified.
Innovation Solution
A voice processing device that includes a sound pickup and circuitry to identify the current sound source location and specify additional locations that have been previously identified as sound sources within a predetermined time period, enhancing sounds from both the current and past sound source locations to maintain clear audio during speaker changes.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If the location identifier identifies only the current sound source location, then the sound from the current source is enhanced, but the sound from the new source location remains unenhanced until identified, and the previous location produces unwanted noise
Solution Approach 1:
The system performs preliminary action by designating a specified source location (previous sound source location) in advance before the actual speaker transition occurs. This allows the system to pre-prepare for the upcoming transition by maintaining enhancement at the previous location while the new source is being identified, ensuring seamless audio continuity without interruption or noise.
2Productivity
If the system enhances sound only from the newly identified sound source location, then the new speaker's voice is clarified, but there is a delay until identification occurs and the previous location produces noise
Solution Approach 1:
The system merges the enhancement processing of two different sound source locations (current and specified/previous) simultaneously. By combining the enhancement outputs from both locations, the system achieves continuous clear audio during transitions while suppressing harmful noise through intelligent signal processing that distinguishes between desired and unwanted sounds.
3Reliability
If the system continuously enhances sound from the identified sound source location, then voice clarity is maintained, but during speaker switches the audio becomes unclear until the new source is identified
Solution Approach 1:
The system maintains continuity of useful action by keeping sound enhancement active at both the current sound source location and the specified source location (previous location) simultaneously during transitions. This continuous dual-location enhancement ensures that voice clarity is maintained without interruption or time delay, as at least one location is always being enhanced regardless of the transition state.
Data Source
AI summary
A voice processing device includes: a sound pickup to receive sounds respectively from a plurality of locations; and circuitry to: identify, from the plurality of locations, a sound source location at which a sound source exists, as a current sound source location; specify at least one location of the plurality of locations other than the current sound source location, from at least one sound source location that has been identified as the sound source location during a past predetermined time period, as a specified source location at which a sound to be enhanced exists; enhance a sound from the current sound source location, and a sound from the specified source location; and output an audio signal including the enhanced sounds.


