Teleconference Display for Overlapping Voice Clarity
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
In voice conference systems, overlapping voices from multiple speakers lead to unclear communication, making it difficult for participants to listen and understand each other, which results in inefficiencies and the need for repeated utterances.
Innovation Solution
A display method that shows side-by-side images of multiple terminals in a first region, with text images indicating the content of overlapping voices displayed in association with each image. When a user moves a text image to a second region, it is displayed there, allowing for easier visualization of overlapping speech content.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If priority-based voice level adjustment is used, then clarity of high-priority speaker is improved, but understanding of other speakers deteriorates
Solution Approach 1:
The patent segments the display area into multiple regions, with the first region showing images of all speakers and the second region displaying text images of overlapping voices. This segmentation allows participants to simultaneously see who is speaking and what is being said, resolving the contradiction between voice clarity and utterance understanding.
Solution Approach 2:
The patent introduces text images as an intermediary element that converts audio content into visual form. These text images are displayed in the second region and can be moved to the first region, providing a visual mediator that helps participants understand overlapping utterances without compromising the ability to identify active speakers.
2Loss of information
If text images of overlapping voices are displayed in the first region, then utterance understanding is improved, but visual clutter increases
Solution Approach 1:
By separating the display into two regions - the first region for speaker images and the second region for text images of overlapping voices - the patent reduces visual clutter in the primary speaker view while still providing comprehensive utterance information in the secondary region.
Solution Approach 2:
The patent implements dynamic interaction where text images can be moved between the first and second regions by user operation. This dynamic repositioning allows participants to adjust the display based on their needs, reducing clutter in the first region when necessary while maintaining accessibility to utterance information.
Data Source
AI summary
A display method includes displaying, side by side, in a first region, a first image corresponding to a first terminal and a second image corresponding to a second terminal, when a first voice detected by the first terminal and a second voice detected by the second terminal overlap, displaying a first text image indicating content of the first voice in the first region in association with the first image and displaying a second text image indicating content of the second voice in the first region in association with the second image, and, when receiving operation for moving the first text image to a second region different from the first region, displaying the first text image in the second region.


