Voice-Driven Dynamic Image Positioning in Web Conferences
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing web conference systems do not dynamically change displayed presentation materials in real time based on bidirectional conversations between participants.
Innovation Solution
A system that includes a voice information input unit, a voice analysis unit, and an image change unit, which analyzes voice information to identify content and changes in content, and uses this information to dynamically change the position, shape, and color of content in shared images in real time.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If a general web conference system displays presentation materials prepared by a speaker, then the presentation can be shared with participants, but the displayed material does not change even when participants make statements regarding the material
Solution Approach 1:
The system segments the presentation material into multiple independent content elements (images, texts, charts) that can be individually manipulated. Each content element can be independently positioned, resized, or modified based on voice input, allowing dynamic reconfiguration without changing the entire presentation structure.
Solution Approach 2:
The presentation material transitions from a static displayed state to a dynamic state where content elements can be automatically repositioned and reconfigured in real-time based on voice analysis. The system continuously monitors voice inputs and dynamically adjusts the display without requiring manual intervention or system reconfiguration.
2Productivity
If voice information is analyzed to change content position in real time, then interactive presentation capability is enhanced, but processing time and computational resources increase
Solution Approach 1:
The system performs preliminary analysis of voice information by continuously monitoring and pre-processing audio inputs even before specific content repositioning is triggered. Voice patterns, keywords, and intent are pre-identified so that when repositioning is needed, the system can execute immediately without delay.
Solution Approach 2:
The system implements selective processing by skipping detailed analysis of voice inputs that do not require content repositioning. Only specific trigger words, phrases, or patterns initiate the full repositioning sequence, allowing the system to rapidly process and respond to critical inputs while ignoring routine conversational elements.
Data Source
AI summary
To provide a system that changes a shared image in real time based on a conversation. A system for changing an image based on a voice that includes a voice information input unit configured to input voice information, a voice analysis unit configured to analyze the voice information input by the voice information input unit, and an image change unit configured to change a position of content in an image representing the content using information on the content included in the voice information analyzed by the voice analysis unit and information on a change in the content.


