Program Image Creation with Two-Way Voice Interaction
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional program images lack two-way communication capabilities, failing to dynamically change based on viewer reactions, limiting interactive engagement.
Innovation Solution
A program image creation method and apparatus that synchronizes voice input information with image, avatar, and decoration selections, allowing for real-time interaction and adaptation based on viewer input, including voice processing, avatar combination, decoration integration, and hyperlink setup, enabling dynamic image changes and enhanced viewer participation.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If conventional program images are distributed through one-way communication, then the system is simple and reliable, but the viewer cannot interact or provide feedback to change the image content
Solution Approach 1:
The program image is divided into multiple independent layers including background layer, character layer, decoration layer, and UI layer. Each layer can be independently controlled and modified based on viewer interaction, allowing the system to maintain simplicity in individual components while achieving complex interactive capabilities through their combination
Solution Approach 2:
The program image transitions from a static one-way distribution model to a dynamic multi-layered structure where each layer can be independently adjusted, animated, or modified in real-time based on viewer feedback, enabling adaptive content delivery without requiring complete system redesign
2Ease of operation
If a program image allows two-way communication with viewer feedback, then viewer engagement increases, but the processing and synchronization of multiple input types becomes more complex
Solution Approach 1:
The program image is designed as a universal container that can accept multiple types of viewer input (voice, text, touch) and process them through a unified event handling system. The same layered structure handles both traditional click interactions and new voice-based interactions, reducing the need for separate processing paths for different input types
Solution Approach 2:
An intermediary processing layer is introduced that receives various input types (voice commands, text input, touch gestures) and converts them into standardized events that the layered image structure can uniformly process. This mediator simplifies the complexity by providing a single interface between diverse input sources and the image processing system
3Adaptability or versatility
If multiple avatars and decoration materials are prepared for personalization, then the program image can be customized according to viewer preferences, but the data storage and management requirements increase
Solution Approach 1:
Instead of providing complete alternative images for each customization option, the system divides the program image into layers with specific local characteristics. Each layer contains only the necessary elements for its function (e.g., avatar layer contains only character elements, decoration layer contains only background elements), reducing redundant data storage while maintaining full personalization capability
Solution Approach 2:
The program image structure implements a nested layered architecture where avatars, decorations, and other elements are organized in hierarchical layers. Each layer can contain sub-layers or nested elements that can be independently controlled, allowing comprehensive customization options to be represented through combinations of smaller data units rather than requiring complete separate image files for each variation
Data Source
AI summary
A program image creation method that allows two-way communication for communication in a format of questions and answers. The method includes: a description image processing step of setting a description image based on image selection information; a voice processing step of synchronizing a voice from voice input information with the description image; an avatar processing step of combining an avatar that is set based on avatar selection information with the description image; a decoration processing step of combining a decoration material that is set based on decoration selection information with the description image; and an interactive processing step of setting a hyperlink based on interactive selection information.


