HTML Component Voice Prompt Conversion via Protocol Translator
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional voice systems are unable to interact with web pages that have interactable components, limiting their functionality to reading text only and preventing full utilization of web pages, especially those used for controlling processes or workflows.
Innovation Solution
A method and system that convert HTML components of a web page into voice prompts using a voice attribute file, allowing interaction through a mobile system by transforming HTML components into parameterized data suitable for voice interaction, enabling voice-driven navigation and control of web pages.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If conventional voice systems are used to access web pages, then text reading capability is provided, but interaction with interactable components is lost
Solution Approach 1:
The patent introduces a protocol translator as an intermediary component that converts HTML protocol commands into voiceXML commands. This translator acts as a bridge between the web page server and the voice-based mobile system, enabling the voice system to interact with web page components by translating the interaction protocols between the two different systems.
Solution Approach 2:
The system changes the protocol parameters by transforming HTML protocol commands (such as form submissions, button clicks) into equivalent voiceXML commands. This parameter transformation allows the voice system to execute web page interactions by changing the command format from HTML protocol to voiceXML protocol while maintaining the functional intent.
2Adaptability or versatility
If web pages are accessed using screen-based interface, then full interactivity with components is achieved, but accessibility without visual interface is limited
Solution Approach 1:
The patent replaces the mechanical screen-based interaction system with an acoustic voice-based system. Instead of requiring visual display and manual input through screens, the system substitutes these with voice commands that are processed through the protocol translator to achieve the same interactive effects on web pages.
Solution Approach 2:
The protocol translator provides universal functionality by enabling a single voice-based interface to interact with multiple types of web page components (forms, buttons, links, tables). This multi-functional capability allows the audio interface to perform various interaction tasks that were previously limited to screen-based interfaces.
3Adaptability or versatility
If HTML components are transformed to voice prompts, then voice interaction is enabled, but manual coding effort is required
Solution Approach 1:
The system implements self-service by allowing web pages to automatically generate their own voiceXML representations through the protocol translator. The translator processes standard HTML components and automatically generates the corresponding voiceXML commands without requiring manual intervention, making the transformation process self-executing and reducing coding effort.
Solution Approach 2:
The patent applies preliminary action by pre-configuring the protocol translator with the transformation rules and mappings between HTML and voiceXML protocols. This preliminary setup enables the system to automatically handle the transformation of various HTML components into voice prompts without requiring ad-hoc coding for each transformation case.
Data Source
Figure 1~2
Figure 3
Figure 4~5
AI summary
Embodiments of the invention address the deficiencies of the prior art by providing a method, apparatus, and program product to of converting components of a web page to voice prompts for a user. In some embodiments, the method comprises selectively determining at least one HTML component from a plurality of HTML components of a web page to transform into a voice prompt for a mobile system (16) based upon a voice attribute file associated with the web page. The method further comprises transforming the at least one HTML component into parameterized data suitable for use by the mobile system (16) based upon at least a portion of the voice attribute file (48) associated with the at least one HTML component and transmitting the parameterized data to the mobile system (16).