Dynamic Voice Navigation for GUI Indexing
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing voice recognition software for navigation in graphical user interfaces is inflexible and impractical due to the need for static voice dialogue applications that anticipate all user grammar and vocabulary choices, limiting browsing and navigation, especially in web content with numerous GUI items.
Innovation Solution
A method and system that converts graphical user interface items into voice searchable indices, allowing for dynamic voice navigation by correlating verbal inputs to phonemic representations, enabling just-in-time voice-enabled browsing without pre-defining all possible user inputs.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If static voice dialogue applications are constructed for each GUI view to anticipate all user grammar and vocabulary choices, then voice recognition coverage is improved, but device complexity and ease of operation deteriorate due to the need to pre-define every possible user input
Solution Approach 1:
The patent transitions from static voice dialogue applications to dynamic voice navigation. Instead of pre-defining all possible grammars and vocabularies for each GUI view, the system dynamically generates voice commands based on the current GUI context and item properties. This allows the voice recognition system to adapt to different views and content without requiring manual configuration of every possible user input scenario.
Solution Approach 2:
The patent creates a universal voice navigation system that can handle multiple GUI views and item types through a single dynamic generation mechanism. Rather than creating separate static dialogue applications for each view, the system uses a unified approach that generates appropriate voice commands based on the current context, making the system versatile across different interfaces and content types.
2Ease of manufacture
If static voice navigation systems are used to enable web pages, then implementation simplicity is improved, but adaptability deteriorates because coverage does not extend to the entire webpage
Solution Approach 1:
The system dynamically extends voice navigation coverage to the entire webpage by generating voice commands based on the current view context. As users navigate through different sections of a webpage, the system automatically adapts the voice recognition capabilities to match the current content, ensuring comprehensive coverage without requiring separate static configurations for each section.
Solution Approach 2:
The patent segments the webpage into manageable views and generates voice navigation commands specific to each view's context. By processing the webpage in segments and creating dynamic voice commands for each segment, the system achieves full webpage coverage while maintaining implementation simplicity through a modular approach.
3Measurement precision
If voice recognition software anticipates every grammar and vocabulary choice statically, then recognition accuracy is improved, but ease of operation worsens due to significant impediment to browsing and navigation
Solution Approach 1:
The system maintains high voice recognition accuracy by dynamically generating context-specific voice commands for each GUI view. Instead of relying on static pre-definition of all possible inputs, the system adapts the recognition parameters to the current context, ensuring accuracy while allowing flexible browsing and navigation without impediments.
Data Source
AI summary
A method, apparatus, and electronic device for voice navigation are disclosed. A voice input mechanism 310 may receive a verbal input from a user to a voice user interface program invisible to the user. A processor 104 may identify in a graphical user interface (GUI) a set of GUI items. The processor 104 may convert the set of GUI items to a set of voice searchable indices 400. The processor 104 may correlate a matching GUI item of the set of GUI items to a phonemic representation of the verbal input.


