Parsing Verbal Expressions for Multiple-Goal Command Execution
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing voice-search interfaces struggle to accurately interpret and execute complex, multiple-goal commands due to limitations in parsing verbal expressions with proper punctuation and lack of visual feedback, making it difficult for users to efficiently direct computing devices to perform multiple tasks.
Innovation Solution
The system analyzes verbal expressions by parsing terms into categories, examining temporal distribution, and comparing confidence levels of non-overlapping terms to identify potential multiple-goal commands, allowing for user review and editing via various interfaces to ensure accurate execution.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If a voice-search interface is used to specify multiple-goal tasks, then the user gains freedom to formulate commands without visual feedback, but the interface cannot accurately interpret complex multiple-goal commands with proper punctuation
Solution Approach 1:
The patent introduces an intermediary processing layer between voice input and task execution. This layer includes confidence level calculation, temporal distribution analysis, and category-based term grouping that mediates between the ambiguous voice input and the precise task interpretation, resolving the contradiction between ease of use and accuracy
Solution Approach 2:
The system provides feedback by presenting parsed multiple-goal commands to users for review and confirmation before execution. This feedback loop allows users to verify accurate interpretation while maintaining the convenience of voice input, addressing both ease of operation and interpretation accuracy
2Measurement precision
If a text interface is used to specify multiple-goal tasks, then the interface can accurately specify commands with punctuation and operators, but the user must correctly handle an intimidating amount of punctuation and operators
Solution Approach 1:
The patent replaces the mechanical typing system with an acoustic/voice-based system. By substituting the physical act of typing punctuation and operators with natural speech, the system maintains command specification accuracy through voice recognition while dramatically improving ease of operation by eliminating the need to manually handle complex punctuation
3Ease of operation
If existing voice interfaces are used for multiple-goal tasks, then the interface is simple to use, but the user cannot correctly specify multiple-goal tasks at all
Solution Approach 1:
The patent implements dynamic processing that adapts to the complexity of the voice input. The system dynamically calculates confidence levels, performs temporal distribution analysis, and adjusts its interpretation strategy based on the parsed structure, enabling it to handle multiple-goal tasks while maintaining simplicity of use
Solution Approach 2:
The system segments the voice command into individual terms, assigns categories to each term, and analyzes their temporal distribution. This segmentation allows the interface to correctly identify and execute multiple-goal tasks by processing each goal as a separate but related unit, enhancing task specification capability while keeping the interface simple
Data Source
AI summary
A method for parsing a verbal expression received from a user to determine whether or not the expression contains a multiple-goal command is described. Specifically, known techniques are applied to extract terms from the verbal expression. The extracted terms are assigned to categories. If two or more terms are found in the parsed verbal expression that are in associated categories and that do not overlap one another temporally, then the confidence levels of these terms are compared. If the confidence levels are similar, then the terms may be parallel entries in the verbal expression and may represent multiple goals. If a multiple-goal command is found, then the command is either presented to the user for review and possible editing or is executed. If the parsed multiple-goal command is presented to the user for review, then the presentation can be made via any appropriate interface including voice and text interfaces.


