Voice Command Editing via Incomplete Field Highlighting
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current speech-to-text systems lack the ability to efficiently edit and complete voice commands for actions, such as sending emails, as they do not effectively handle incomplete data fields or provide users with clear guidance on providing complete syntax information.
Innovation Solution
A system that generates a graphical user interface (GUI) to prompt users for additional voice input, allowing them to edit and complete voice commands by highlighting incomplete syntax fields and providing training actions to ensure accurate command execution.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Speed
If speech-to-text systems automatically convert voice input to text, then conversion speed is improved, but accuracy of complete command execution deteriorates due to incomplete data fields
Solution Approach 1:
The system performs preliminary analysis of the voice command structure before full execution, identifying incomplete data fields in advance. It then proactively presents guidance interfaces to users to supply missing information before the command is finalized, ensuring complete and accurate command execution while maintaining fast conversion speed.
Solution Approach 2:
The system provides real-time feedback to users about the completeness of their voice commands by highlighting incomplete data fields. This feedback loop allows users to understand what information is missing and provides guidance on how to supply it, thereby improving command execution accuracy without sacrificing conversion speed.
2Reliability
If the system requests additional voice input for incomplete fields, then command accuracy is improved, but user interaction time increases
Solution Approach 1:
The system applies partial action by selectively requesting only the specific missing information needed to complete the command, rather than requiring users to restate the entire command. This targeted approach minimizes additional user interaction time while still achieving complete and accurate command execution.
3Ease of operation
If the system provides detailed guidance interfaces, then user understanding is improved, but interface complexity increases
Solution Approach 1:
The guidance interface is segmented into discrete, focused elements that address each incomplete data field separately. Rather than presenting a complex monolithic interface, the system breaks down the guidance into manageable segments corresponding to specific missing information, making it easier for users to understand and provide the required input.
Data Source
AI summary
A computer-implemented method of generating a voice command to perform an action includes receiving a voice request to perform the action, wherein the voice request comprises first audio information for one or more first data fields associated with the action; generating a GUI that when rendered on a display device comprises a prompt message prompting a user to speak second audio information for one or more second data fields associated with the action; and inserting into the one or more second data fields data indicative of one or more of (i) the first audio information, and (ii) the second audio information.


