Voice Command Editing via Incomplete Field Highlighting

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current speech-to-text systems lack the ability to efficiently edit and complete voice commands for actions, such as sending emails, as they do not effectively handle incomplete data fields or provide users with clear guidance on providing complete syntax information.

Innovation Solution

A system that generates a graphical user interface (GUI) to prompt users for additional voice input, allowing them to edit and complete voice commands by highlighting incomplete syntax fields and providing training actions to ensure accurate command execution.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Speed

If speech-to-text systems automatically convert voice input to text, then conversion speed is improved, but accuracy of complete command execution deteriorates due to incomplete data fields

Engineering Contradiction:
Improvevoice-to-text conversion speedVSAvoidcommand execution accuracy
Core Design Contradiction:
SpeedVSReliability

Solution Approach 1:

The system performs preliminary analysis of the voice command structure before full execution, identifying incomplete data fields in advance. It then proactively presents guidance interfaces to users to supply missing information before the command is finalized, ensuring complete and accurate command execution while maintaining fast conversion speed.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system provides real-time feedback to users about the completeness of their voice commands by highlighting incomplete data fields. This feedback loop allows users to understand what information is missing and provides guidance on how to supply it, thereby improving command execution accuracy without sacrificing conversion speed.

Inventive Principle:
Principle #23Feedback

2Reliability

If the system requests additional voice input for incomplete fields, then command accuracy is improved, but user interaction time increases

Engineering Contradiction:
Improvecommand execution accuracyVSAvoiduser interaction time
Core Design Contradiction:
ReliabilityVSLoss of time

Solution Approach 1:

The system applies partial action by selectively requesting only the specific missing information needed to complete the command, rather than requiring users to restate the entire command. This targeted approach minimizes additional user interaction time while still achieving complete and accurate command execution.

Inventive Principle:
Principle #16Partial or excessive action

3Ease of operation

If the system provides detailed guidance interfaces, then user understanding is improved, but interface complexity increases

Engineering Contradiction:
Improveuser understandingVSAvoidinterface complexity
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The guidance interface is segmented into discrete, focused elements that address each incomplete data field separately. Rather than presenting a complex monolithic interface, the system breaks down the guidance into manageable segments corresponding to specific missing information, making it easier for users to understand and provide the required input.

Inventive Principle:
Principle #1Segmentation

Data Source

PatentUS9111539B1Editing voice input
Publication Date: 2015.08.18 GOOGLE LLC
  • US9111539B1 patent drawing
  • US9111539B1 patent drawing
  • US9111539B1 patent drawing

AI summary

A computer-implemented method of generating a voice command to perform an action includes receiving a voice request to perform the action, wherein the voice request comprises first audio information for one or more first data fields associated with the action; generating a GUI that when rendered on a display device comprises a prompt message prompting a user to speak second audio information for one or more second data fields associated with the action; and inserting into the one or more second data fields data indicative of one or more of (i) the first audio information, and (ii) the second audio information.