Voice Authentication and Command Processing in Locked Devices
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing voice authentication and command systems for electronic devices require separate steps for authentication and command execution, which can be time-consuming and inconvenient, especially when the device is in a locked mode.
Innovation Solution
A method that allows a user to use a single vocal utterance for both authentication and command execution, where the device processes the utterance to determine the action and authenticate the user simultaneously, eliminating the need for additional input once authentication is successful.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If separate steps are used for authentication and command execution, then security and reliability are improved, but time consumption and operational complexity increase
Solution Approach 1:
The patent combines authentication and command execution into a single integrated voice processing step. The voice biometric engine and voice action engine process the utterance simultaneously, allowing the user to authenticate and issue a command in one continuous action rather than requiring separate authentication followed by separate command input.
Solution Approach 2:
The voice utterance serves multiple functions simultaneously: it acts as both the authentication credential (biometric template) and the command input. This multi-functionality eliminates the need for separate authentication and command steps, reducing time loss while maintaining security through the voice biometric verification process.
2Reliability
If separate steps are used for authentication and command execution, then security is improved, but ease of operation deteriorates
Solution Approach 1:
The patent merges authentication and command execution into a single integrated process where the voice utterance is processed by both the voice biometric engine and voice action engine simultaneously. This eliminates the need for users to complete separate authentication steps before issuing commands, significantly improving ease of operation while maintaining security through biometric verification.
Solution Approach 2:
The voice utterance performs dual functions as both authentication credential and command input. This multi-functionality allows users to authenticate and issue commands in a single natural speech act, making the system much easier to operate compared to traditional separate-step approaches.
3Reliability
If multiple inputs are required for authentication and command, then security is improved, but device complexity increases
Solution Approach 1:
The patent combines multiple processing functions into a single integrated voice processing pipeline. The voice utterance is simultaneously provided to the voice biometric engine for authentication and the voice action engine for command interpretation, eliminating the need for separate processing stages and reducing overall system complexity.
Solution Approach 2:
The voice processing system is designed to handle multiple functions through a single unified input stream. The same voice utterance is processed for both authentication and command execution, reducing the number of separate input channels and processing paths needed, thereby simplifying the overall device architecture.
Data Source
AI summary
Methods, systems, and apparatus for voice authentication and command. In an aspect, a method comprises: receiving, by a data processing apparatus that is operating in a locked mode, audio data that encodes an utterance of a user, wherein the locked mode prevents the data processing apparatus from performing at least one action; providing, while the data processing apparatus is operating in the locked mode, the audio data to a voice biometric engine and a voice action engine; receiving, while the data processing apparatus is operating in the locked mode, an indication from the voice biometric engine that the user has been biometrically authenticated; and in response to receiving the indication, triggering the voice action engine to process a voice action that is associated with the utterance.


