Voice Credential Recognition via Segmented Input Processing
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current voice recognition systems face challenges in accurately identifying voice inputs providing credentials, which often include a mix of characters, words, and commands, leading to difficulties in hands-free entry and increased user effort.
Innovation Solution
A method and system for identifying voice inputs on a user device, involving the reception of voice inputs, identification of characters and phrases, and conversion to text, with the text displayed in the sequence corresponding to the voice input, utilizing a digital assistant that can interpret natural language and perform tasks based on user intent.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If speech recognition is used for credential entry, then hands-free operation is enabled, but recognition accuracy deteriorates due to mixed input of characters, words, and phrases
Solution Approach 1:
The system segments the credential input process into distinct phases: collecting voice inputs during a time window, identifying whether each input represents a character, word, or phrase, and processing them through appropriate recognition pathways. This segmentation allows the system to handle mixed input types (characters, words, phrases) separately, improving overall recognition accuracy while maintaining hands-free operation.
2Reliability
If voice inputs are converted to text for display, then user verification is improved, but processing complexity increases
Solution Approach 1:
The system performs preliminary actions by collecting and pre-processing voice inputs during a defined time window before final credential verification. Voice inputs are captured, transcribed to text, and displayed in sequence during this preliminary phase, allowing the system to prepare multiple candidate credentials for verification without adding complexity to the core authentication logic.
Data Source
Figure 1
Figure 2A
Figure 2B
AI summary
Systems and processes for identifying of a voice input providing one or more user credentials are provided. In one example process, a voice input can be received. A first character, a phrase identifying a second character, and a word can be identified based on the voice input. In response to the identification, the first character, the second character, and the word can be converted to text. The text can be caused to display, with a display, in a sequence corresponding to an order of the first character, the second character, and the word in the voice input.