Voice Credential Recognition via Segmented Input Processing

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current voice recognition systems face challenges in accurately identifying voice inputs providing credentials, which often include a mix of characters, words, and commands, leading to difficulties in hands-free entry and increased user effort.

Innovation Solution

A method and system for identifying voice inputs on a user device, involving the reception of voice inputs, identification of characters and phrases, and conversion to text, with the text displayed in the sequence corresponding to the voice input, utilizing a digital assistant that can interpret natural language and perform tasks based on user intent.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If speech recognition is used for credential entry, then hands-free operation is enabled, but recognition accuracy deteriorates due to mixed input of characters, words, and phrases

Engineering Contradiction:
Improvehands-free operationVSAvoidrecognition accuracy
Core Design Contradiction:
Ease of operationVSMeasurement precision

Solution Approach 1:

The system segments the credential input process into distinct phases: collecting voice inputs during a time window, identifying whether each input represents a character, word, or phrase, and processing them through appropriate recognition pathways. This segmentation allows the system to handle mixed input types (characters, words, phrases) separately, improving overall recognition accuracy while maintaining hands-free operation.

Inventive Principle:
Principle #1Segmentation

2Reliability

If voice inputs are converted to text for display, then user verification is improved, but processing complexity increases

Engineering Contradiction:
Improveuser verificationVSAvoidprocessing complexity
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

The system performs preliminary actions by collecting and pre-processing voice inputs during a defined time window before final credential verification. Voice inputs are captured, transcribed to text, and displayed in sequence during this preliminary phase, allowing the system to prepare multiple candidate credentials for verification without adding complexity to the core authentication logic.

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentEP3394852B1Identification of voice inputs providing credentials
Publication Date: 2022.09.21 APPLE INC
  • EP3394852B1 patent drawingFigure 1
  • EP3394852B1 patent drawingFigure 2A
  • EP3394852B1 patent drawingFigure 2B

AI summary

Systems and processes for identifying of a voice input providing one or more user credentials are provided. In one example process, a voice input can be received. A first character, a phrase identifying a second character, and a word can be identified based on the voice input. In response to the identification, the first character, the second character, and the word can be converted to text. The text can be caused to display, with a display, in a sequence corresponding to an order of the first character, the second character, and the word in the voice input.