Voice Authentication With Password Content and Voice Pattern Matching
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional authentication methods, such as passwords and biometrics, are either easily discoverable, intrusive, or difficult for users to manage, compromising the security of sensitive digital resources and user experience.
Innovation Solution
A multi-dimensional voice-based digital authentication system that uses a central server to verify a spoken password expression, checking both the content and voice pattern representation against stored data, ensuring multiple points of authentication.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If conventional password authentication is used, then ease of operation is improved, but reliability deteriorates due to ease of discovery and circumvention
Solution Approach 1:
The patent combines two authentication factors into a single voice input: the password phrase (knowledge factor) and the voice pattern biometric (biological factor). The system extracts both the semantic content (password verification) and acoustic characteristics (voice print verification) from the same spoken input, merging knowledge-based and biometric authentication into one unified process that enhances security without requiring additional user actions
Solution Approach 2:
The patent transitions from traditional 2D password authentication (text-based) to multi-dimensional voice authentication by adding acoustic, spectral, and temporal dimensions. The voice pattern analysis extracts features across multiple dimensions including frequency spectrum, time-domain characteristics, and spectral envelope, creating a high-dimensional authentication space that is significantly harder to compromise than traditional passwords
2Reliability
If biometric authentication such as fingerprint or iris scanners is used, then reliability is improved, but device complexity and ease of operation worsen due to additional equipment and intrusiveness
Solution Approach 1:
The patent makes the voice authentication system universally applicable across multiple platforms and devices by using only the microphone, which is a standard component in virtually all modern computing devices. The same voice-based authentication mechanism can be deployed on desktops, laptops, mobile phones, tablets, and servers without requiring device-specific hardware modifications, unlike fingerprint scanners or iris cameras
Solution Approach 2:
The patent creates a digital copy of the user's voice pattern characteristics and stores it as a template in the database. This voice print template serves as a replicable authentication credential that can be verified across different sessions and devices without requiring the physical presence of the original biometric sensor hardware used in other biometric systems
3Reliability
If biometric authentication such as fingerprint or iris scanners is used, then reliability is improved, but ease of operation worsens due to intrusiveness
Solution Approach 1:
The patent enables the system to automatically extract and verify both password and voice pattern features from the user's spontaneous speech without requiring the user to perform separate actions. The system autonomously performs speech-to-text conversion, voice feature extraction, pattern matching, and authentication decision-making, making the process as effortless as simply speaking the password while achieving enhanced security
Data Source
AI summary
Systems and methods for multi-dimensional voice-based digital authentication are provided. Methods include receiving, at a central server from a remote user device, a request to access a protected digital resource; displaying a prompt, at the device, requesting a spoken password expression; capturing, at the device, an expression spoken in response to the prompt; determining that the expression satisfies predetermined attributes of the password expression; calculating a voice pattern representation of the expression; determining that the voice pattern representation of the expression satisfies a threshold similarity to a voice pattern representation of the user that is stored in a database at the central server; and, in response to the expression satisfying the predetermined attributes of the password expression and satisfying the threshold similarity to the voice pattern representation of the user that is stored in the database at the central server, authorizing the device to access the protected digital resource.


