Speech Processing Device for Hands-Free Secret Info
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Users in public spaces face privacy concerns when providing private or secret information via speech communication, as existing solutions often require non-verbal modes like text input, which are not hands-free.
Innovation Solution
A method and device that use keyword detection, an audio database, and audio signal processing for audio splicing or synthesis, allowing secret information to be inserted into a voice data stream without requiring users to pronounce it aloud, using a combination of audio scan and detection, substitution determination, and insertion units to create a seamless substituted data stream.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Productivity
If users provide private information via speech communication in public spaces, then communication efficiency is improved, but privacy security deteriorates
Solution Approach 1:
The patent introduces an intermediary system that processes speech input through keyword detection and audio substitution. When a delimiter keyword is detected, the system automatically substitutes it with pre-recorded audio containing sensitive information, preventing the user from directly pronouncing private data in public spaces while maintaining seamless communication flow.
2Object-affected harmful factors
If users switch to non-verbal input modes like text entry, then privacy security is improved, but ease of operation deteriorates
Solution Approach 1:
The patent replaces the mechanical action of typing or manual text entry with an automated audio substitution system. The system detects delimiter keywords in speech and automatically inserts pre-recorded sensitive information, eliminating the need for manual keyboard input while maintaining privacy security.
3Object-affected harmful factors
If users are required to type secret information on a keyboard, then privacy security is improved, but productivity deteriorates
Solution Approach 1:
The system acts as an intermediary between the user's speech and the communication system. By detecting delimiter keywords and automatically substituting them with pre-recorded sensitive information, it eliminates the need for manual typing while maintaining privacy security, thus restoring communication efficiency.
4Ease of operation
If a hands-free solution is implemented, then ease of operation is improved, but device complexity increases
Solution Approach 1:
The patent implements preliminary action by requiring users to pre-record sensitive information and associate it with delimiter keywords during a setup phase. This pre-prepared audio database enables hands-free operation during actual use without requiring complex real-time processing, as the system simply needs to detect keywords and retrieve pre-stored audio segments.
Data Source
AI summary
During telephone calls in a public space, users may be reluctant to provide private or secret information, due to the risk of eavesdropping. A hands-free solution for entering secret information to electronic speech communication devices is based on speech processing. A method for speech processing a voice input data stream comprises steps of scanning the voice input data stream and detecting a spoken delimiter therein, determining a predefined audio sample corresponding to the detected spoken delimiter, inserting the determined predefined audio sample into the voice input data stream at the spoken delimiter, wherein a substituted voice data stream is obtained and wherein speech portions of the voice input data stream at least before the spoken delimiter remain in the substituted voice data stream, and providing the substituted voice data stream for output towards a recipient.

