Voice Classification for Whisper-Mode Word Replacement
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Voice control devices often inadvertently disclose user notifications to people other than the intended recipient, compromising privacy when the user wishes to keep information private.
Innovation Solution
An output-content control device that analyzes user voice inputs to detect intentions and generates output sentences based on notification information, replacing specific words when the voice is determined to be a whisper, making the content difficult for others to understand.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If voice control devices output notification information to users, then users receive necessary information feedback, but people other than the user can also hear the notification, compromising privacy
Solution Approach 1:
The patent applies local quality by detecting the user's voice volume level and selectively applying word replacement only when the voice is determined to be a whisper. The output-content generating unit replaces specific words in notification information with alternative words only under this specific condition, rather than uniformly for all notifications. This localized application of word replacement based on voice characteristics achieves privacy protection precisely when needed (whispered queries) while maintaining normal information delivery for regular voice inputs.
2Measurement precision
If notification information is output clearly to the user, then the user understands the information accurately, but the same clear output allows others to understand the content, losing privacy
Solution Approach 1:
The patent implements dynamics by making the output content adaptive based on the detected voice characteristics. The voice classifying unit dynamically determines whether the input voice is a whisper, and the output-content generating unit dynamically adjusts the notification content accordingly - using original clear wording for normal voice inputs and replaced ambiguous wording for whispered inputs. This dynamic adjustment allows the system to optimize between clarity and privacy based on real-time voice analysis.
Solution Approach 2:
The patent applies parameter changes by modifying the linguistic parameters of the output notification based on the voice volume parameter detected from the user's input. When the voice volume indicates a whisper, the system changes the semantic parameters of the output by replacing specific words with alternative words that convey the same information to the user but are less intelligible to others. This parameter transformation resolves the contradiction between clear communication and privacy protection.
3Object-affected harmful factors
If word replacement is applied to notification output, then privacy is protected from others, but the user may have difficulty understanding the modified content
Solution Approach 1:
The patent applies copying by maintaining the original notification information as a reference and creating a modified version for output. The output-content generating unit generates a copy of the notification with specific words replaced by alternative words that have similar or related meanings. This copying approach ensures that the essential information is preserved and can be understood by the user while the specific wording changes prevent others from easily comprehending the content. The system effectively creates a privacy-protected copy rather than completely altering the meaning.
Data Source
AI summary
An output-content control device includes a voice classifying unit configured to analyze a voice spoken by a user and acquired by a voice acquiring unit to determine whether the voice is a predetermined voice; an intention analyzing unit configured to analyze the voice acquired by the voice acquiring unit to detect intention information indicating what kind of information is wished to be acquired by the user; a notification-information acquiring unit configured to acquire notification information to be notified to the user based on the intention information; and an output-content generating unit configured to generate an output sentence as sentence data to be output to the user based on the notification information and also configured to generate the output sentence in which at least one word selected among words included in the notification information is replaced with another word when the voice is determined to be the predetermined voice.


