Adaptive Voice Prompt Adjustment for Driver State
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional user interactive systems fail to effectively communicate with users as they do not adjust their voice prompts based on the user's emotional or physical state, leading to suboptimal interaction.
Innovation Solution
A method and system that analyze user utterances to determine the user's state and adjust voice prompts, including tone, content, speed, prosody, and gender, to match the user's state, using utterance parameter vectors and linear functions to generate appropriate audio waveforms.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If the system uses the same voice prompt tone and content for all users, then the system operation is simple, but the communication effectiveness with users deteriorates when users are in different states
Solution Approach 1:
The voice prompt system transitions from a static, fixed tone and content approach to a dynamic system that automatically adjusts voice characteristics based on real-time detection of user state. The system monitors user utterances and driving conditions, then dynamically modifies tone, content, speed, and prosody to match the user's current state, thereby maintaining communication effectiveness across varying user conditions.
Solution Approach 2:
The system changes multiple voice prompt parameters including tone, content, speed, and prosody based on detected user state. By modifying these acoustic parameters dynamically, the system adapts its communication style to match user conditions such as stress, happiness, or fatigue levels, improving communication effectiveness without requiring manual system reconfiguration.
2Ease of operation
If the system manually changes voice by user choice, then the user can control the voice prompt, but the system cannot automatically adjust to user state
Solution Approach 1:
The system performs self-adjustment by automatically detecting user state through analysis of utterances and driving conditions, then autonomously modifying voice prompt parameters without requiring user intervention. This self-service capability eliminates the need for manual voice selection while maintaining adaptability to user conditions.
Solution Approach 2:
The system implements a feedback loop where user utterances and driving conditions are continuously monitored, user state is determined based on this feedback, and voice prompt parameters are adjusted accordingly. This closed-loop feedback mechanism enables automatic adaptation to changing user states without requiring explicit user commands.
3Reliability
If the system adjusts voice prompt based on user state detection, then communication effectiveness improves, but the device complexity increases
Solution Approach 1:
The system employs a multi-functional approach where a single state detection module handles multiple user states (stress, happiness, fatigue), and the voice adjustment mechanism serves multiple purposes by modifying various voice parameters (tone, content, speed, prosody) simultaneously. This universal design improves communication effectiveness across diverse user conditions while managing system complexity through integrated processing.
Data Source
AI summary
The voice prompt of an interactive system is adjusted based upon a state of a user. An utterance of the user is received, and the state of the user is determined based upon signal processing of the utterance of the user. Once the state of the user is determined, the voice prompt is adjusted by adjusting at least one of a tone of voice of the voice prompt, a content of the voice prompt, a prosody of the voice prompt, and a gender of the voice prompt based upon the determined state of the user.


