Voice Assistant Tracking for Updated Response Relevance
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing AI assistant technologies in electronic devices provide responses only based on user input, failing to recognize changes in information unless the user actively requests it, leading to outdated responses.
Innovation Solution
The electronic device includes a processor that identifies tracking elements in user voice inputs, stores relevant text and natural language understanding results, and continuously updates responses based on time and context to ensure accuracy.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If the AI assistant operates only based on user input, then the device complexity is reduced, but the reliability of information provision deteriorates because users fail to receive updated information when specific information changes
Solution Approach 1:
The system performs preliminary actions by storing the user's original voice input, corresponding text, and natural language understanding results before any information changes occur. This preliminary storage enables the system to later compare updated information against the original query context, ensuring users receive relevant updates without needing to rephrase their requests.
Solution Approach 2:
The system implements a feedback mechanism where it continuously monitors for changes in specific information related to the user's original query. When changes are detected, the system compares the updated information against the stored original understanding results and provides follow-up responses, creating a closed-loop feedback system that maintains information reliability.
2Reliability
If the AI assistant continuously verifies response appropriateness, then the reliability of information provision is improved, but the loss of time increases due to continuous verification processes
Solution Approach 1:
The system performs preliminary analysis by storing the natural language understanding results of the user's original voice input, including key entities and context. This preliminary processing eliminates the need for time-consuming re-analysis when verifying if updated information remains relevant to the user's original intent.
Solution Approach 2:
The system creates a copy of the original user voice, corresponding text, and understanding results for comparison purposes. By copying the original context rather than re-processing the user's input, the system efficiently verifies response appropriateness without incurring the full time cost of complete re-analysis.
3Productivity
If the AI assistant provides immediate responses based on user voice, then the productivity is improved, but the loss of information occurs because users fail to recognize changes in specific information
Solution Approach 1:
The system implements a feedback mechanism that actively monitors for changes in specific information related to the user's original query. When changes are detected, the system automatically provides follow-up responses to inform users of the updates, ensuring information change recognition without requiring users to actively seek out changes.
Solution Approach 2:
The system performs self-service by autonomously monitoring information sources for changes related to the user's stored queries and automatically notifying users of relevant updates, eliminating the need for users to manually re-query or check for information changes.
Data Source
AI summary
An electronic device includes a microphone, a memory, and a processor configured to obtain a first natural language understanding result for a first user voice obtained through the microphone based on a first text corresponding to the first user voice, provide a first response corresponding to the first user voice based on the first natural language understanding result, identify whether the first user voice includes a tracking element based on the first natural language understanding result and a second text corresponding to the first response, based on identifying that the first user voice includes the tracking element, store the first text, the first natural language understanding result and the second text in the memory, and obtain a third text corresponding to the first response based on the first natural language understanding result.


