Electronic Device Predicts User Utterance to Reduce Response Latency
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current electronic devices experience delays in responding to user utterances due to the series of processes involved in analyzing voice data, leading to increased user waiting times and reduced reliability of speech recognition services.
Innovation Solution
An electronic device equipped with a communication module, microphone, wake-up recognition modules, and a processor that predicts user utterances to prepare responses in advance, reducing waiting times by transmitting recorded data to an external device for processing and receiving predicted response information.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If the electronic device performs complete voice data analysis using linguistic models and algorithms, then the response accuracy is improved, but the response time increases
Solution Approach 1:
The electronic device performs preliminary voice data analysis using wake-up recognition modules before the user actually utteres the command. By pre-processing and predicting potential user inputs, the system prepares response data in advance, so when the user speaks, the device can quickly retrieve and deliver the pre-computed response, thus maintaining high accuracy while significantly reducing waiting time.
Solution Approach 2:
The speech recognition process is divided into multiple segments: wake-up recognition module for initial detection, linguistic model for semantic understanding, and algorithm processing for response generation. Each segment operates independently and can be processed in parallel or pre-computed, allowing the system to maintain accuracy while reducing overall response time through efficient task division.
2Reliability
If the electronic device processes voice data through multiple analysis stages, then the speech recognition reliability is improved, but the user perception of normal operation deteriorates
Solution Approach 1:
The system performs voice data analysis and prepares responses in advance before the user actually issues the command. By pre-processing potential user inputs and having responses ready, the device creates the perception of immediate responsiveness while still maintaining reliable multi-stage analysis in the background, thus improving both reliability and user perception simultaneously.
Data Source
AI summary
In an embodiment of the disclosure, disclosed is an electronic device including a communication module, a microphone, a first and a second wake-up recognition module, a memory, and a processor. The processor is configured to receive a first user utterance through the microphone, recognize the first user utterance based on at least one of the first or the second wake-up recognition module, when the recognized first user utterance includes specified at least one first trigger information, record at least part of the first user utterance by activating the recording function, transmit recorded data to an external device, and receive at least one of second user utterance information, which is predicted to occur at a time after the function of the speech recognition service is activated by the first wake-up recognition module, or at least one response information associated with the second user utterance from the external device.


