Smart Speaker Output Method Prioritizing Voice Replies
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional smart speaker technologies require users to interact solely through voice, limiting the ability to obtain and present diverse information effectively, such as daily information, weather, or environmental conditions, which can be cumbersome and unsatisfactory in providing comprehensive feedback.
Innovation Solution
An output method and electronic device configuration that processes voice information to determine if it is a request, obtaining reply and supplemental information, and transmitting both to an output device for prioritized and diversified feedback, including environmental sounds or additional data, to enhance user interaction and experience.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If voice interaction is used for all information queries, then the system maintains simple operation, but the information completeness and user experience deteriorate
Solution Approach 1:
The patent segments the information output into two distinct types: reply information (direct answers to voice queries) and supplemental information (additional contextual data). This segmentation allows the system to maintain simple voice interaction while delivering comprehensive information through multiple channels including text, images, and audio, thereby resolving the contradiction between operational simplicity and information completeness.
Solution Approach 2:
The patent transitions from a single-dimension voice-only interaction to multi-dimensional output by incorporating text displays, image presentations, and supplemental audio information. This dimensional expansion enables the system to preserve voice interaction simplicity while significantly enhancing information completeness through additional output modalities.
2Loss of time
If only reply information is provided, then the response is quick and simple, but the contextual understanding and user immersion deteriorate
Solution Approach 1:
The patent applies preliminary action by pre-processing and categorizing information into reply and supplemental components before output. This allows the system to deliver immediate reply information for quick responses while simultaneously preparing and delivering supplemental contextual information, thereby maintaining fast response times while enhancing contextual understanding.
Solution Approach 2:
The patent introduces supplemental information as an intermediary element that bridges the gap between brief voice queries and comprehensive understanding. This intermediary layer provides additional context, background information, and related data without interfering with the speed of the primary reply, thus resolving the contradiction between quick responses and contextual richness.
3Loss of information
If multiple information types are output simultaneously, then the information completeness improves, but the device complexity increases
Solution Approach 1:
The patent implements universality by designing an output system that can handle multiple information types (text, images, audio) through a unified processing framework. The system uses a single processor that can generate and manage different output formats, and a versatile output device capable of presenting various information types, thereby achieving information completeness without proportionally increasing device complexity.
Solution Approach 2:
The patent merges the output of reply information and supplemental information into a coordinated presentation system. By combining text displays, image presentations, and audio outputs into an integrated system managed by a unified processor, the patent achieves comprehensive information delivery while avoiding the complexity that would arise from separate independent systems for each output type.
Data Source
AI summary
An output method includes obtaining voice information, determining whether the voice information is a voice request, in response to the voice information being the voice request, obtaining reply information for replying to the voice request and supplemental information, transmitting the reply information and the supplemental information to an output device for outputting the reply information and the supplemental information using different parameters, such that an output of the reply information is prioritized over an output of the supplemental information, and in response to receiving a predetermined operation, outputting the reply information and the supplemental information using different parameters, such that the output of the supplemental information is prioritized over the output of the reply information. The supplemental information is information that needs to be outputted in association with the reply information.


