Wearable Message Composition via Cloud Speech Inference
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current wearable multimedia devices require extensive user interaction and resources to compose and transmit electronic messages, especially when using speech inputs, as they need explicit specification of content and recipients, leading to inefficiencies in processing and battery usage.
Innovation Solution
A wearable multimedia device with a cloud computing platform and application ecosystem that automatically infers content and processes multimedia data, allowing users to compose and transmit messages using fewer and shorter speech inputs by leveraging contextual information and machine learning algorithms, reducing the need for explicit content specification and optimizing resource usage.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If speech input is used to compose electronic messages, then ease of operation is improved, but device complexity increases due to need for speech processing and content inference systems
Solution Approach 1:
The patent introduces a cloud-based assistant service as an intermediary between the user's speech input and the message composition system. The wearable device captures speech and transmits it to the assistant service, which performs speech-to-text conversion, content inference, and message generation. This mediator handles the complex processing tasks remotely, reducing the burden on the wearable device's local resources while still providing hands-free message composition.
Solution Approach 2:
The patent replaces traditional mechanical input methods (typing, button pressing) with acoustic input (speech). The microphone subsystem captures speech signals, which are then converted to text by the assistant service. This substitution eliminates the need for physical interaction with the device, significantly improving ease of operation for message composition.
2Productivity
If automatic content inference is implemented, then productivity is improved by reducing input time, but use of energy increases due to additional processing requirements
Solution Approach 1:
The patent extracts the computationally intensive speech processing and content inference tasks from the wearable device and relocates them to the cloud-based assistant service. The wearable device only performs lightweight functions such as speech capture, packet formation, and transmission. This extraction allows automatic content inference to improve productivity while minimizing the energy consumption on the wearable device, as the heavy processing occurs remotely on the assistant service's servers.
3Speed
If speech-to-text conversion is performed locally, then speed of processing is improved, but device complexity and resource requirements worsen
Solution Approach 1:
The assistant service acts as an intermediary that handles speech-to-text conversion remotely. The wearable device transmits captured speech to the assistant service, which performs the conversion and returns the text. This approach maintains fast processing by leveraging the assistant service's dedicated speech recognition resources, while avoiding the need for complex speech processing hardware and software on the wearable device itself.
Data Source
AI summary
Systems, methods, devices and non-transitory, computer-readable storage mediums are disclosed for a wearable multimedia device and cloud computing platform with an application ecosystem for processing multimedia data captured by the wearable multimedia device. In an embodiment, a wearable multimedia device receives a first speech input from a user, including a first command to generate a message, and first content for inclusion in the message. The device determines second content for inclusion in the message based on the first content, and generates the message such that the messages includes the first and second content. The device receives a second speech input from the user, including a second command to modify the message. In response, the device determines third content for inclusion in the message based on the first content and/or the second content, and modifies the message using the third content. The device transmits the modified message to a recipient.


