Voice-Controlled Device Content Transition to Display

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing voice interaction systems lack the ability to seamlessly transition content between devices, particularly from audio-only output on one device to visual output on another device in response to user voice commands, limiting the user's access to comprehensive information across multiple devices in a connected environment.

Innovation Solution

A voice-controlled device identifies proximate devices with display capabilities and instructs them to output visual content associated with user voice commands, allowing for the transfer of content from one device to another, such as from a summary on a voice-controlled device to detailed information on a tablet or other display-capable device, through a content-transition engine and speech recognition technology.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Loss of information

If content is output only through audio on a voice-controlled device, then the device can maintain simple hardware design, but the user cannot access comprehensive visual information

Engineering Contradiction:
Improveinformation completenessVSAvoiddevice functionality
Core Design Contradiction:
Loss of informationVSDevice complexity

Solution Approach 1:

The system divides the information delivery function across multiple devices: the voice-controlled device handles audio output and voice commands, while a separate display device handles visual content presentation. This segmentation allows each device to remain specialized and simple while the system as a whole provides comprehensive information.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The voice-controlled device acts as an intermediary that receives voice commands, processes them, and then directs appropriate content to either its own audio output or to a display device. This mediator approach enables seamless transitions between audio and visual output without requiring the user to manually switch devices.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Adaptability or versatility

If the voice-controlled device includes display capabilities, then visual content can be provided directly, but the device complexity increases

Engineering Contradiction:
Improveoutput modality flexibilityVSAvoiddevice structure
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The voice-controlled device is designed to perform multiple functions: it can output content through its own audio speaker, transfer content to a display device for visual presentation, and handle voice command processing. This multi-functionality allows the system to adapt to different user needs without adding physical display hardware to the voice-controlled device itself.

Inventive Principle:
Principle #6Universality (Multi-functionality)

3Ease of operation

If the system transitions content between devices based on user commands, then user experience is enhanced, but the system complexity increases

Engineering Contradiction:
Improvecontent access convenienceVSAvoidsystem architecture
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The system automatically detects user intent through voice commands and autonomously determines the appropriate output device and content format. The voice-controlled device self-manages the content transition process, selecting whether to output audio locally or transfer visual content to a display device without requiring manual device switching by the user.

Inventive Principle:
Principle #25Self-service

Data Source

PatentUS20240360512A1Providing Content on Multiple Devices
Publication Date: 2024.10.31 AMAZON TECH INC
  • US20240360512A1 patent drawing
  • US20240360512A1 patent drawing
  • US20240360512A1 patent drawing

AI summary

Techniques for receiving a voice command from a user and, in response, providing audible content to the user via a first device and providing visual content for the user via a second device. In some instances, the first device includes a microphone for generating audio signals that include user speech, as well as a speaker for outputting audible content in response to identified voice commands from the speech. However, the first device might not include a display for displaying graphical content. As such, the first device may be configured to identify devices that include displays and that are proximate to the first device. The first device may then instruct one or more of these other devices to output visual content associated with a user's voice command.