Cross-Device Media Playback via Voice-Based Target Device Selection
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing systems lack the ability to play media content on a target device based on a voice command received at a different device, limiting user flexibility in controlling media playback across devices.
Innovation Solution
A system and method that enables media content playback on a target device by processing voice commands at a separate device, utilizing a speech proxy, automatic speech recognition, natural language understanding, and connect services to identify and control the target device for media playback.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If voice commands are processed at the device where media playback is requested, then the control process is simple, but the user cannot play media on different devices
Solution Approach 1:
A cloud-based server acts as an intermediary between the user's mobile device and the target playback device. The server receives voice commands, processes them through speech recognition and natural language understanding, identifies the target device, and sends control instructions to play media content on the specified device, enabling cross-device media playback without direct peer-to-peer complexity
2Ease of operation
If manual device selection is required for media playback, then device control is precise, but user convenience and operation speed are reduced
Solution Approach 1:
The system automatically identifies the target playback device by parsing the voice command for device location information (e.g., room name, device name) and autonomously selects the appropriate device without requiring manual user input or selection, thereby enhancing convenience and reducing time loss
Solution Approach 2:
The system performs preliminary actions by pre-configuring device profiles and locations in advance, allowing the voice command processing to directly match and identify the target device without requiring real-time manual selection or configuration by the user
Data Source
Figure 1
Figure 2
Figure 3
AI summary
A first device receives a voice command from a first user of a second device. The first device determines, from content in the voice command, one or more characteristics of a target device and media content to be played on the target device. The first device identifies, using the characteristics of the target device, a third device. In response to identifying the third device: the first device modifies account information for the third device to associate the third device with the first user and transmits instructions to the third device to play the media content.