Voice-Based Custom Device Action Execution System

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current digital virtual assistants (DVAs) face limitations in expanding voice-based interactions across a wide range of second-party devices and third-party applications due to technical barriers, such as the need for OS modifications, cloud presence maintenance, and limited vocabulary scope, which restricts customization and flexibility.

Innovation Solution

A data processing system that enables voice-based interactions by receiving device action data and identifiers from client devices, mapping these to device actions and commands, and transmitting executable commands to perform actions, allowing second-party device providers and third-party application providers to register custom device actions and application actions, enabling on-device execution without cloud dependency.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If current DVA systems use traditional voice interaction methods, then basic voice commands can be processed, but the system cannot expand interactions across diverse second-party devices and third-party applications due to OS modification requirements and cloud dependency

Engineering Contradiction:
Improveexpandability across devices and applicationsVSAvoidOS modification requirements
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent extracts the voice interaction processing functionality from cloud-based systems and relocates it to on-device execution. The DVA processor is integrated directly into client devices, eliminating the need for cloud presence maintenance and reducing dependency on complex OS modifications. This allows the system to process voice commands locally while maintaining the ability to interact with diverse devices and applications.

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The patent introduces a device action customization component that acts as an intermediary between the voice query and device execution. This component maps voice-based queries to device actions and executable commands, enabling seamless interactions with second-party devices and third-party applications without requiring direct OS modifications. The intermediary layer standardizes the interaction interface while preserving device-specific capabilities.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Extent of automation

If the DVA system maintains cloud presence for voice processing, then centralized processing can be achieved, but the system requires continuous cloud maintenance and cannot execute actions directly on devices

Engineering Contradiction:
Improvecloud-based processing capabilityVSAvoidcloud dependency
Core Design Contradiction:
Extent of automationVSReliability

Solution Approach 1:

The patent implements self-service by enabling client devices to process voice queries and execute device actions autonomously without continuous cloud dependency. The DVA processor and device action customization component are integrated into the client device, allowing it to independently receive voice queries, process them through the natural language processor, and execute corresponding actions on connected devices. This eliminates the need for constant cloud presence while maintaining automated processing capabilities.

Inventive Principle:
Principle #25Self-service

3Ease of operation

If the system uses a limited vocabulary scope for voice commands, then the system remains simple to implement, but customization and flexibility for different devices are restricted

Engineering Contradiction:
Improvevoice command simplicityVSAvoidcustomization capability
Core Design Contradiction:
Ease of operationVSAdaptability or versatility

Solution Approach 1:

The patent implements dynamics by making the vocabulary scope adaptable rather than fixed. The device action customization component dynamically determines the applicable vocabulary based on the specific device or application context. When a voice query is received, the system identifies the target device or application and retrieves the corresponding customized action set. This allows the same voice processing infrastructure to support diverse devices with different capabilities and terminologies without requiring separate rigid vocabularies for each.

Inventive Principle:
Principle #15Dynamics

Data Source

PatentUS20240282306A1Systems and methods for voice-based initiation of custom device actions
Publication Date: 2024.08.22 GOOGLE LLC
  • US20240282306A1 patent drawing
  • US20240282306A1 patent drawing
  • US20240282306A1 patent drawing

AI summary

Systems and methods for enabling voice-based interactions with electronic devices can include a data processing system maintaining a plurality of device action data sets and a respective identifier for each device action data set. The data processing system can receive, from an electronic device, an audio signal representing a voice query and an identifier. The data processing system can identify, using the identifier, a device action data set. The data processing system can identify a device action from device action data set based on content of the audio signal. The data processing system can then identify, from the device action dataset, a command associated with the device action and send the command to the for execution device for execution.