Voice Interface Multi-Media Output Coordination

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing voice-based interface systems, such as AI speakers, have limited information delivery capabilities due to a simple and consistent output scheme, which restricts the effectiveness of response information provided to voice requests.

Innovation Solution

The system manages multiple media sources, including AI speakers, smartphones, IPTVs, and smart devices, to provide both auditory and additional outputs, such as visual or tactile responses, based on the type and efficiency of information delivery, allowing for adaptive and expanded information output methods.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Loss of information

If only a simple answering voice is output through a single medium, then the device complexity is reduced and ease of operation is improved, but the information delivery capability is limited

Engineering Contradiction:
Improveinformation delivery capabilityVSAvoidsystem complexity
Core Design Contradiction:
Loss of informationVSDevice complexity

Solution Approach 1:

The system segments information delivery across multiple media types (voice output through speakers, visual output through displays, text output through interfaces). Each medium handles specific types of response information, allowing comprehensive information delivery while maintaining simple control through a unified voice-based interface.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The voice-based interface serves multiple functions: it can request information, control devices, and receive responses through various media. The system universally handles different types of outputs (auditory, visual, textual) through a single integrated platform, improving information delivery without requiring multiple specialized interfaces.

Inventive Principle:
Principle #6Universality (Multi-functionality)

2Loss of information

If multiple media sources are managed for varied outputs, then information delivery capability is enhanced, but device complexity increases

Engineering Contradiction:
Improveinformation delivery capabilityVSAvoidmedia management complexity
Core Design Contradiction:
Loss of informationVSDevice complexity

Solution Approach 1:

The system merges multiple media outputs (voice, visual, text) under a single voice-based interface control. Different electronic devices (smartphones, IPTVs, smart refrigerators) are integrated into one unified system that manages and coordinates their outputs, enhancing information delivery while centralizing control to minimize management complexity.

Inventive Principle:
Principle #5Merging (Combining)

Solution Approach 2:

The voice-based interface acts as an intermediary between the user and multiple electronic devices. It translates user voice requests into appropriate outputs across different media, and conversely, consolidates responses from various devices into unified voice answers, simplifying the interaction complexity despite managing multiple media sources.

Inventive Principle:
Principle #24Intermediary (Mediator)

3Ease of operation

If synchronized outputs across multiple media are provided, then user interaction is improved, but the loss of time for coordination increases

Engineering Contradiction:
Improveuser interaction qualityVSAvoidcoordination time
Core Design Contradiction:
Ease of operationVSLoss of time

Solution Approach 1:

The system performs preliminary actions by pre-establishing communication channels and coordination protocols between the voice-based interface and multiple electronic devices. Media information is prepared and synchronized in advance, allowing rapid coordinated output when a voice request is made, thus improving user interaction without significant time loss.

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS11341966B2Output for improving information delivery corresponding to voice request
Publication Date: 2022.05.24 NAVER CORP
  • US11341966B2 patent drawing
  • US11341966B2 patent drawing
  • US11341966B2 patent drawing

AI summary

An output technology for improving information delivery corresponding to a voice request is provided. In one embodiment, a method by which an electronic device comprising a voice-based interface provides information comprises the steps of: receiving a voice request from a user through the voice-based interface; acquiring response information corresponding to the voice request; outputting the response information in a reply voice, which is an auditory output form, through at least one medium among a plurality of media including a main medium corresponding to the voice-based interface and a sub medium included in other electronic devices linkable with the electronic device; and providing other outputs for at least a part of the response information through at least one medium, which is the same as or different from the medium through which the reply voice is being outputted, among the plurality of media.