Multi-modal Input Processing for Electronic Devices

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Voice recognition services struggle to process user inputs with insufficient information, failing to recognize intent and requiring additional utterances, and lack integration with other input methods like keyboards and touch inputs, making them difficult to use.

Innovation Solution

An electronic device and method that integrates voice input processing with other input means, such as keyboards and touch inputs, by generating text data from user utterances, selecting appropriate applications, and performing operations based on this data, while allowing additional inputs to complete tasks.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If voice recognition service processes only simple user inputs, then it can provide results based on recognized voice, but it cannot process user inputs that require multiple applications or additional information

Engineering Contradiction:
Improvecapability to process complex user inputsVSAvoidrecognition accuracy
Core Design Contradiction:
Adaptability or versatilityVSReliability

Solution Approach 1:

The patent combines voice input processing with other input methods (keyboard, touch screen, mouse) into an integrated interface. The system merges multiple input modalities to handle complex user inputs that require information from both voice and other input devices, thereby improving adaptability while maintaining reliability through complementary input channels.

Inventive Principle:
Principle #5Merging (Combining)

Solution Approach 2:

The voice recognition service is enhanced to handle multiple types of user inputs beyond simple voice commands. The system is designed to process diverse input types (voice, text, touch) and coordinate them to execute complex tasks involving multiple applications, making the service universal and multi-functional.

Inventive Principle:
Principle #6Universality (Multi-functionality)

2Measurement precision

If voice recognition service requests additional utterances for insufficient information, then it can grasp user intent, but users perceive the service as difficult to use

Engineering Contradiction:
Improveintent recognition accuracyVSAvoiduser experience
Core Design Contradiction:
Measurement precisionVSEase of operation

Solution Approach 1:

The system merges voice input with other input methods to gather sufficient information for intent recognition without requiring multiple voice utterances. Users can supplement voice input with keyboard typing or touch screen interactions, improving ease of operation while maintaining precise intent recognition through combined input data.

Inventive Principle:
Principle #5Merging (Combining)

Solution Approach 2:

The patent introduces an integrated interface as an intermediary between the user and the voice recognition service. This interface coordinates multiple input methods and processes them together, mediating the interaction to reduce the burden on users while accurately grasping user intent through synthesized input data.

Inventive Principle:
Principle #24Intermediary (Mediator)

3Adaptability or versatility

If voice recognition service uses only voice input/output interface, then it can process natural language, but it cannot integrate with other input means like keyboard and mouse

Engineering Contradiction:
Improveintegration with multiple input methodsVSAvoidinterface complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The interface is designed to be universal, supporting multiple input methods (voice, keyboard, touch screen, mouse) within a single integrated system. This multi-functional interface can handle various input types without requiring separate processing systems, achieving integration while managing complexity through unified design.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The patent segments the input processing into modular components, each handling a specific input method (voice recognition module, keyboard input module, touch screen module). These segmented modules communicate through a standardized interface, enabling integration of multiple input methods while keeping individual module complexity manageable.

Inventive Principle:
Principle #1Segmentation

Data Source

PatentUS11561763B2Electronic device for processing multi-modal input, method for processing multi-modal input and server for processing multi-modal input
Publication Date: 2023.01.24 SAMSUNG ELECTRONICS CO LTD
  • US11561763B2 patent drawing
  • US11561763B2 patent drawing
  • US11561763B2 patent drawing

AI summary

An electronic device is provided. The electronic device includes a housing, a touchscreen display exposed through a first portion of the housing, a microphone disposed at a second portion of the housing, a speaker disposed at a third portion of the housing, a memory disposed inside the housing, a processor disposed inside the housing, and electrically connected to the display, the microphone, the speaker, and the memory. The memory is configured to store a plurality of application programs, each of which includes a graphic user interface (GUI).