Reusable Multimodal Application With Universal Voice and Touch Input

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing mobile devices face challenges in providing user-friendly multimodal interactions due to constrained form factors and the need for custom development for each application, limiting accessibility and revenue opportunities for content providers and carriers.

Innovation Solution

A reusable multimodal application on mobile devices that accepts multimodal inputs, synchronizes and processes them, and integrates with existing infrastructure to provide seamless multimodal experiences without requiring hardware or software replacements.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If custom development is performed for each application to provide multimodal functionality, then multimodal interaction capability is improved, but device complexity and development cost increase

Engineering Contradiction:
Improvemultimodal interaction capabilityVSAvoiddevelopment complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent implements a universal multimodal input framework that can be applied across multiple applications without custom development. The system provides a standardized interface layer that handles voice, touch, and other input modes uniformly, allowing any application to benefit from multimodal capabilities through a common architecture rather than individual customization for each app.

Inventive Principle:
Principle #6Universality (Multi-functionality)

2Productivity

If hardware or software replacement is performed to enable multimodal functionality, then interaction efficiency is improved, but device compatibility and infrastructure cost worsen

Engineering Contradiction:
Improveinteraction efficiencyVSAvoidinfrastructure compatibility
Core Design Contradiction:
ProductivityVSAdaptability or versatility

Solution Approach 1:

The patent introduces an intermediary framework layer between the existing application layer and the hardware layer. This mediator translates various hardware input modes (voice, touch, keypad) into standardized input commands that existing applications can process, enabling multimodal interaction without requiring changes to the underlying hardware or software infrastructure.

Inventive Principle:
Principle #24Intermediary (Mediator)

3Ease of operation

If multiple modes of communication are integrated, then user interaction efficiency is improved, but system complexity increases

Engineering Contradiction:
Improveuser interaction efficiencyVSAvoidsystem complexity
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The patent segments the multimodal input system into distinct modular components, each handling a specific input mode (voice recognition module, touch input module, keypad module). Each module operates independently and communicates through a standardized interface, reducing system complexity by avoiding the need for intricate integration logic while maintaining ease of operation through coordinated multi-mode input.

Inventive Principle:
Principle #1Segmentation

Data Source

PatentUS12418583B2Reusable multimodal application
Publication Date: 2025.09.16 GULA CONSULTING LLC
  • US12418583B2 patent drawing
  • US12418583B2 patent drawing
  • US12418583B2 patent drawing

AI summary

A method and system are disclosed herein for accepting multimodal inputs and deriving synchronized and processed information. A reusable multimodal application is provided on the mobile device. A user transmits a multimodal command to the multimodal platform via the mobile network. The one or more modes of communication that are inputted are transmitted to the multimodal platform(s) via the mobile network(s) and thereafter synchronized and processed at the multimodal platform. The synchronized and processed information is transmitted to the multimodal application. If required, the user verifies and appropriately modifies the synchronized and processed information. The verified and modified information are transferred from the multimodal application to the visual application. The final result(s) are derived by inputting the verified and modified results into the visual application.