Speech Recognition Engine Manager for Adaptive User Interface Selection

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Complex graphical user interfaces (GUIs) are often mouse- and keyboard-intensive, making them difficult or impossible for individuals with physical disabilities to use, while existing speech recognition systems lack the ability to dynamically select the most suitable engine based on user preferences and environmental factors.

Innovation Solution

A speech recognition engine manager that detects available engines, selects the most suitable one based on user preferences, heuristics, and environmental factors, and manages sessions to ensure optimal speech recognition, allowing for seamless integration with various user interfaces and applications.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If a single speech recognition engine is used, then the system is simple to manage, but it cannot adapt to different user preferences and environmental factors

Engineering Contradiction:
Improveadaptability to user preferences and environmental factorsVSAvoidcomplexity of managing multiple speech recognition engines
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent combines multiple speech recognition engines into a unified system managed by a speech recognition engine manager. This manager detects available engines, selects the most suitable one based on user preferences and environmental factors, and coordinates their operation, thereby achieving adaptability while maintaining manageable system complexity through centralized control.

Inventive Principle:
Principle #5Merging (Combining)

Solution Approach 2:

The speech recognition engine manager serves multiple functions: it detects available engines, selects appropriate engines based on various criteria, manages sessions, and coordinates recognition operations. This multi-functional approach allows the system to adapt to different scenarios without requiring separate management systems for each function.

Inventive Principle:
Principle #6Universality (Multi-functionality)

2Adaptability or versatility

If multiple speech recognition engines are managed manually, then flexibility is improved, but ease of operation deteriorates

Engineering Contradiction:
Improveflexibility in selecting speech recognition enginesVSAvoidease of managing speech recognition engines
Core Design Contradiction:
Adaptability or versatilityVSEase of operation

Solution Approach 1:

The speech recognition engine manager automatically detects available engines, selects the most suitable one based on user preferences and environmental factors, and manages sessions without requiring manual intervention. This self-service capability maintains flexibility in engine selection while significantly improving ease of operation by eliminating manual management tasks.

Inventive Principle:
Principle #25Self-service

3Ease of operation

If speech recognition is implemented, then accessibility for individuals with physical disabilities is improved, but system complexity increases

Engineering Contradiction:
Improveaccessibility for individuals with physical disabilitiesVSAvoidcomplexity of integrating speech recognition systems
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The speech recognition engine manager acts as an intermediary between the user interface and multiple speech recognition engines. It handles the complexity of engine detection, selection, and coordination, while presenting a simple interface to users. This mediator approach improves accessibility by enabling speech recognition without requiring users to manage the underlying system complexity.

Inventive Principle:
Principle #24Intermediary (Mediator)

Data Source

PatentUS7340395B2Multiple speech recognition engines
Publication Date: 2008.03.04 SAP SE
  • US7340395B2 patent drawing
  • US7340395B2 patent drawing
  • US7340395B2 patent drawing

AI summary

A system having multiple speech recognition engines, each operable to recognize spoken data, is described. A speech recognition engine manager detects the speech recognition engines, and selects at least one for recognizing spoken input from a user, via a user interface. In this way, a speech recognition engine that is particularly suited to a current environment may be selected. For example, a speech recognition engine that is particularly suited for, or preferred by, the user may be selected, or a speech recognition engine that is particularly suited for a particular type of interface, interface element, or application, may be selected. Multiple ones of the speech recognition engines may be selected and simultaneously maintained in an active state, by maintaining a session associated with each of the engines. Accordingly, users' experience with voice applications may be enhanced, and, in particular, users with physical disabilities may more easily interact with software applications.