Gesture Recognition Input for Interactive Media Devices

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Interactive media devices, especially those without keyboards, face challenges in inputting data, particularly for East Asian languages, due to the complexity of characters and the time-consuming nature of using virtual keyboards.

Innovation Solution

A sensing device tracks the movement of an object in a trajectory indicative of the desired input, which is then translated by a tracking module and recognition software into recognizable inputs, such as characters or navigational commands, displayed for user selection, with suggested characters ranked by likelihood for easier input.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If a virtual keyboard is used for data input, then text input capability is provided, but input time is excessive and language support is limited

Engineering Contradiction:
Improvelanguage supportVSAvoidinput time
Core Design Contradiction:
Adaptability or versatilityVSLoss of time

Solution Approach 1:

The patent replaces the mechanical virtual keyboard input system with an optical gesture recognition system. A sensing device captures video data of hand gestures, and software translates these gestures into text input or navigational commands. This substitution enables support for East Asian languages through natural handwriting gestures while dramatically reducing input time compared to character-by-character virtual keyboard selection.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Ease of operation

If a virtual keyboard is used for data input, then text input capability is provided, but input time is excessive

Engineering Contradiction:
Improvetext input capabilityVSAvoidinput time
Core Design Contradiction:
Ease of operationVSLoss of time

Solution Approach 1:

The patent substitutes the tedious virtual keyboard navigation mechanism with a gesture-based optical recognition system. Users can input text through natural hand gestures captured by a sensing device, eliminating the need to navigate through virtual keyboard characters. This maintains text input capability while reducing input time from minutes to seconds.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Solution Approach 2:

The system performs preliminary recognition and translation of gestures into text or commands before final input is required. The software continuously monitors and interprets gesture trajectories, preparing suggested characters or commands in advance, which are then confirmed or selected by the user, streamlining the overall input process.

Inventive Principle:
Principle #10Preliminary action

3Productivity

If gesture recognition is implemented, then input time is reduced and language support is expanded, but device complexity increases

Engineering Contradiction:
Improveinput speedVSAvoidsystem complexity
Core Design Contradiction:
ProductivityVSDevice complexity

Solution Approach 1:

The patent implements a universal gesture recognition system that handles multiple functions through a single mechanism. The same sensing device and software framework support both text input (including East Asian languages) and navigational commands, eliminating the need for separate input systems and reducing overall device complexity despite the advanced functionality.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Data Source

PatentUS9207765B2Recognizing interactive media input
Publication Date: 2015.12.08 MICROSOFT TECHNOLOGY LICENSING LLC
  • US9207765B2 patent drawing
  • US9207765B2 patent drawing
  • US9207765B2 patent drawing

AI summary

Techniques and systems for inputting data to interactive media devices are disclosed herein. In some aspects, a sensing device senses an object as it moves in a trajectory indicative of a desired input to an interactive media device. Recognition software may be used to translate the trajectory into various suggested characters or navigational commands. The suggested characters may be ranked based on a likelihood of being an intended input. The suggested characters may be displayed on a user interface at least in part based on the rank and made available for selection as the intended input.