Visual Cue Input for Banking Accessibility

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing mobile and online banking systems face challenges in accessibility for users who have difficulty entering input commands, particularly those who cannot type or prefer not to use audio speech recognition, especially in sensitive or noisy environments, and lack convenience for users wanting to streamline mobile application processes.

Innovation Solution

A method and system that utilize a computing device's camera to capture and parse visual cues such as lip movements and gestures to map user inputs to available operations on a mobile or online banking application, allowing users to execute commands without traditional input methods, with verification and customization options.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If traditional input methods (typing, audio speech recognition) are used in banking applications, then command input capability is provided, but accessibility for users with input difficulties is poor and convenience in sensitive/noisy environments is reduced

Engineering Contradiction:
Improveaccessibility for users with input difficultiesVSAvoidsuitability for sensitive/noisy environments
Core Design Contradiction:
Ease of operationVSAdaptability or versatility

Solution Approach 1:

The patent replaces traditional mechanical input methods (typing on keyboard, audio speech recognition) with a visual-based input system that captures and interprets user gestures and facial expressions through camera imaging, eliminating the need for physical or vocal input while maintaining command input capability

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Solution Approach 2:

The system introduces an intermediary visual processing layer between the user and the banking application, where camera-captured visual cues are translated into command inputs, providing an alternative communication path that works in environments where traditional methods are inconvenient or inaccessible

Inventive Principle:
Principle #24Intermediary (Mediator)

2Ease of operation

If visual cue recognition system is implemented, then accessibility and convenience are enhanced, but system complexity increases

Engineering Contradiction:
Improveusability for users with input difficultiesVSAvoidsystem architecture complexity
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The patent leverages the existing camera component of mobile devices, which is already a universal feature in modern smartphones and tablets, to perform multiple functions including visual cue capture, gesture recognition, and command translation, thereby adding functionality without requiring additional specialized hardware

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The system utilizes the device's own existing camera and processing capabilities to perform visual recognition and command translation, allowing the device to serve itself for input recognition without requiring external specialized equipment or complex additional system components

Inventive Principle:
Principle #25Self-service

3Adaptability or versatility

If visual cues are mapped to application operations, then input method versatility is improved, but mapping accuracy and reliability may be compromised

Engineering Contradiction:
Improveinput method optionsVSAvoidcommand recognition accuracy
Core Design Contradiction:
Adaptability or versatilityVSReliability

Solution Approach 1:

The system performs preliminary mapping between visual cues and application operations during system setup or usage, creating a predefined association database that enables accurate and reliable translation of captured visual gestures into specific banking application commands, reducing ambiguity in real-time recognition

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS11755118B2Input commands via visual cues
Publication Date: 2023.09.12 CAPITAL ONE SERVICES LLC
  • US11755118B2 patent drawing
  • US11755118B2 patent drawing
  • US11755118B2 patent drawing

AI summary

Embodiments disclosed herein generally relate to a method and system of generating text input via facial recognition. A computing system receives a video stream of a user operating an application on a client device. The video stream includes a time series of images of the user. The computing system parses the video stream to identify one or more visual cues of the user. The computing system identifies a current page of the application. The computing system maps the identified on or more visual cues to an operation available on the current page of the application. The computing system executes the mapped operation.