Speech Assistance Apparatus for Accessible VoIP Communication
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Users with speech issues, cognitive difficulties, or language barriers face challenges in communicating effectively during multi-player video games or other network-based interactions, leading to social isolation and dissatisfaction.
Innovation Solution
A speech assistance apparatus and method that transmits predefined or user-generated phrases on behalf of the user by recognizing easier words or phrases spoken and associating them with more difficult ones, allowing for seamless communication through a network using a system comprising a storage unit, recognition unit, evaluation unit, and transmission unit.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If users communicate using Voice over Internet Protocol (VoIP) during gameplay, then real-time communication is enabled, but users with speech issues, cognitive issues, or language barriers find communication inaccessible and feel socially isolated
Solution Approach 1:
The patent introduces an intermediary system that converts spoken words into text messages automatically. This mediator handles the communication process for users who struggle with speech, allowing them to communicate effectively without directly speaking. The system translates the user's intended message into appropriate text communications that are then transmitted to other players, resolving the accessibility issue while maintaining real-time communication capability.
Solution Approach 2:
The patent replaces the mechanical speech-based communication system with an automated text generation system. Instead of requiring users to physically speak and articulate words, the system uses speech recognition and natural language processing to automatically generate and transmit text messages. This substitution eliminates the barrier for users with speech difficulties while preserving the real-time communication function.
2Loss of information
If users are required to speak fluently to communicate effectively, then clear communication is achieved, but users with speech issues or cognitive difficulties are excluded from effective communication
Solution Approach 1:
The patent implements a self-service communication system where the automated text generation handles the complex task of formulating clear messages. Users simply need to indicate their intent, and the system automatically generates grammatically correct, contextually appropriate text communications. This self-service approach eliminates the need for users to manually craft clear speech, automatically resolving the issue of communication clarity while reducing the operational difficulty.
3Extent of automation
If speech recognition is used to identify user intent, then automated text generation is enabled, but the system complexity increases with multiple processing units
Solution Approach 1:
The patent integrates multiple functions into a unified communication assistance system. The speech recognition unit, evaluation unit, and text generation unit work together as a coordinated system that handles the entire communication process. By designing the system with multi-functionality, the patent reduces overall complexity compared to having separate independent systems for each function, as the components are optimized to work together efficiently within a single integrated architecture.
Data Source
AI summary
An apparatus, for assisting at least a first user in communicating with one or more other users via a network, includes: a storage unit configured to store: phrase data corresponding to one or more phrases, where each phrase comprises one or more words, tag data corresponding to one or more tags, where each tag comprises at least part of one word, and first association data corresponding to one or more associations between one or more of the phrases and one or more of the tags; an input unit configured to receive one or more audio signals from the at least first user; a recognition unit configured to recognise one or more spoken words included within the received audio signals; an evaluation unit configured to evaluate whether a given recognised spoken word corresponds to a given tag; and if so, a transmission unit configured to transmit one or more of the phrases associated with the given tag to one or more of the other users.

