Wearable Audio-to-Text Visualizer for Communication Support

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Individuals with conditions such as autism and Asperger's syndrome face challenges in maintaining socially appropriate eye contact, managing tantrums, and communicating effectively, due to difficulties with speech disorders and prosody, which existing technologies have not adequately addressed.

Innovation Solution

A wearable system that captures and processes audio data to provide real-time feedback and reports on conversation improvements, tantrum prediction, and speech analysis, using sensors to detect and interpret nonverbal cues and audio patterns, and visually present auditory information to assist individuals and caregivers.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If audio data is processed and visually presented to assist individuals with communication disorders, then communication skills and social interaction are improved, but device complexity and processing requirements increase

Engineering Contradiction:
Improvecommunication skillsVSAvoidsystem complexity
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The patent introduces an intermediary system that converts audio data into visual representations. The system uses sensors to capture audio, processes the audio signals, and displays visual feedback to individuals with communication disorders. This intermediary approach translates complex auditory information into accessible visual formats, improving communication without requiring the individuals to process complex audio signals directly.

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The patent replaces direct auditory processing with visual presentation mechanisms. Instead of relying on the individual's auditory processing capabilities, the system substitutes audio information with visual displays that are easier to process and interpret. This substitution reduces the cognitive load on individuals with communication disorders while maintaining the essential communication function.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Productivity

If real-time audio processing and feedback are provided, then conversation improvements and tantrum management are enhanced, but energy consumption and processing time increase

Engineering Contradiction:
Improveconversation improvement rateVSAvoidenergy consumption
Core Design Contradiction:
ProductivityVSUse of energy by moving object

Solution Approach 1:

The system provides continuous audio capture and processing throughout the interaction, maintaining constant monitoring of speech patterns and conversational flow. This continuous operation enables real-time detection of communication patterns and tantrum indicators, allowing for immediate feedback and intervention without interrupting the natural flow of interaction.

Inventive Principle:
Principle #20Continuity of useful action

Solution Approach 2:

The patent implements feedback mechanisms that provide real-time information about audio patterns and conversational dynamics. The system analyzes audio data and delivers feedback to both the individual and caregivers, enabling continuous adjustment and improvement of communication strategies. This feedback loop accelerates learning and skill development while managing energy consumption through targeted processing.

Inventive Principle:
Principle #23Feedback

3Measurement precision

If multiple sensors and processing components are integrated, then measurement precision of audio patterns and nonverbal cues is improved, but device complexity increases

Engineering Contradiction:
Improveaudio pattern detection accuracyVSAvoidsensor integration complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The patent divides the audio processing system into separate functional modules, each responsible for specific tasks such as audio capture, signal processing, pattern recognition, and feedback generation. This segmentation allows for specialized optimization of each component while maintaining overall system manageability. The modular architecture reduces complexity by enabling independent development and maintenance of each processing stage.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The system employs multi-functional sensors and processing components that serve multiple purposes. For example, audio sensors not only capture speech but also detect nonverbal cues and emotional states. The processing system handles various functions including pattern recognition, tantrum detection, and communication analysis simultaneously. This multi-functionality reduces the total number of components needed while maintaining high measurement precision.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Data Source

PatentUS11837249B2Visually presenting auditory information
Publication Date: 2023.12.05 RPX CORP
  • US11837249B2 patent drawing
  • US11837249B2 patent drawing
  • US11837249B2 patent drawing

AI summary

Systems, methods and non-transitory computer readable media for processing audio and visually presenting information are provided. Audio data captured by one or more audio sensors included in a wearable apparatus from an environment of a wearer of the wearable apparatus may be obtained. The audio data may be analyzed to obtain textual information. The audio data may be analyzed to associate different portions of the textual information with different speakers. A head mounted display system may be used to present each portion of the textual information in a presentation region associated with the speaker associated with the portion of the textual information.