Captioning System with Dual Input for Hearing-Impaired Users

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

During captioning communication sessions, miscommunication can occur due to the hearing user speaking unclearly or in a non-native language, making it difficult for the hearing-impaired user and human assistants to understand and transcribe spoken words accurately.

Innovation Solution

A method that presents captions of spoken words on a display device for a hearing-impaired user during a call, allowing the hearing user to either speak or type words that are then displayed, enabling clear communication by avoiding difficulties in spoken language understanding.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Loss of information

If a human assistant transcribes spoken words during a captioning communication session, then the hearing-impaired user can read captions of the words spoken by the hearing user, but miscommunication can occur due to the hearing user speaking unclearly or in a non-native language

Engineering Contradiction:
Improvetranscription accuracyVSAvoidcommunication accuracy
Core Design Contradiction:
Loss of informationVSReliability

Solution Approach 1:

The patent introduces a text-to-speech conversion system as an intermediary between the hearing user and the captioning system. The hearing user's spoken words are converted to text, then back to speech through a different voice, providing a reliable transcription that bypasses the issues of unclear or non-native speech. This intermediary process ensures accurate captioning without relying on the original speaker's clarity or language proficiency.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Reliability

If the hearing user types words instead of speaking them, then communication accuracy is improved, but the complexity of the communication process increases

Engineering Contradiction:
Improvecommunication accuracyVSAvoidcommunication process complexity
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

The patent enables the hearing user to type messages directly, allowing them to serve their own communication needs without requiring a human assistant for transcription. This self-service approach gives the hearing user direct control over the communication process, ensuring accurate transmission of their intended message while reducing the complexity of requiring multiple intermediaries.

Inventive Principle:
Principle #25Self-service

3Ease of operation

If real-time captioning is provided during audio communications, then hearing-impaired users can participate effectively, but the system complexity increases due to the need for human assistants and transcription processes

Engineering Contradiction:
Improveaccessibility for hearing-impaired usersVSAvoidcaptioning system complexity
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The patent replaces the mechanical system of human assistants performing manual transcription with an automated text-to-speech conversion system. The hearing user's spoken words are automatically converted to text and then to speech through a synthesized voice, eliminating the need for human transcriptionists. This automation maintains real-time captioning capability while significantly reducing system complexity and operational overhead.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Data Source

PatentUS12132855B2Presentation of communications
Publication Date: 2024.10.29 SORENSON IP HOLDINGS LLC
  • US12132855B2 patent drawing
  • US12132855B2 patent drawing
  • US12132855B2 patent drawing

AI summary

A method to present communications may include captioning, by a human assistant during a call between a first user using a first captioning telephone device and a second user using a second telephone device, words spoken by the second user into the second telephone device. The method may also include presenting the captioned words on a first display of the first captioning telephone device, receiving text typed into the second telephone device by the second user, and presenting the received text on the first display of the first captioning telephone device.