Voice-to-Text Feedback for Resuming Interrupted Utterances

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Users face difficulties in resuming voice-to-text conversion after interruptions, as they often forget the content of their utterances and struggle to determine where to continue, leading to the need to start anew, which is cumbersome.

Innovation Solution

The voice information processing apparatus automatically outputs the previously uttered content as voice when an interruption is detected, allowing users to listen and resume their sentences without canceling the existing text conversion process.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If the user cancels the sentence and starts the utterance from the beginning after an interruption, then the text conversion can be restarted, but the operation becomes complicated and time-consuming

Engineering Contradiction:
Improvetext conversion accuracyVSAvoidoperation complexity
Core Design Contradiction:
ReliabilityVSEase of operation

Solution Approach 1:

The system performs preliminary action by automatically outputting the already-converted text as voice when an interruption is detected. This allows the user to hear what has been converted so far and resume from the correct position, eliminating the need to cancel and restart the entire process.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system provides feedback by converting the already-processed text back to voice and outputting it to the user. This feedback mechanism helps the user understand the current conversion state and determine where to resume uttering, reducing operational complexity.

Inventive Principle:
Principle #23Feedback

2Reliability

If the user cancels and restarts the utterance after an interruption, then complete text conversion can be achieved, but the time required increases

Engineering Contradiction:
Improvesentence completionVSAvoidtime for text conversion
Core Design Contradiction:
ReliabilityVSLoss of time

Solution Approach 1:

The system preserves the preliminary conversion work by automatically outputting the already-converted text as voice. This allows the user to resume from where they left off rather than restarting, significantly reducing the time required to complete the sentence conversion.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system maintains continuity of useful action by preserving the conversion state and enabling the user to resume uttering from the interruption point. The already-converted text remains available and is output as voice to guide the user, ensuring the conversion process continues without unnecessary repetition.

Inventive Principle:
Principle #20Continuity of useful action

3Productivity

If the user does not cancel after an interruption, then the conversion process continues, but the user cannot accurately recall where to continue the sentence

Engineering Contradiction:
Improveconversion efficiencyVSAvoidutterance content recall
Core Design Contradiction:
ProductivityVSLoss of information

Solution Approach 1:

The system provides feedback by automatically outputting the already-converted text as voice. This auditory feedback helps the user recall what has been converted and where to continue, preventing information loss about the utterance content while maintaining conversion efficiency.

Inventive Principle:
Principle #23Feedback

Solution Approach 2:

The system uses a form of state indication by providing auditory feedback (voice output) about the current conversion state. This is analogous to visual indicators that show system state, helping the user understand where they are in the conversion process without having to mentally track it.

Inventive Principle:
Principle #32Color changes

Data Source

PatentUS12142272B2Voice information processing apparatus and voice information processing method
Publication Date: 2024.11.12 ALPS ALPINE CO LTD
  • US12142272B2 patent drawing
  • US12142272B2 patent drawing
  • US12142272B2 patent drawing

AI summary

A voice information processing apparatus sequentially converts an utterance of a user into text during a voice reception period that is a period in which an uttered voice to be converted into text is received from a user, and in a case where it can be regarded that the utterance of the user has been interrupted, the voice information processing apparatus automatically causes utterance content already uttered by the user to be output by a voice during the voice reception period. As a result, the voice information processing apparatus can cause the user to recognize a content of a sentence that has been uttered by the user so far and converted into text, when it can be regarded that the utterance of the user has been interrupted.