Voice Transcription Error Correction via Sound Similarity Matching

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional methods for inputting character strings, whether through keyboards or voice recognition, often result in typographical errors and false recognitions, requiring manual correction by users, which is time-consuming and inefficient.

Innovation Solution

A computer-based transcription method that accepts voice input, performs voice recognition, and corrects character strings by specifying portions with similarity to the original sound information, allowing users to easily identify and correct errors using a graphical user interface.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Productivity

If voice recognition is used for inputting character strings, then input speed is improved, but accuracy deteriorates due to false recognitions and typographical errors

Engineering Contradiction:
Improveinput speedVSAvoidaccuracy
Core Design Contradiction:
ProductivityVSMeasurement precision

Solution Approach 1:

The system performs voice recognition to generate character strings, then uses sound information matching to provide feedback by identifying portions that match the original voice input. This feedback mechanism allows automatic correction of recognition errors while preserving the speed advantage of voice input.

Inventive Principle:
Principle #23Feedback

Solution Approach 2:

The patent replaces manual mechanical correction (keyboard input) with an automatic sound-based correction system. By substituting the mechanical keyboard correction process with automated sound information comparison, the system maintains high input speed while improving accuracy.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Measurement precision

If manual correction is performed to improve accuracy, then accuracy is improved, but time consumption increases

Engineering Contradiction:
ImproveaccuracyVSAvoidtime consumption
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

The system performs self-correction by automatically comparing sound information and identifying matching portions without requiring manual user intervention. This self-service mechanism improves accuracy while minimizing the time loss associated with manual correction.

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The system performs preliminary sound information matching and identification of correctable portions before final correction is applied. This preliminary action prepares the correction data in advance, reducing the time needed for actual correction operations.

Inventive Principle:
Principle #10Preliminary action

3Ease of operation

If voice recognition is used, then ease of operation is improved, but reliability deteriorates due to typographical errors and false recognitions

Engineering Contradiction:
Improveease of inputVSAvoidreliability
Core Design Contradiction:
Ease of operationVSReliability

Solution Approach 1:

The system provides automatic feedback by comparing sound information and identifying portions that match the original voice input. This feedback loop enhances reliability by automatically detecting and correcting errors while maintaining the ease of voice-based operation.

Inventive Principle:
Principle #23Feedback

4Productivity

If automatic correction is implemented, then productivity is improved, but device complexity increases

Engineering Contradiction:
Improvetranscription efficiencyVSAvoidsystem complexity
Core Design Contradiction:
ProductivityVSDevice complexity

Solution Approach 1:

The patent replaces complex manual correction operations with an automated sound information matching system. By substituting manual processes with automated acoustic comparison, the system improves productivity while managing complexity through algorithmic rather than mechanical means.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Data Source

PatentUS11798558B2Recording medium recording program, information processing apparatus, and information processing method for transcription
Publication Date: 2023.10.24 FUJITSU LTD
  • US11798558B2 patent drawing
  • US11798558B2 patent drawing
  • US11798558B2 patent drawing

AI summary

A method for transcription is performed by a computer. The method includes: accepting input of a voice after causing a display unit to display a sentence including a plurality of words; acquiring first sound information being information concerning sounds corresponding to the sentence; acquiring second sound information being information concerning sounds of the voice accepted in the accepting; specifying a portion in the first sound information having a prescribed similarity to the second sound information; and correcting a character string in the sentence corresponding to the specified portion based on a character string corresponding to the second sound information.