Speech Recognition Correction Estimation for Accuracy

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing speech recognition technologies face challenges in improving accuracy when appropriate learning has not been performed, leading to reduced convenience in speech recognition services.

Innovation Solution

An information processing apparatus and method that includes a speech recognition unit, a correction portion estimation unit, and a presenting unit to collate speech recognition results with necessary collation information, estimating and presenting correction portions to users, enhancing accuracy and user interaction.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If word replacement is performed based on past learning results, then speech recognition accuracy can be improved when appropriate learning has been performed, but the system cannot contribute to improving accuracy when appropriate learning has not been performed

Engineering Contradiction:
Improvespeech recognition accuracyVSAvoidservice convenience
Core Design Contradiction:
Measurement precisionVSAdaptability or versatility

Solution Approach 1:

The system presents estimated correction portions to users, who can then confirm or correct them. This feedback loop allows the system to learn from user corrections and improve future recognition accuracy, making the service adaptable even without extensive pre-learning

Inventive Principle:
Principle #23Feedback

Solution Approach 2:

The system performs preliminary estimation of correction portions before final speech recognition completion. By proactively identifying and presenting potential errors to users, the system prepares corrections in advance, improving overall accuracy without requiring extensive pre-learning

Inventive Principle:
Principle #10Preliminary action

2Measurement precision

If speech recognition results are presented without correction estimation, then the system is simpler to operate, but the accuracy of the speech recognition result cannot be improved

Engineering Contradiction:
Improvespeech recognition accuracyVSAvoidservice convenience
Core Design Contradiction:
Measurement precisionVSEase of operation

Solution Approach 1:

The system automatically estimates and presents correction portions without requiring users to manually review every recognition result. This self-service approach improves accuracy while maintaining ease of operation, as users only need to interact when corrections are presented

Inventive Principle:
Principle #25Self-service

Data Source

PatentUS11107469B2Information processing apparatus and information processing method
Publication Date: 2021.08.31 SONY GROUP CORP
  • US11107469B2 patent drawing
  • US11107469B2 patent drawing
  • US11107469B2 patent drawing

AI summary

The present disclosure relates to an information processing apparatus and an information processing method for enabling provision of a more convenient speech recognition service. The information processing apparatus includes a speech recognition unit that performs speech recognition for speech information based on an utterance of a user, and a correction portion estimation unit that collates content of a sentence obtained as a speech recognition result with collation information necessary for determining accuracy of the content to estimate, for the sentence, a correction portion that requires correction. The sentence obtained as a speech recognition result is displayed together with the correction portion estimated by the correction portion estimation unit and presented to the user.