Speech Recognition Correction Estimation for Accuracy
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing speech recognition technologies face challenges in improving accuracy when appropriate learning has not been performed, leading to reduced convenience in speech recognition services.
Innovation Solution
An information processing apparatus and method that includes a speech recognition unit, a correction portion estimation unit, and a presenting unit to collate speech recognition results with necessary collation information, estimating and presenting correction portions to users, enhancing accuracy and user interaction.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If word replacement is performed based on past learning results, then speech recognition accuracy can be improved when appropriate learning has been performed, but the system cannot contribute to improving accuracy when appropriate learning has not been performed
Solution Approach 1:
The system presents estimated correction portions to users, who can then confirm or correct them. This feedback loop allows the system to learn from user corrections and improve future recognition accuracy, making the service adaptable even without extensive pre-learning
Solution Approach 2:
The system performs preliminary estimation of correction portions before final speech recognition completion. By proactively identifying and presenting potential errors to users, the system prepares corrections in advance, improving overall accuracy without requiring extensive pre-learning
2Measurement precision
If speech recognition results are presented without correction estimation, then the system is simpler to operate, but the accuracy of the speech recognition result cannot be improved
Solution Approach 1:
The system automatically estimates and presents correction portions without requiring users to manually review every recognition result. This self-service approach improves accuracy while maintaining ease of operation, as users only need to interact when corrections are presented
Data Source
AI summary
The present disclosure relates to an information processing apparatus and an information processing method for enabling provision of a more convenient speech recognition service. The information processing apparatus includes a speech recognition unit that performs speech recognition for speech information based on an utterance of a user, and a correction portion estimation unit that collates content of a sentence obtained as a speech recognition result with collation information necessary for determining accuracy of the content to estimate, for the sentence, a correction portion that requires correction. The sentence obtained as a speech recognition result is displayed together with the correction portion estimated by the correction portion estimation unit and presented to the user.


