Speech Intent Recognition for Foreign Language Learners

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current voice recognition devices for foreign language conversation systems have low recognition performance for non-native speakers, limiting their effectiveness and making it difficult for learners to engage in free speech, resulting in unsatisfactory learning outcomes.

Innovation Solution

An apparatus and method that includes a voice recognition device, a speech intent recognition device using skill level and dialogue context-based models, and a feedback processing device to provide customized responses, recommended, right, or alternative expressions, enabling accurate intent determination and adaptive feedback for learners.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If existing voice recognition devices are used for non-native speakers, then the device can recognize speech, but the recognition performance is very low

Engineering Contradiction:
Improvespeech recognition accuracyVSAvoidadaptability to different skill levels
Core Design Contradiction:
Measurement precisionVSAdaptability or versatility

Solution Approach 1:

The patent applies local quality by creating different speech-based models tailored to specific skill levels (beginner, intermediate, advanced) rather than using a single universal model. Each model is optimized for the characteristics and error patterns of speakers at that particular skill level, thereby improving recognition accuracy for non-native speakers without requiring a completely separate system for each proficiency level.

Inventive Principle:
Principle #3Local quality

Solution Approach 2:

The system dynamically changes parameters by selecting different speech-based models based on the user's skill level information. When a user's skill level changes or when the system detects variations in speech quality, it adjusts the model parameters accordingly, allowing the recognition system to adapt to different proficiency levels and maintain high accuracy across diverse user groups.

Inventive Principle:
Principle #35Parameter changes

2Ease of operation

If voice recognition is limited to native speakers, then recognition accuracy is high, but learners cannot freely enter speech

Engineering Contradiction:
Improvefreedom of speech inputVSAvoidrecognition performance
Core Design Contradiction:
Ease of operationVSMeasurement precision

Solution Approach 1:

The patent implements universality by designing a speech recognition system that serves multiple functions: it accurately recognizes native speakers while simultaneously providing tailored support for non-native speakers at various skill levels. The system universally applies intent recognition and feedback mechanisms across all user types, enabling free speech input from learners without sacrificing recognition performance for any group.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The system uses feedback by providing real-time corrections and suggestions to non-native speakers when recognition confidence is low or errors are detected. This feedback loop allows learners to freely input speech while the system continuously improves recognition accuracy by learning from user corrections and adapting the speech models accordingly.

Inventive Principle:
Principle #23Feedback

3Productivity

If the system provides the same response based on scenario, then implementation is simple, but learning effect is not satisfactory

Engineering Contradiction:
Improvelearning effectivenessVSAvoidsystem complexity
Core Design Contradiction:
ProductivityVSDevice complexity

Solution Approach 1:

The patent applies dynamics by making the system response adaptive rather than static. The feedback processing device dynamically generates responses based on the user's skill level, dialogue context, and speech intent, rather than providing fixed scenario-based responses. This dynamic adaptation significantly improves learning effectiveness by tailoring feedback to each learner's specific needs and progress stage.

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The system segments the feedback mechanism into multiple components: speech-based models for different skill levels, dialogue context-based models for situational understanding, and intent recognition modules for determining user goals. This segmentation allows the complex system to be managed through modular components, each handling a specific aspect of the interaction, thereby improving learning effectiveness without overwhelming system complexity.

Inventive Principle:
Principle #1Segmentation

4Measurement precision

If skill level-specific models are used, then speech intent recognition accuracy is improved, but device complexity increases

Engineering Contradiction:
Improveintent recognition accuracyVSAvoidmodel management complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The patent manages model complexity through parameter changes by dynamically selecting and switching between different speech-based models based on the user's skill level. Rather than maintaining all models simultaneously active, the system changes the active model parameters according to user profile and context, achieving high intent recognition accuracy while keeping the operational complexity manageable through parameter-driven model selection.

Inventive Principle:
Principle #35Parameter changes

Data Source

PatentUS9767710B2Apparatus and system for speech intent recognition
Publication Date: 2017.09.19 POSTECH ACADEMY INDUSTRY FOUNDATION
  • US9767710B2 patent drawing
  • US9767710B2 patent drawing
  • US9767710B2 patent drawing

AI summary

The apparatus for foreign language study includes: a voice recognition device configured to recognize a speech entered by a user and convert the speech into a speech text; a speech intent recognition device configured to extract a user speech intent for the speech text using skill level information of the user and dialog context information; and a feedback processing device configured to extract a different expression depending on the user speech intent and a speech situation of the user. According to the present invention, the intent of a learner's speech may be determined even though the learner's skill is low, and customized expressions for various situations may be provided to the learner.