Real-Time Speech Recognition for Intralingual Supertitling

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Students often struggle to follow and understand the meaning of complete sentences or clauses in a target language, as they focus on understanding the later portions of a sentence and may forget the beginning, leading to a loss of overall meaning, despite speaking in the target language promoting language learning over time.

Innovation Solution

Employing speech recognition technology to convert spoken content from a teacher in a target language into corresponding text in real time, allowing students to see the text alongside hearing it, providing a multi-sensory language learning experience that persists the teacher's speech for reference, and also converting student speech for self-assessment.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Productivity

If the teacher speaks in long sentences in the target language, then the language learning effectiveness is improved, but the students lose track of the overall meaning as they focus on later portions and forget the beginning

Engineering Contradiction:
Improvelanguage learning effectivenessVSAvoidoverall meaning of sentence
Core Design Contradiction:
ProductivityVSLoss of information

Solution Approach 1:

The system performs preliminary action by displaying the text of the teacher's speech to students before they need to process it for understanding. The text appears on screen as the teacher speaks, allowing students to preview and retain the beginning portions of sentences while listening, thus preventing information loss about the overall meaning.

Inventive Principle:
Principle #10Preliminary action

2Productivity

If the teacher speaks primarily in the target language, then the students' listening and oral production skills are developed, but the students have difficulty understanding complete sentences or clauses

Engineering Contradiction:
Improvelanguage skill developmentVSAvoidcomprehension of speech
Core Design Contradiction:
ProductivityVSEase of operation

Solution Approach 1:

The system introduces text as an intermediary between the teacher's spoken language and the students' comprehension. The displayed text serves as a mediator that bridges the gap between hearing the target language and understanding complete sentences, allowing students to follow along without needing to perfectly process every auditory detail in real-time.

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The system adds another dimension to language instruction by combining auditory input (teacher's speech) with visual input (displayed text). This multi-sensory approach allows students to process language information through both hearing and reading channels simultaneously, making comprehension easier while maintaining immersion in the target language.

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

3Productivity

If speech recognition technology is used to convert spoken content to text in real time, then students can see and hear the content simultaneously, but the system complexity increases

Engineering Contradiction:
Improvelanguage learning effectivenessVSAvoidsystem complexity
Core Design Contradiction:
ProductivityVSDevice complexity

Solution Approach 1:

The system employs self-service by using speech recognition technology that automatically converts the teacher's spoken words into text without requiring manual transcription. The system serves itself by capturing audio input, processing it through recognition algorithms, and displaying the resulting text, eliminating the need for human transcribers or complex manual preparation of lesson materials.

Inventive Principle:
Principle #25Self-service

Data Source

PatentUS10026329B2Intralingual supertitling in language acquisition
Publication Date: 2018.07.17 ISSLA ENTERPRISES
  • US10026329B2 patent drawing
  • US10026329B2 patent drawing
  • US10026329B2 patent drawing

AI summary

A technique for facilitating language instruction employs speech recognition technology to convert spoken content from a teacher in a target language to corresponding text in the target language, substantially in real time, and to project the converted text for viewing by the students. Students are thus able both to hear the spoken content from the teacher and to see the corresponding text, thus enjoying a multi-sensory, intralingual language learning experience that combines both listening and reading.