Text Processing System for Automated Student Dictation

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Traditional dictation methods for students to reinforce word recognition are hindered by nonstandard parental pronunciation and lack of immediate feedback, requiring parental cooperation and failing to correct mistakes in a timely manner.

Innovation Solution

A text processing method and apparatus that collects and recognizes text images through gestures, broadcasts the text in voice form for dictation, and automatically checks the dictation results using image recognition, ensuring standard pronunciation and timely feedback.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If traditional parental dictation is used, then students can practice word recognition, but parental nonstandard pronunciation causes errors and misleads students

Engineering Contradiction:
Improvepronunciation accuracyVSAvoidparental cooperation requirement
Core Design Contradiction:
ReliabilityVSEase of operation

Solution Approach 1:

The system enables students to independently perform dictation practice by automatically selecting words, playing standard audio recordings, and providing automated checking through image recognition, eliminating the need for parental participation while ensuring pronunciation accuracy through pre-recorded standard audio

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The system introduces an electronic device as an intermediary between the student and the dictation process, using pre-recorded standard audio recordings as a mediator to transmit correct pronunciation without requiring parental involvement, thus resolving the contradiction between reliability and ease of operation

Inventive Principle:
Principle #24Intermediary (Mediator)

2Productivity

If parental dictation is used, then students can practice words, but parents lack time to cooperate

Engineering Contradiction:
Improvedictation practice efficiencyVSAvoidparental time investment
Core Design Contradiction:
ProductivityVSLoss of time

Solution Approach 1:

The system allows students to independently complete the entire dictation workflow including word selection, audio playback, writing practice, and result checking without requiring parental time investment, thereby maintaining productivity while eliminating the time loss associated with parental cooperation

Inventive Principle:
Principle #25Self-service

3Measurement precision

If traditional dictation is used, then students write words, but there is no automatic checking of dictation results

Engineering Contradiction:
Improvedictation checking accuracyVSAvoidfeedback time
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

The system implements automated feedback by using image recognition technology to immediately check the student's handwritten dictation results against the correct answer, providing instant measurement of accuracy and enabling immediate correction without time loss

Inventive Principle:
Principle #23Feedback

Solution Approach 2:

The system replaces the manual mechanical checking process with automated image recognition technology, which captures the student's writing and automatically compares it with the correct answer, achieving precise measurement of dictation accuracy instantaneously without time loss

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Data Source

PatentUS12136285B2Text processing method and apparatus, and electronic device and non-transitory computer-readable medium
Publication Date: 2024.11.05 DOUYIN VISION CO LTD
  • US12136285B2 patent drawing
  • US12136285B2 patent drawing

AI summary

Provided are a text processing method and apparatus, an electronic device and a non-transitory computer-readable medium. The method includes: collecting a to-be-processed text image, and performing gesture recognition on the to-be-processed text image to obtain a to-be-processed text, where the to-be-processed text is a text selected from the to-be-processed text image through a gesture; performing voice broadcasting on the to-be-processed text to prompt a user to perform dictation processing on the to-be-processed text; and collecting a dictation text image, performing recognition on the dictation text image, and determining a dictation check result according to a recognition result and the to-be-processed text.