Image Text Broadcasting Navigation and Fluency
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current text-to-speech (TTS) broadcasting technologies do not support forward and backward functions, which are essential for users like visually impaired and hearing-impaired individuals, and lack semantic cohesion and fluency in broadcasting text data.
Innovation Solution
An image text broadcasting method that performs character recognition on text lines, stores the data in a storage space, and uses association information to enable sequential and specified broadcasting, allowing for forward and backward navigation and improving broadcast speed and coherence.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If current TTS broadcasting technology is used, then text data can be broadcast, but forward and backward navigation functions are not supported
Solution Approach 1:
The patent applies preliminary action by pre-processing the original text into segmented text data with associated position information stored in memory before broadcasting. This allows the system to quickly retrieve and navigate to specific text segments during broadcasting without requiring re-processing, enabling efficient forward and backward navigation while maintaining complete broadcasting functionality.
Solution Approach 2:
The patent segments the original text into multiple text data units, each with position information, and stores them separately in memory. This segmentation enables independent retrieval and broadcasting of specific text segments, making forward and backward navigation feasible while preserving the ability to broadcast complete text when needed.
2Reliability
If text data is broadcast sequentially without segmentation, then broadcasting can be performed, but semantic cohesion and fluency are lacking
Solution Approach 1:
The patent implements feedback by continuously tracking the current broadcasting position and using it to determine which text segment to broadcast next. This position-based feedback mechanism ensures that text segments are broadcast in the correct sequential order, maintaining semantic cohesion while enabling efficient navigation and improving overall broadcasting fluency.
Solution Approach 2:
The patent performs preliminary segmentation and position tagging of text data before broadcasting. This pre-organization of text with position information allows the system to maintain semantic coherence by systematically retrieving segments in order while improving broadcasting efficiency through direct access to pre-processed text units.
3Adaptability or versatility
If all text data is stored in memory for complete broadcasting, then broadcasting completeness is achieved, but navigation efficiency is reduced
Solution Approach 1:
The patent segments text data into smaller units with position information and stores them in memory. This segmentation allows the system to store complete text content while enabling rapid navigation to specific segments by position, reducing navigation time compared to processing entire text blocks, while still supporting complete broadcasting when required.
Solution Approach 2:
The patent performs preliminary segmentation and position tagging of text data before broadcasting operations. This pre-processing creates an efficient data structure that enables both complete broadcasting and quick navigation to specific segments, reducing navigation time without sacrificing broadcasting completeness.
Data Source
AI summary
An image text broadcasting method, an electronic device, and a storage medium are provided. The image text broadcasting method includes: receiving a specified broadcast indication; determining a current broadcast progress about broadcast data in response to the specified broadcast indication; acquiring a next piece of broadcast data from a first text according to the current broadcast progress and the specified broadcast indication, in which the first text is composed of text data recognized and stored for a text in a text area of an image.


