Teleprompter Text Synchronization via Speech Acoustic Features

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Teleprompters often display text at inappropriate speeds, leading to poor broadcast performance due to serialization issues, where users struggle to follow the content in real-time.

Innovation Solution

A method and apparatus for text content matching that determines a target utterance in a target text based on collected speech information, using acoustic features processed by audio following methods to differentiate and display the corresponding text, enabling intelligent prompting and improved broadcast synchronization.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If the teleprompter displays broadcast text at a fixed scrolling speed, then the display mechanism is simple, but the broadcast user cannot follow the content in real-time resulting in poor broadcast effect

Engineering Contradiction:
Improvebroadcast synchronizationVSAvoidteleprompter control system
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The system collects speech information from the broadcast user, processes it through acoustic feature extraction and audio following methods, and uses the recognition results to dynamically adjust text display positioning. This feedback loop enables the teleprompter to adapt to the user's actual speaking pace and improve synchronization without requiring manual speed adjustments.

Inventive Principle:
Principle #23Feedback

Solution Approach 2:

The teleprompter transitions from a fixed-speed scrolling mechanism to a dynamic display system that adjusts text positioning based on real-time speech recognition results. The text display location is no longer static but changes dynamically according to the user's speech progress, allowing flexible adaptation to varying speaking speeds.

Inventive Principle:
Principle #15Dynamics

2Productivity

If the teleprompter displays content at high speed, then more content can be shown, but the broadcast user cannot follow the problem in time

Engineering Contradiction:
Improvecontent display rateVSAvoiduser following ability
Core Design Contradiction:
ProductivityVSEase of operation

Solution Approach 1:

The system continuously monitors the user's speech through acoustic feature collection and audio following, providing real-time feedback on speech progress. This feedback enables the teleprompter to adjust text display positioning dynamically, ensuring that the displayed content matches the user's actual speaking pace rather than following a predetermined speed schedule.

Inventive Principle:
Principle #23Feedback

Solution Approach 2:

The system changes the display parameter from fixed scrolling speed to dynamic text positioning based on speech recognition results. By adjusting the text position according to the recognized speech content and timing, the system effectively adapts the display rate to match the user's speaking ability without requiring explicit speed control.

Inventive Principle:
Principle #35Parameter changes

3Device complexity

If the teleprompter uses traditional scrolling display, then the display mechanism is simple, but serialization issues occur resulting in poor broadcast effect

Engineering Contradiction:
Improvedisplay mechanismVSAvoidbroadcast effect
Core Design Contradiction:
Device complexityVSReliability

Solution Approach 1:

The system implements a feedback mechanism where speech information is collected, processed through acoustic feature extraction, and used to control text display positioning. This closed-loop control eliminates serialization issues by ensuring that text is displayed only when and where the user is actually speaking, rather than following a predetermined scrolling schedule.

Inventive Principle:
Principle #23Feedback

Solution Approach 2:

The system replaces the traditional mechanical scrolling display mechanism with an intelligent positioning system based on audio following and speech recognition. Instead of continuously scrolling text at a fixed rate, the system uses acoustic processing to determine text position, substituting a simple mechanical approach with a more complex but reliable intelligent control system.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Data Source

PatentUS20240428784A1Method, apparatus, electronic device and storage medium for text content matching
Publication Date: 2024.12.26 BEIJING ZITIAO NETWORK TECH CO LTD
  • US20240428784A1 patent drawing
  • US20240428784A1 patent drawing

AI summary

Embodiments of the present disclosure provide a method, apparatus, electronic device, and storage medium for text content matching. The method of text content matching includes: in accordance with a collection of to-be-processed speech information, determining a to-be-processed acoustic feature corresponding to the to-be-processed speech information (S110); processing, based on an audio following method, the to-be-processed acoustic feature to obtain a to-be-matched utterance corresponding to the to-be-processed acoustic feature (S120); and determining a target utterance associated with the to-be-matched utterance in target text and differentiating a display of the target utterance in the target text (S130).