Text-Audio Mapping for Transcription Verification
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current transcription technologies, whether computer-generated or human-generated, lack the ability to easily correlate specific text in transcriptions with corresponding locations in audio files, making it difficult for users to verify or retrieve specific information from audio communications.
Innovation Solution
A method and apparatus that generate a mapping of transcribed text to corresponding audio communications, identifying offset times for each word or phrase, allowing users to easily locate specific points in the audio based on the text, independent of the transcription process.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If human transcription is used to improve transcription quality, then transcription accuracy is improved, but the ability to locate specific text in audio becomes difficult
Solution Approach 1:
The patent introduces a mapping file as an intermediary between the transcription and audio file. This mapping file contains timestamp information that links specific text segments to their corresponding audio locations, enabling users to easily navigate to and verify specific portions of the transcription without manually searching through the entire audio file.
2Productivity
If computer-generated transcription is used to reduce cost and time, then productivity is improved, but transcription quality deteriorates
Solution Approach 1:
The mapping file serves as an intermediary that enables verification of computer-generated transcriptions against the original audio. Users can quickly navigate to specific timestamps to verify accuracy, allowing efficient use of automated transcription while maintaining quality control through selective audio verification.
3Device complexity
If no mapping is generated between text and audio, then device complexity is reduced, but information verification becomes difficult
Solution Approach 1:
The patent segments the transcription into discrete text segments, each associated with a specific timestamp in the mapping file. This segmentation allows users to verify specific portions of the transcription independently, improving reliability without requiring complex verification systems.
Data Source
AI summary
In one embodiment, a method includes receiving at a communication device an audio communication and a transcribed text created from the audio communication, and generating a mapping of the transcribed text to the audio communication independent of transcribing the audio. The mapping identifies locations of portions of the text in the audio communication. An apparatus for mapping the text to the audio is also disclosed.


