Speaker-Specific Transcript Generation for Record Population

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing systems face challenges in accurately populating electronic records with information from communication sessions between individuals, particularly in distinguishing between speakers' contributions, leading to errors in data tagging and record field population.

Innovation Solution

A computer-implemented method generates transcripts for each participant in a communication session and identifies specific words or phrases to populate record fields, ensuring accuracy by distinguishing between speakers' inputs and transmitting this data to an electronic record database.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Device complexity

If a single transcript is generated from combined audio of multiple speakers, then the transcription process is simplified, but the accuracy of attributing information to specific speakers deteriorates

Engineering Contradiction:
Improvetranscription system complexityVSAvoidspeaker attribution accuracy
Core Design Contradiction:
Device complexityVSMeasurement precision

Solution Approach 1:

The patent divides the audio stream into separate segments for each speaker and generates individual transcripts for each speaker. This segmentation allows the system to maintain simple transcription processes for each speaker while accurately attributing information to the correct speaker, thereby resolving the contradiction between system simplicity and attribution accuracy.

Inventive Principle:
Principle #1Segmentation

2Manufacturing precision

If manual review of transcripts is performed to ensure accurate record population, then data accuracy improves, but processing time and labor costs increase

Engineering Contradiction:
Improverecord population accuracyVSAvoidprocessing time
Core Design Contradiction:
Manufacturing precisionVSLoss of time

Solution Approach 1:

The patent implements an automated feedback mechanism where the system generates transcripts, identifies potential record field values, and validates the population process automatically. This feedback loop ensures high accuracy in record population without requiring manual review, thereby maintaining precision while reducing processing time and eliminating manual labor costs.

Inventive Principle:
Principle #23Feedback

3Adaptability or versatility

If all words from both speakers are considered for record field population, then more potential values are available, but the precision of matching words to correct record fields deteriorates

Engineering Contradiction:
Improvevalue selection flexibilityVSAvoidfield-value matching accuracy
Core Design Contradiction:
Adaptability or versatilityVSMeasurement precision

Solution Approach 1:

The patent applies local quality by analyzing the transcript of each speaker individually and identifying record field values based on the specific context and role of each speaker. This approach maintains versatility in value selection while ensuring high precision in matching words to the correct record fields, as each speaker's contributions are evaluated in their local context rather than mixed together.

Inventive Principle:
Principle #3Local quality

Data Source

PatentUS9824691B1Automated population of electronic records
Publication Date: 2017.11.21 SORENSON IP HOLDINGS LLC
  • US9824691B1 patent drawing
  • US9824691B1 patent drawing
  • US9824691B1 patent drawing

AI summary

A computer-implemented method to populate an electronic record may include generating first transcript data of first audio of a first speaker during a conversation between the first speaker and a second speaker. The method may also include generating second transcript data of second audio of the second speaker during the conversation and identifying one or more words from the first transcript data as being a value for a record field based on the identified words corresponding to the record field and the one or more words being from the first transcript data and not being from the second transcript data. The method may further include providing the identified words to an electronic record database as a value for the record field of a user record of the first speaker.