Audio Trigger Action Item Generation in Conference Calls

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current methods for capturing and tracking action items during meetings and conversations are inefficient and lack systematic follow-up features, leading to potential non-completion of tasks and decreased productivity.

Innovation Solution

An automated system that detects audio triggers during conversations to generate and notify participants of action items, using a conference bridge controller to monitor audio content, generate action items based on detected triggers, and provide alerts to participants.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If manual entry of action items is used through server-based solutions, then action items can be stored centrally, but efficiency and accuracy decrease due to manual entry requirements

Engineering Contradiction:
Improveaccuracy of action item captureVSAvoidefficiency of action item capture
Core Design Contradiction:
ReliabilityVSProductivity

Solution Approach 1:

The patent replaces manual mechanical entry processes with automated speech recognition and natural language processing systems. The system automatically transcribes conversation audio, identifies action items through pattern recognition, and populates tracking fields without human intervention, thereby improving both accuracy and efficiency simultaneously

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Solution Approach 2:

The system enables self-service automation where the conversation itself generates the action items. The speech recognition system automatically captures speaker identities, action item descriptions, and assignees from the conversation content without requiring participants to manually document anything

Inventive Principle:
Principle #25Self-service

2Device complexity

If no automated system is used, then conference call providers can maintain simple features during meetings, but follow-up functionality is insufficient leading to lost actions

Engineering Contradiction:
Improvesimplicity of conference systemVSAvoidcompletion tracking of action items
Core Design Contradiction:
Device complexityVSReliability

Solution Approach 1:

The system performs preliminary action by automatically capturing and structuring action items during the conference call itself. The speech recognition system identifies and extracts action items in real-time, creating structured data with assignees and descriptions before the meeting ends, ensuring nothing is lost

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent introduces an intermediary automated processing layer between the conversation and the action item tracking system. This intermediary system transcribes audio, identifies action items through natural language processing, and bridges the gap between casual conversation and structured task management

Inventive Principle:
Principle #24Intermediary (Mediator)

3Ease of operation

If participants rely on static text for action items, then simple communication is maintained, but systematic tracking and follow-up become difficult

Engineering Contradiction:
Improvesimplicity of action item communicationVSAvoidsystematic tracking capability
Core Design Contradiction:
Ease of operationVSProductivity

Solution Approach 1:

The system transforms static text action items into dynamic, structured data objects. Action items captured from conversation are automatically populated with metadata including assignee identities, due dates, and status fields, enabling systematic tracking while maintaining ease of creation through natural language input

Inventive Principle:
Principle #15Dynamics

Data Source

PatentUS9432517B2Methods, apparatuses, and systems for generating an action item in response to a detected audio trigger during a conversation
Publication Date: 2016.08.30 ARLINGTON TECHNOLOGIES LLC
  • US9432517B2 patent drawing
  • US9432517B2 patent drawing
  • US9432517B2 patent drawing

AI summary

Embodiments include methods, apparatuses, and systems for generating an action item in response to a detected audio trigger during a conversation. Embodiments relate to generation of one or more action items in response to detection of an audio trigger, such as a spoken command, keyword, audio tone or other indicator, which is detected during a conversation, such as an audio or video conference or peer-to-peer conversation. The audio trigger and a portion of the conversation are then used to generate an action item relating to the audio trigger and an accompanying portion of the conversation. By automatically generating action items in real time as part of a conversation, action items can be captured and stored more efficiently, and the participants in the conversation are allowed greater confidence that all items requiring follow up actions are properly stored and organized.