Event Video Focus Selection via Participant Action Detection

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Remote meeting systems typically focus on a limited number of locations during events, failing to capture the viewer's focus as in-person participants can control their view by turning their heads or directing their gaze, while remote viewers lack control over the content focus.

Innovation Solution

A processor monitors actions of participants at an event, interpreting gestures and vocalizations to automatically identify and record locations of interest, combining these into a video for indirect viewers, allowing them to see what direct viewers are focused on.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If remote meeting systems record a limited number of fixed locations, then the system complexity is reduced and recording is simplified, but the ability to capture viewer focus and adapt to participant interest is lost

Engineering Contradiction:
Improveability to capture viewer focusVSAvoidsystem complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The system automatically monitors participant actions, interprets gestures and vocalizations, and determines locations of interest without human intervention. The processor self-manages the recording focus by detecting participant engagement and autonomously switching between locations based on interpreted participant behavior, eliminating the need for manual operator control.

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The system performs multiple functions using a single integrated approach: it records video, monitors participant actions, interprets gestures and vocalizations, determines locations of interest, and dynamically adjusts recording focus all through one system. This multi-functional approach allows the system to adapt to different participant behaviors and locations without requiring separate specialized systems.

Inventive Principle:
Principle #6Universality (Multi-functionality)

2Loss of information

If the system dynamically adjusts recording locations based on participant actions, then the ability to match participant focus is improved, but the processing complexity and computational requirements increase

Engineering Contradiction:
Improveinformation about participant focusVSAvoidprocessing complexity
Core Design Contradiction:
Loss of informationVSDevice complexity

Solution Approach 1:

The system replaces manual mechanical control of recording devices with automated electronic processing. Instead of operators physically adjusting cameras or microphones, the system uses processors to monitor participant actions, interpret gestures and vocalizations, and automatically control recording devices, substituting mechanical operations with electronic automation.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Solution Approach 2:

The system continuously monitors participant actions and uses this feedback to dynamically adjust recording focus. By detecting gestures, vocalizations, and other participant behaviors in real-time, the system creates a closed-loop control mechanism where participant actions directly influence recording location selection, ensuring the recording always reflects current participant interest.

Inventive Principle:
Principle #23Feedback

3Loss of information

If multiple locations are monitored and recorded simultaneously, then complete event coverage is improved, but the ability to maintain viewer engagement with focus points is reduced

Engineering Contradiction:
Improveevent coverageVSAvoidviewer engagement
Core Design Contradiction:
Loss of informationVSEase of operation

Solution Approach 1:

The system extracts and focuses on only the most relevant locations of interest from the entire event space at any given time. Rather than presenting all recorded locations simultaneously, it selectively extracts the specific locations where participant attention is concentrated, presenting only those extracted focus points to maintain viewer engagement while still providing comprehensive event coverage over time.

Inventive Principle:
Principle #2Taking out (Extraction)

Data Source

PatentUS11182600B2Automatic selection of event video content
Publication Date: 2021.11.23 INTERNATIONAL BUSINESS MACHINE CORPORATION
  • US11182600B2 patent drawing
  • US11182600B2 patent drawing
  • US11182600B2 patent drawing

AI summary

A processor may record a first location at an event with at least one person. The processor may monitor a plurality of actions of that at least one person at the first location. The processor may interpret at least one action of the at least one person that indicates a change of interest to a second location at the event. Based on the at least one action, the processor may determine the second location at the event. The processor may record the second location at the event.