Event Video Focus Selection via Participant Action Detection
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Remote meeting systems typically focus on a limited number of locations during events, failing to capture the viewer's focus as in-person participants can control their view by turning their heads or directing their gaze, while remote viewers lack control over the content focus.
Innovation Solution
A processor monitors actions of participants at an event, interpreting gestures and vocalizations to automatically identify and record locations of interest, combining these into a video for indirect viewers, allowing them to see what direct viewers are focused on.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If remote meeting systems record a limited number of fixed locations, then the system complexity is reduced and recording is simplified, but the ability to capture viewer focus and adapt to participant interest is lost
Solution Approach 1:
The system automatically monitors participant actions, interprets gestures and vocalizations, and determines locations of interest without human intervention. The processor self-manages the recording focus by detecting participant engagement and autonomously switching between locations based on interpreted participant behavior, eliminating the need for manual operator control.
Solution Approach 2:
The system performs multiple functions using a single integrated approach: it records video, monitors participant actions, interprets gestures and vocalizations, determines locations of interest, and dynamically adjusts recording focus all through one system. This multi-functional approach allows the system to adapt to different participant behaviors and locations without requiring separate specialized systems.
2Loss of information
If the system dynamically adjusts recording locations based on participant actions, then the ability to match participant focus is improved, but the processing complexity and computational requirements increase
Solution Approach 1:
The system replaces manual mechanical control of recording devices with automated electronic processing. Instead of operators physically adjusting cameras or microphones, the system uses processors to monitor participant actions, interpret gestures and vocalizations, and automatically control recording devices, substituting mechanical operations with electronic automation.
Solution Approach 2:
The system continuously monitors participant actions and uses this feedback to dynamically adjust recording focus. By detecting gestures, vocalizations, and other participant behaviors in real-time, the system creates a closed-loop control mechanism where participant actions directly influence recording location selection, ensuring the recording always reflects current participant interest.
3Loss of information
If multiple locations are monitored and recorded simultaneously, then complete event coverage is improved, but the ability to maintain viewer engagement with focus points is reduced
Solution Approach 1:
The system extracts and focuses on only the most relevant locations of interest from the entire event space at any given time. Rather than presenting all recorded locations simultaneously, it selectively extracts the specific locations where participant attention is concentrated, presenting only those extracted focus points to maintain viewer engagement while still providing comprehensive event coverage over time.
Data Source
AI summary
A processor may record a first location at an event with at least one person. The processor may monitor a plurality of actions of that at least one person at the first location. The processor may interpret at least one action of the at least one person that indicates a change of interest to a second location at the event. Based on the at least one action, the processor may determine the second location at the event. The processor may record the second location at the event.


