Interactive Video Frame Generation for User Engagement
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current video playback experiences are hindered by the need to pause and zoom in on content, which deteriorates visibility and fails to provide additional information, leading to a suboptimal user engagement and missed opportunities for marketers.
Innovation Solution
A method to track user interactions with video frames, such as pausing and attention-activity, to generate and store interactive versions, allowing for dynamic replacement of frames with enhanced interactivity based on user engagement and metadata, providing additional information without overwhelming users.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If the user pauses the video and varies zoom level to explore content, then the user can examine details more closely, but the visibility of content deteriorates and user engagement experience becomes poor
Solution Approach 1:
The patent transitions from 2D video playback to a 3D immersive virtual environment where users can navigate and interact with content spatially. This dimensional change allows users to explore content details without pausing or zooming, as they can physically move through the virtual space to examine objects from different angles and distances, thereby maintaining continuous video playback while improving both content exploration capability and user engagement experience
Solution Approach 2:
The patent introduces a virtual environment as an intermediary layer between the user and the video content. This virtual environment acts as a mediator that provides interactive exploration capabilities without requiring the user to interrupt video playback. The virtual environment translates traditional video content into an immersive spatial context, allowing users to examine details through natural navigation rather than pausing and zooming operations
2Loss of information
If no additional information is provided about content in the video frame, then the video remains simple and easy to play, but the user experience is suboptimal and marketing opportunities are lost
Solution Approach 1:
The patent pre-generates interactive versions of video frames in advance, creating enriched content with embedded interactive elements before delivery. This preliminary action allows the system to provide comprehensive information (product details, pricing, specifications) within the video itself without requiring complex real-time processing during playback. The interactive versions are prepared beforehand, so when users interact with the content, all necessary information is already available, eliminating the need for additional complexity during video playback
3Loss of information
If interactive versions of frames are pre-generated and stored, then user engagement is enhanced with additional information, but storage requirements and processing complexity increase
Solution Approach 1:
The patent embeds interactive elements and additional information within the existing video frame structure, creating a nested hierarchy where interactive content is contained within the video data. This nesting approach allows the system to provide comprehensive information without requiring completely separate storage systems. The interactive versions are integrated into the video file structure, with metadata and interactive elements nested within the existing video container, thereby reducing overall storage requirements compared to maintaining entirely separate interactive content files
Data Source
AI summary
In one embodiment, at least one of number of times a frame is paused by a plurality of users (users) and attention-activity of the users for the frame is tracked for each frame of an asset (video etc.). An interactive version of at least one frame is pre-generated based on the tracking. The interactive version is stored to enable playing of the interactive version of the at least one frame. In another embodiment, pausing of a currently playing frame (frame) of the asset is determined. A determination to replace the frame is made based on at least one of attention-activity of a user in the frame, and detecting metadata, of the frame, specifying that the frame is to be replaced. An interactive version of the frame is generated, based on at least one of the attention of the user and the metadata, to replace the frame with the interactive version.


