Automated Scene Change Marker Generation for Streaming Accessibility
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
The over-the-top (OTT) streaming market lacks automated scene change markers and user-friendly navigation features, particularly for individuals with visual impairments or blindness, as current technologies rely on manual input and static image grids, which are inefficient and inaccessible.
Innovation Solution
A system that automatically and programmatically generates scene change markers and trailers using crowdsourced data from user interactions, leveraging machine learning and AI to determine viewer interest and provide audio cues for navigation, enhancing accessibility for visually impaired users.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Extent of automation
If manual input is used for scene change markers, then implementation simplicity is maintained, but automation and scalability are lost
Solution Approach 1:
The system automatically generates scene change markers by analyzing user interaction data itself, without requiring external manual input. The crowdsource server performs automated analysis of playback patterns, pauses, and user behavior to identify scene changes, allowing the system to serve itself rather than relying on manual annotation processes.
Solution Approach 2:
The patent replaces the mechanical/manual process of scene change marker creation with an automated computational system. Instead of manual analysis and input, the system uses algorithms to process user interaction data, detect patterns, and automatically generate scene change markers through computational analysis of playback behavior.
2Ease of operation
If static image grid views are used for content navigation, then interface simplicity is maintained, but accessibility for visually impaired users deteriorates
Solution Approach 1:
The patent introduces audio cues as an intermediary element that bridges the gap between visual content and visually impaired users. Audio cues provide information about scene changes and content transitions, serving as a mediator that conveys navigation information to users who cannot rely on visual feedback from static image grids.
Solution Approach 2:
The system changes the parameter of information delivery from purely visual (static images) to include auditory feedback (audio cues). By introducing a different sensory modality, the system maintains navigation functionality while making it accessible to visually impaired users who cannot process visual information effectively.
3Loss of information
If conventional search commands are used, then device simplicity is maintained, but user awareness of content progress is lost
Solution Approach 1:
The patent implements feedback by providing audio cues that inform users about their progress through content during search operations. When users perform fast-forward or rewind actions, the system analyzes user interactions and provides audible feedback about scene changes and content transitions, allowing users to understand their navigation progress without visual feedback.
Data Source
AI summary
Disclosed herein are various embodiments for providing content searching for people with visual impairments or blindness. An embodiment operates by receiving a command to search multimedia content including both video content and audio content. One or more scene changes, including a first scene change, corresponding to the video content are determined. The search command is executed on the multimedia content. It is detected that the multimedia content has reached the first scene change responsive to the executing the search command. An audio cue t is audibly output responsive to the detection.


