Automated Scene Change Marker Generation for Streaming Accessibility

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

The over-the-top (OTT) streaming market lacks automated scene change markers and user-friendly navigation features, particularly for individuals with visual impairments or blindness, as current technologies rely on manual input and static image grids, which are inefficient and inaccessible.

Innovation Solution

A system that automatically and programmatically generates scene change markers and trailers using crowdsourced data from user interactions, leveraging machine learning and AI to determine viewer interest and provide audio cues for navigation, enhancing accessibility for visually impaired users.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Extent of automation

If manual input is used for scene change markers, then implementation simplicity is maintained, but automation and scalability are lost

Engineering Contradiction:
Improveautomated scene change marker generationVSAvoidsystem complexity
Core Design Contradiction:
Extent of automationVSDevice complexity

Solution Approach 1:

The system automatically generates scene change markers by analyzing user interaction data itself, without requiring external manual input. The crowdsource server performs automated analysis of playback patterns, pauses, and user behavior to identify scene changes, allowing the system to serve itself rather than relying on manual annotation processes.

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The patent replaces the mechanical/manual process of scene change marker creation with an automated computational system. Instead of manual analysis and input, the system uses algorithms to process user interaction data, detect patterns, and automatically generate scene change markers through computational analysis of playback behavior.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Ease of operation

If static image grid views are used for content navigation, then interface simplicity is maintained, but accessibility for visually impaired users deteriorates

Engineering Contradiction:
Improvenavigation accessibilityVSAvoidcontent information accessibility
Core Design Contradiction:
Ease of operationVSLoss of information

Solution Approach 1:

The patent introduces audio cues as an intermediary element that bridges the gap between visual content and visually impaired users. Audio cues provide information about scene changes and content transitions, serving as a mediator that conveys navigation information to users who cannot rely on visual feedback from static image grids.

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The system changes the parameter of information delivery from purely visual (static images) to include auditory feedback (audio cues). By introducing a different sensory modality, the system maintains navigation functionality while making it accessible to visually impaired users who cannot process visual information effectively.

Inventive Principle:
Principle #35Parameter changes

3Loss of information

If conventional search commands are used, then device simplicity is maintained, but user awareness of content progress is lost

Engineering Contradiction:
Improvecontent progress feedbackVSAvoidsearch system complexity
Core Design Contradiction:
Loss of informationVSDevice complexity

Solution Approach 1:

The patent implements feedback by providing audio cues that inform users about their progress through content during search operations. When users perform fast-forward or rewind actions, the system analyzes user interactions and provides audible feedback about scene changes and content transitions, allowing users to understand their navigation progress without visual feedback.

Inventive Principle:
Principle #23Feedback

Data Source

PatentUS20220353588A1Program searching for people with visual impairments or blindness
Publication Date: 2022.11.03 ROKU INC
  • US20220353588A1 patent drawing
  • US20220353588A1 patent drawing
  • US20220353588A1 patent drawing

AI summary

Disclosed herein are various embodiments for providing content searching for people with visual impairments or blindness. An embodiment operates by receiving a command to search multimedia content including both video content and audio content. One or more scene changes, including a first scene change, corresponding to the video content are determined. The search command is executed on the multimedia content. It is detected that the multimedia content has reached the first scene change responsive to the executing the search command. An audio cue t is audibly output responsive to the detection.