Dynamic Media Captioning With User-Generated AR/VR Feedback in Home Networks

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Consumers lack the capability to obtain or create caption content for media content that is tailored to their specific requests, limiting personalized engagement and accessibility, especially in noisy or complex audio environments.

Innovation Solution

A media control device that communicates with microphones and cameras to identify user-specific caption requests, using machine learning algorithms to generate or utilize user-provided caption content, and integrate it with media content in a customizable format.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If captioning is provided for all media content, then accessibility is improved, but device complexity and processing requirements increase

Engineering Contradiction:
ImproveaccessibilityVSAvoiddevice complexity
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

The system enables users to self-generate caption content by recording and transcribing their own commentary, eliminating the need for professional captioning services. The media control device automatically processes user recordings into caption content, allowing users to serve their own captioning needs without external assistance.

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The captioning system transitions from static pre-generated captions to dynamic user-generated captions that can be created and modified in real-time. Users can record captions during media playback and have them automatically processed, making the captioning process adaptive and flexible rather than fixed and rigid.

Inventive Principle:
Principle #15Dynamics

2Adaptability or versatility

If user-generated caption content is allowed, then personalization is improved, but ease of operation deteriorates

Engineering Contradiction:
ImprovepersonalizationVSAvoidease of operation
Core Design Contradiction:
Adaptability or versatilityVSEase of operation

Solution Approach 1:

The system replaces manual caption creation with automatic speech-to-text conversion. Users simply speak their captions, and the system automatically transcribes and processes them into formatted caption content, eliminating the need for manual typing or editing of caption text.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Solution Approach 2:

The media control device acts as an intermediary that automatically processes user speech into formatted caption content. It handles the complex tasks of speech recognition, text formatting, and synchronization with media content, shielding users from technical complexity while enabling personalized captioning.

Inventive Principle:
Principle #24Intermediary (Mediator)

3Productivity

If caption content is integrated with media content, then engagement is improved, but loss of information increases

Engineering Contradiction:
ImproveengagementVSAvoidloss of information
Core Design Contradiction:
ProductivityVSLoss of information

Solution Approach 1:

The system separates caption content from the original media content, allowing users to independently control, activate, or deactivate captions. This segmentation enables users to engage with captioned content when needed while preserving access to the original uncaptioned media content, preventing information loss.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The system allows dynamic adjustment of caption parameters such as visibility, timing, and content display. Users can modify caption parameters to optimize engagement while maintaining the integrity of the original media content, ensuring that caption integration enhances rather than degrades the viewing experience.

Inventive Principle:
Principle #35Parameter changes

Data Source

PatentUS12395708B2Providing dynamic media captioning and augmented/virtual reality feedback in home network environments
Publication Date: 2025.08.19 ARRIS ENTERPRISES LLC
  • US12395708B2 patent drawing
  • US12395708B2 patent drawing
  • US12395708B2 patent drawing

AI summary

Technologies are disclosed for providing captioning for media content that may be performed by a media control device. An input may be received indicating at least one request for caption content for at least a part of the media content. One or more frames of the media content that may correspond to the request for caption content may be ascertained. Specific content from the one or more frames of the media content that may correspond to the request for caption content may be ascertained. At least one source of the caption content may be identified. The caption content may be provided in a format such that the caption content may be displayable with the one or more frames of the media content, for example in modified presentation of the media content.