Method, system, and computer-readable medium for training a captioner model to generate captions for video content by analyzing and predicting cinematic elements

WO2026117452A1PCT designated stage Publication Date: 2026-06-04NETFLIX INC

Patent Information

Authority / Receiving Office
WO · WO
Patent Type
Applications
Current Assignee / Owner
NETFLIX INC
Filing Date
2025-11-21
Publication Date
2026-06-04

Smart Images

  • Figure US2025056535_04062026_PF_FP_ABST
    Figure US2025056535_04062026_PF_FP_ABST
Patent Text Reader

Abstract

A method trains a captioner model to generate captions for video content by organizing a dataset, extracting frames, associating metadata, segmenting video, applying labels, aggregating labels, training the model, refining it, deploying it for labeling, and post-processing labels. A computing system trains a captioner model by organizing datasets, extracting frames, associating metadata, segmenting videos, applying labels, aggregating labels, training the model, refining it, deploying it for labeling, and post-processing labels. A computer-readable medium has instructions for training a captioner model by organizing datasets, extracting frames, associating metadata, segmenting videos, applying labels, aggregating labels, training the model, refining it, deploying it for labeling, and post-processing labels.
Need to check novelty before this filing date? Find Prior Art