Media File Extractor Structure for Track Referencing
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing solutions for signaling and composing media data across multiple tracks are complex and not fully compliant with existing mechanisms, particularly when one track references another, leading to overhead and inefficiencies in data transmission and parsing.
Innovation Solution
A method for encapsulating media data into a media file that includes a first track with media samples containing NAL units and a second track with an extractor structure that references data entities in the first track, using a copy mode attribute to specify how data entities are referenced and extracted, allowing for efficient data extraction and composition.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Loss of information
If existing mechanisms for referencing data entities across tracks are used, then data transmission can be achieved, but the signaling becomes complex and generates overhead
Solution Approach 1:
The patent extracts the complexity of data entity referencing by introducing an extractor structure that separates the reference mechanism from the media sample data. The extractor contains a copy mode attribute that explicitly specifies how data entities should be referenced, removing the need for complex implicit parsing rules and reducing signaling overhead while maintaining referential integrity across tracks
Solution Approach 2:
The extractor structure acts as an intermediary between tracks, providing a standardized interface for referencing data entities. The copy mode attribute within the extractor serves as a mediator that clearly defines the referencing behavior, simplifying the interaction between tracks and reducing parsing complexity compared to direct track-to-track referencing
2Reliability
If existing solutions for composing tracks from groups of tracks are used, then track composition can be achieved, but the solutions are not fully compliant with existing mechanisms and lack clear definition
Solution Approach 1:
The patent segments the track composition process into distinct, well-defined components: the extractor structure with explicit copy mode attributes, and the referenced media samples. This segmentation provides clear implementation boundaries and compliance checkpoints, making it easier to implement correctly while ensuring full compliance with existing ISOBMFF mechanisms
Solution Approach 2:
The patent introduces the copy mode attribute as a explicit parameter that changes the behavior of data entity referencing. By providing clear parameter definitions and value meanings, the patent makes the composition process more implementable while maintaining compliance with existing standards, resolving the ambiguity in previous solutions
Data Source
AI summary
A method and device for encapsulating media data into a media file and parsing a media file. The method comprising according to one of its aspects: including, in the media file, a first track comprising media samples, each media sample contains a set of one or more NAL units; including, in the media file, a second track comprising an extractor, the extractor is a structure referencing a data entity in a media sample contained in the first track; and including, in the extractor, a copy mode attribute that identifies, in the media sample, the referenced data entity relatively to one or more NAL units contained in the media sample.


