Media File Aggregation via Content Feature Recognition

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current automatic aggregation of media files on mobile devices is inflexible and inefficient, often requiring manual selection and aggregation, which is time-consuming and does not meet user needs effectively.

Innovation Solution

A method that recognizes content features of media files, such as image and sound features, to automatically determine aggregation themes and synthesize media files into videos, allowing for more flexible and efficient aggregation based on content similarity and quality.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Extent of automation

If automatic aggregation is performed based on description information (shooting time and shooting location), then aggregation can be performed automatically, but the aggregation results are monotonous and inflexible

Engineering Contradiction:
Improveautomatic aggregationVSAvoidaggregation flexibility
Core Design Contradiction:
Extent of automationVSAdaptability or versatility

Solution Approach 1:

The patent changes the aggregation parameters from traditional description information (shooting time, shooting location) to content features extracted through image recognition and sound recognition. This parameter change enables the system to aggregate media files based on content similarity, scene type, character appearance, and audio characteristics, thereby achieving both automatic aggregation and flexible aggregation results that meet diverse user needs.

Inventive Principle:
Principle #35Parameter changes

2Adaptability or versatility

If manual selection and aggregation of media files is performed, then aggregation flexibility can be improved, but it requires a lot of time and energy from the user

Engineering Contradiction:
Improveaggregation flexibilityVSAvoidtime consumption
Core Design Contradiction:
Adaptability or versatilityVSLoss of time

Solution Approach 1:

The patent implements self-service by enabling the system to automatically recognize content features of media files, determine aggregation themes, and perform aggregation without user intervention. The image recognition module and sound recognition module automatically extract features, and the aggregation module automatically groups media files based on these features, freeing users from manual selection and aggregation operations while maintaining flexible aggregation results.

Inventive Principle:
Principle #25Self-service

3Adaptability or versatility

If media files are aggregated based on content features recognized by the system, then aggregation flexibility and abundance of dimensions are improved, but the system complexity increases

Engineering Contradiction:
Improveaggregation dimension abundanceVSAvoidsystem complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent segments the complex media file aggregation system into independent functional modules: an image recognition module for extracting visual features, a sound recognition module for extracting audio features, and an aggregation module for grouping media files. This segmentation allows each module to perform its specific function independently, making the overall system more manageable and maintainable while achieving multi-dimensional content-based aggregation.

Inventive Principle:
Principle #1Segmentation

4Ease of operation

If aggregated media files are synthesized into a video to display results, then user experience is improved, but additional processing time and resources are required

Engineering Contradiction:
Improveuser experienceVSAvoidprocessing time
Core Design Contradiction:
Ease of operationVSLoss of time

Solution Approach 1:

The patent applies preliminary action by pre-processing media files to extract and store content features (image features and sound features) before aggregation. This pre-extraction of features enables the aggregation module to quickly match and group media files based on content similarity without requiring re-processing during aggregation, thereby reducing overall processing time while still providing enhanced user experience through video synthesis of aggregation results.

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS12067053B2Media file processing method, device, readable medium, and electronic apparatus
Publication Date: 2024.08.20 DOUYIN VISION CO LTD
  • US12067053B2 patent drawing
  • US12067053B2 patent drawing
  • US12067053B2 patent drawing

AI summary

A media file processing method includes: recognizing content features of a target media file, wherein the content features include an image feature and/or a sound feature; determining a target aggregation theme of the target media file according to the recognized content features of the target media file; determining the target media file as media files under the target aggregation theme; and synthesizing the media files under the target aggregation theme in response to a video clip instruction with respect to the target aggregation theme, to obtain a target video corresponding to the target aggregation theme.