Active Experience File Generation for Music Files
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional music playback methods limit users to passive appreciation of music within predetermined lengths, restricting active participation and immersive experiences.
Innovation Solution
An electronic device and method that acquire multiple sound sources and graphic sources from uploaded music and graphic files, generating active experience content by synchronizing these sources, and utilizing generative AI to reduce operational burdens.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If conventional music playback methods are used, then the system is simple and easy to operate, but users are limited to passive appreciation within predetermined lengths and cannot actively participate
Solution Approach 1:
The patent segments music files into multiple sound sources (e.g., vocals, instruments, beats) that can be independently manipulated. This allows users to actively participate by selecting, reordering, and remixing segments, transforming passive playback into active engagement while managing complexity through automated segmentation algorithms
Solution Approach 2:
The system dynamically generates playable content by allowing users to interactively rearrange and remix sound sources in real-time. The playback structure transitions from fixed predetermined lengths to flexible user-defined sequences, enabling active participation while the system automatically handles the complexity of dynamic content generation
2Adaptability or versatility
If active experience content is generated for each music file using multiple sound sources and graphic sources, then user engagement and immersion are improved, but the operation burden for producing content increases
Solution Approach 1:
The system automatically separates music files into multiple sound sources using AI-based audio separation technology, eliminating the need for manual tracking of individual instruments and vocals. The system self-generates the structured data and playable content from uploaded music files, significantly reducing the operation burden while maintaining high music experience quality
Solution Approach 2:
The system transforms the complexity of content generation by changing parameters from manual source separation to automated AI-based separation. By adjusting the separation quality and number of extracted sound sources as configurable parameters, the system delivers high-quality active music experiences while keeping the user interface simple and the operation burden low
3Adaptability or versatility
If sound sources and graphic sources are provided with different time units, then content generation flexibility is improved, but synchronization accuracy deteriorates
Solution Approach 1:
The patent introduces a synchronization module that acts as an intermediary between sound sources and graphic sources with different time units. This module automatically aligns temporal references, converts between different time units (e.g., audio samples to video frames), and ensures precise synchronization while preserving the flexibility of independent content generation for each media type
Data Source
AI summary
According to various embodiments, there may be provided an operation method of a server, including: obtaining an audio file and a graphic file; obtaining a plurality of audio sources and a plurality of visual sources based on at least one of the audio file or the graphic file; obtaining a plurality of processed audio sources based on changing at least some characteristics of the plurality of audio sources; obtaining a plurality of processed visual sources based on changing at least some characteristics of the plurality of visual sources; and generating at least one active experience file based on generating at least one first audio source selected from the plurality of processed audio sources, at least one first visual source selected from the plurality of processed visual sources, and a specific kind of interaction selected by a user in a related form, wherein, when the specific kind of interaction is received based on the at least one active experience file, the electronic device of the user is set to provide the at least one first audio source and the at least one first visual source.


