Audio Style Segmentation and Splicing for Seamless Playback

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current audio processing technologies, particularly in mobile devices, lack the ability to accurately clip and splice songs based on preset styles, resulting in suboptimal user experience due to manual operation and imprecise determination of song segments.

Innovation Solution

A method and device that analyze audio files to determine start and end times of segments corresponding to specific styles, allowing for clipping and splicing of audio files according to preset styles, thereby enhancing user experience by creating seamless and customized audio compositions.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If manual operation is used for clipping songs via network software, then the operation can be performed, but the accuracy of determining song segment locations deteriorates

Engineering Contradiction:
Improvemanual operation capabilityVSAvoidaccuracy of song segment location determination
Core Design Contradiction:
Ease of operationVSMeasurement precision

Solution Approach 1:

The system performs automatic style recognition and segment location determination without requiring manual operation. The audio file is automatically analyzed to identify different musical styles and their transition points, enabling the clipping function to work autonomously with high precision.

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The manual mechanical operation of clipping is replaced by an automated digital signal processing system. The system uses style recognition algorithms to automatically identify segment boundaries and perform clipping operations, substituting human manual control with automated computational methods.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Measurement precision

If style recognition analysis is performed on audio files, then the precision of segment determination is improved, but the processing time and complexity increase

Engineering Contradiction:
Improveprecision of segment determinationVSAvoidcomplexity of processing system
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The audio processing system is divided into distinct functional modules: style recognition module, segment determination module, and clipping module. Each module handles a specific aspect of the processing, making the overall complex system manageable and efficient through functional decomposition.

Inventive Principle:
Principle #1Segmentation

3Measurement precision

If automatic style recognition is implemented, then the accuracy of clipping locations is improved, but the processing time increases

Engineering Contradiction:
Improveaccuracy of clipping locationVSAvoidprocessing time
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

The style recognition and segment determination are performed in advance during the clipping process. By pre-identifying style transition points before the actual clipping operation, the system ensures accurate location determination without adding significant time to the final output generation.

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentEP3142031B1Preset style song processing method and apparatus
Publication Date: 2022.04.13 GUANGDONG OPPO MOBILE TELECOMMUNICATIONS CORP LTD
  • EP3142031B1 patent drawingFigure 1
  • EP3142031B1 patent drawingFigure 2~3
  • EP3142031B1 patent drawingFigure 4

AI summary

The present disclosure provides a method for processing songs of preset styles. The method can include the follows. A setting instruction from a user is received. Preset styles of spliced songs are set in accordance with the setting instruction. N pieces of audio files are read out, wherein N≥1 and is an integer. The N pieces of audio files are analyzed to obtain styles of the N pieces of audio files. The start time and the end time of each paragraph corresponding to each of the styles of the N pieces of audio files are determined. In accordance with the start time and the end time of each paragraph corresponding to each of the styles of the N pieces of audio files, N pieces of audio files are clipped, so as to obtain K pieces of clipped paragraph, wherein K≥1 and is an integer. The K pieces of clipped paragraph are spliced in accordance with the order of the preset styles to obtain audio files of the spliced songs. Therefore, splicing of songs of preset styles can be achieved and user experience can be improved.