Variable-Size Audio Segment Transmission for Music Recognition
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current methods for searching music information in audio streams face challenges in accurately determining music sections, leading to increased network traffic and resource consumption, as users struggle to select precise sections and existing equipment performance is influenced by audio characteristic analysis.
Innovation Solution
A content processing device that extracts an audio signal, determines characteristic music sections based on a music-to-noise ratio, generates segments of variable size, and transmits them to a music recognition server, adjusting segment size within a threshold range to optimize network traffic and resource usage, and prioritizes segments for efficient recognition.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If the user selects an exact music section manually, then the user can determine the section, but it becomes difficult for the user to search the section as desired and increases operation complexity
Solution Approach 1:
The system performs automatic music section detection and segmentation without requiring user intervention. The processor automatically identifies music sections based on audio characteristics and generates segments, eliminating the need for manual user selection while maintaining high precision in music section identification
Solution Approach 2:
The patent replaces manual user interaction (mechanical operation) with automated audio signal processing. The system uses audio characteristic analysis and automatic segmentation algorithms to substitute the manual selection process, thereby improving ease of operation while maintaining or enhancing section selection accuracy
2Reliability
If the user selects a long music section, then the section can be searched, but excessive network traffic and resource consumption occur
Solution Approach 1:
The patent divides the audio stream into multiple segments based on detected music sections. Instead of transmitting the entire audio stream or long continuous sections, the system segments the audio into smaller units and transmits only relevant segments containing music information, thereby reducing network traffic and resource consumption while maintaining recognition reliability
Solution Approach 2:
The system extracts only the essential music-containing segments from the audio stream for transmission to the server. By identifying and extracting specific music sections rather than transmitting the complete audio data, the patent reduces network traffic and energy consumption while preserving the reliability needed for accurate music recognition
3Productivity
If the audio section is divided by client device, then the section can be separated, but data cost increases due to excessive traffic and device resources are consumed
Solution Approach 1:
The patent implements dynamic segment size adjustment based on music section characteristics. The processor adaptively determines segment boundaries and sizes according to the detected music content, optimizing the balance between processing efficiency and resource consumption. This dynamic approach allows the system to process music sections efficiently while minimizing unnecessary device resource usage
4Adaptability or versatility
If a fixed-size segment is transmitted, then the transmission is simple, but it cannot adapt to variable music section lengths and reduces recognition accuracy
Solution Approach 1:
The patent employs dynamic segment size determination based on music section detection. The processor adjusts segment lengths according to the actual music content boundaries, allowing segments to be neither too short nor too long. This dynamic adaptation ensures that each segment contains appropriate music information for accurate recognition while maintaining transmission efficiency
Solution Approach 2:
The system changes the segment size parameter dynamically based on music section characteristics. Instead of using a fixed segment size, the patent adjusts the segment length parameter according to the detected music content, thereby achieving both adaptability to variable music section lengths and high recognition accuracy through optimized segment parameters
Data Source
AI summary
A content processing device is provided. The content processing device includes a receiver configured to receive a content, an audio processor configured to extract an audio signal by decoding audio data included in the content, a processor configured to determine a characteristic section in the audio signal based on a ratio of music information of the audio signal, and detect a segment corresponding to the characteristic section in the audio signal; and a communicator configured to transmit the segment to a music recognition server, and a size of the segment is determined variably within a threshold range.


