Variable-Size Audio Segment Transmission for Music Recognition

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current methods for searching music information in audio streams face challenges in accurately determining music sections, leading to increased network traffic and resource consumption, as users struggle to select precise sections and existing equipment performance is influenced by audio characteristic analysis.

Innovation Solution

A content processing device that extracts an audio signal, determines characteristic music sections based on a music-to-noise ratio, generates segments of variable size, and transmits them to a music recognition server, adjusting segment size within a threshold range to optimize network traffic and resource usage, and prioritizes segments for efficient recognition.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If the user selects an exact music section manually, then the user can determine the section, but it becomes difficult for the user to search the section as desired and increases operation complexity

Engineering Contradiction:
Improvemusic section selection accuracyVSAvoiduser operation difficulty
Core Design Contradiction:
Measurement precisionVSEase of operation

Solution Approach 1:

The system performs automatic music section detection and segmentation without requiring user intervention. The processor automatically identifies music sections based on audio characteristics and generates segments, eliminating the need for manual user selection while maintaining high precision in music section identification

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The patent replaces manual user interaction (mechanical operation) with automated audio signal processing. The system uses audio characteristic analysis and automatic segmentation algorithms to substitute the manual selection process, thereby improving ease of operation while maintaining or enhancing section selection accuracy

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Reliability

If the user selects a long music section, then the section can be searched, but excessive network traffic and resource consumption occur

Engineering Contradiction:
Improvemusic recognition reliabilityVSAvoidnetwork traffic and resource consumption
Core Design Contradiction:
ReliabilityVSLoss of energy

Solution Approach 1:

The patent divides the audio stream into multiple segments based on detected music sections. Instead of transmitting the entire audio stream or long continuous sections, the system segments the audio into smaller units and transmits only relevant segments containing music information, thereby reducing network traffic and resource consumption while maintaining recognition reliability

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The system extracts only the essential music-containing segments from the audio stream for transmission to the server. By identifying and extracting specific music sections rather than transmitting the complete audio data, the patent reduces network traffic and energy consumption while preserving the reliability needed for accurate music recognition

Inventive Principle:
Principle #2Taking out (Extraction)

3Productivity

If the audio section is divided by client device, then the section can be separated, but data cost increases due to excessive traffic and device resources are consumed

Engineering Contradiction:
Improvemusic section processing efficiencyVSAvoiddevice CPU and battery consumption
Core Design Contradiction:
ProductivityVSLoss of energy

Solution Approach 1:

The patent implements dynamic segment size adjustment based on music section characteristics. The processor adaptively determines segment boundaries and sizes according to the detected music content, optimizing the balance between processing efficiency and resource consumption. This dynamic approach allows the system to process music sections efficiently while minimizing unnecessary device resource usage

Inventive Principle:
Principle #15Dynamics

4Adaptability or versatility

If a fixed-size segment is transmitted, then the transmission is simple, but it cannot adapt to variable music section lengths and reduces recognition accuracy

Engineering Contradiction:
Improvesegment size adaptabilityVSAvoidmusic recognition accuracy
Core Design Contradiction:
Adaptability or versatilityVSMeasurement precision

Solution Approach 1:

The patent employs dynamic segment size determination based on music section detection. The processor adjusts segment lengths according to the actual music content boundaries, allowing segments to be neither too short nor too long. This dynamic adaptation ensures that each segment contains appropriate music information for accurate recognition while maintaining transmission efficiency

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The system changes the segment size parameter dynamically based on music section characteristics. Instead of using a fixed segment size, the patent adjusts the segment length parameter according to the detected music content, thereby achieving both adaptability to variable music section lengths and high recognition accuracy through optimized segment parameters

Inventive Principle:
Principle #35Parameter changes

Data Source

PatentUS9910919B2Content processing device and method for transmitting segment of variable size, and computer-readable recording medium
Publication Date: 2018.03.06 SAMSUNG ELECTRONICS CO LTD
  • US9910919B2 patent drawing
  • US9910919B2 patent drawing
  • US9910919B2 patent drawing

AI summary

A content processing device is provided. The content processing device includes a receiver configured to receive a content, an audio processor configured to extract an audio signal by decoding audio data included in the content, a processor configured to determine a characteristic section in the audio signal based on a ratio of music information of the audio signal, and detect a segment corresponding to the characteristic section in the audio signal; and a communicator configured to transmit the segment to a music recognition server, and a size of the segment is determined variably within a threshold range.