Audio Source Identification via Server-Mediated Fingerprinting
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Users face difficulties in identifying nearby parties playing music, especially when multiple parties are present and devices use different operating systems or experience Wi-Fi connectivity issues, making it hard to join or influence the music being played.
Innovation Solution
A method using a portable electronic device to determine position coordinates and sample audio, converting it into a fingerprint, and sending these to a computer server to search for matching audio sources within a predefined range, allowing for quick identification of the closest audio source.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If all currently streaming parties are offered to the user, then the user can choose from more options, but this becomes inconvenient for the user
Solution Approach 1:
The patent segments the list of available parties by geographical location, presenting parties in a structured manner (e.g., closest party first, or organized by distance). This segmentation allows users to see multiple options without being overwhelmed, as the information is divided into manageable groups based on location.
Solution Approach 2:
The patent applies local quality by providing location-specific information about parties. Each party listing includes geographical context (distance, location details), allowing users to quickly assess relevance based on their current position. This localized information quality helps users make informed decisions without examining all parties in detail.
2Reliability
If device-to-device communication using Bluetooth or WiFi is used, then direct connection can be established, but this cannot be easily established in some situations
Solution Approach 1:
The patent introduces a server as an intermediary between portable electronic devices and audio sources. Instead of requiring direct device-to-device communication, the server mediates the connection by receiving audio samples from devices, comparing them with audio sources, and facilitating joins. This intermediary approach maintains connection reliability while eliminating the complexity of direct pairing.
Solution Approach 2:
The patent replaces the mechanical Bluetooth/WiFi direct connection system with a network-based server-mediated system. Instead of relying on proximity-based wireless protocols that require manual pairing, the system uses internet-based communication where the server handles connection logistics, substituting the mechanical pairing process with automated server-side matching.
3Measurement precision
If audio sampling and fingerprinting is performed for all audio sources, then accurate identification is achieved, but this increases processing time and complexity
Solution Approach 1:
The patent applies preliminary action by having the server perform audio sampling and fingerprinting of all audio sources in advance, before user queries. The server continuously monitors and stores audio fingerprints of active audio sources, so when a user wants to identify a party, the matching process is already prepared. This preliminary preparation maintains high identification accuracy while reducing real-time processing requirements.
Solution Approach 2:
The server acts as an intermediary that handles the complex audio processing tasks. Instead of requiring portable devices to perform sophisticated audio fingerprinting and comparison, the server centralizes this functionality. The device simply sends audio samples and receives identification results, transferring the computational complexity to the server while maintaining accurate identification.
Data Source
AI summary
The present disclosure relates in general to media streaming, such as music streaming. In particular, this disclosure presents various embodiments of methods, portable electronic devices (200), computer servers (300) and computer program that allow for identifying an audio source (e.g., a music source) that is currently outputting audio (e.g., playing music). In an example scenario, this may allow a user to identify a social gathering such as a party where music is being played.


