Audio Query System Using Automatic Environmental Capture

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional audio information query methods require manual input of basic information, which is inefficient and often inaccurate, especially when users are unaware of the audio details, limiting the intelligent functions of Internet-enabled devices.

Innovation Solution

A method and device that detect environmental audio data using a microphone, transmit it to a server, and retrieve attribute information such as metadata, time indicators, and stream information, allowing for automatic identification and display of audio item details without user input.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Productivity

If manual input of basic audio information is required, then the query can be performed with existing information, but the efficiency and accuracy of audio information query deteriorates

Engineering Contradiction:
Improveaudio information query efficiencyVSAvoiduser input requirement
Core Design Contradiction:
ProductivityVSEase of operation

Solution Approach 1:

The system automatically collects environmental audio data and performs queries without requiring user manual input. The electronic device autonomously captures audio samples, transmits them to the server, and retrieves attribute information, enabling the system to serve itself rather than relying on user-provided basic information.

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The patent replaces the mechanical manual input process with automated audio capture and processing. Instead of requiring users to manually input basic audio information, the system uses a microphone to automatically collect environmental audio data and processes it through server-based recognition to obtain query results.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Productivity

If automatic audio collection is implemented, then query efficiency improves, but device complexity increases

Engineering Contradiction:
Improveaudio information query efficiencyVSAvoidsystem structure complexity
Core Design Contradiction:
ProductivityVSDevice complexity

Solution Approach 1:

The patent introduces a server as an intermediary component that handles the complex tasks of audio processing, recognition, and information retrieval. By offloading these complex operations to the server, the electronic device maintains relative simplicity while achieving automated efficient querying through the coordinated interaction between the device and server.

Inventive Principle:
Principle #24Intermediary (Mediator)

Applied Scientific Principles

This section explains which scientific principles are used to turn an abstract innovation direction into a practical engineering solution.

Function Achieved in This Case

This approach simplifies and accelerates audio information queries, improving efficiency and accuracy by eliminating the need for manual input and enabling users to quickly identify unknown audio sources, enhancing the intelligent functions of electronic devices.

Implementation Method 1

collecting an audio sample of environmental audio data with a microphone of the electronic device

Methodology Applied
Scientific EffectMicrophone transduction:

Data Source

PatentUS9348906B2Method and system for performing an audio information collection and query
Publication Date: 2016.05.24 TENCENT TECHNOLOGY (SHENZHEN) CO LTD
  • US9348906B2 patent drawing
  • US9348906B2 patent drawing
  • US9348906B2 patent drawing

AI summary

An electronic device with one or more processors, memory, and a display detects a first trigger event and, in response to detecting the first trigger event, collects a audio sample of environmental audio data associated with a media item. The device transmits information corresponding to the audio sample to a server. In response to transmitting the information, the device obtains attribute information corresponding to the audio sample, where the attribute information includes metadata for the media item, a time indicator of a position of the audio sample in the media item, and stream information for the media item. The device displays a portion of the attribute information. The device detects a second trigger event and, in response to detecting the second trigger event: determines a last obtained time indicator; streams the media item based on the stream information; and presents the media item from the last obtained time indicator.