Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

71results about "Audio data querying" patented technology

Music database retrieval method and system based on feature extraction

The invention discloses a music database retrieval method and system based on feature extraction, and relates to the technical field of data analysis, and the method comprises the steps: carrying out the preprocessing of an audio signal of a query track input by a user, dividing the audio into equal-length frame sequences, and generating three-channel spectrogram tensors, including a Mel spectrogram, a logarithmic amplitude spectrogram and a CQT spectrogram; extracting multi-modal features of the equal-length frame sequence, dividing melody motivation segments according to pitch change and stability of the audio signal, performing differential coding, generating a melody differential sequence, and modeling the melody differential sequence into a melody topological graph; and calculating the structural similarity between the melody topological graph and a database melody graph, screening a similar melody candidate set P, mapping a melody curve into a topological manifold through a topological data analysis method, extracting persistent homology features, and converting the persistent homology features into a persistent bar graph. According to the method, the retrieval precision and the matching credibility of the complex melody in the music database are remarkably improved.
Owner:BODA COLLEGE OF JILIN NORMAL UNIV

Bundled search processing method and apparatus, computer device, and storage medium

The application relates to a bundle search processing method and device, computer equipment and a storage medium. The method comprises the following steps: performing bundle search on a word table based on a current to-be-matched object and a current bundle width to obtain a current search result, wherein the current search result comprises a plurality of target words which are matched successfully with the current to-be-matched object and the number of which matches the current bundle width; performing width attenuation on the current bundle width according to a bundle width attenuation mode matched by the number of executed bundle searches to obtain a bundle width for next bundle search; taking each target word as a to-be-matched object for next bundle search, and iteratively performing bundle search until the bundle width obtained by attenuation is equal to a bundle width threshold value, and taking the search result at this time as a target to-be-matched object, and performing bundle search on the word table based on the bundle width threshold value to obtain a target search result. By adopting the method, the search efficiency can be improved while ensuring the search effect in an ultra-long sequence scene.
Owner:ZHAOLIAN CONSUMER FINANCE CO LTD

Audio playing control method and system based on identifier triggering, microphone, sound box equipment and storage medium

The invention discloses an audio playing control method and system based on identifier triggering, a microphone, sound box equipment and a storage medium. The method comprises the following steps: acquiring a non-contact identification signal; acquiring identification information corresponding to the non-contact identification signal; determining a corresponding target audio resource based on the identification information; and controlling an audio playing device to play the target audio resource. According to the method, the target audio resource is determined by using the non-contact identification signal, and the audio playing device is controlled to play the target audio resource, so that the sound box device can accurately position the audio content needing to be played according to the non-contact identification signal, a user does not need to carry out tedious manual operation or contact interaction, and the user experience is improved. The use convenience of the sound box equipment is improved, and the problems that the audio resource on-demand efficiency is low and even the audio resource on-demand operation cannot be completed due to the fact that the user is not familiar with the position of each function key on the sound box equipment or the screen operation logic are solved.
Owner:广东台德智联科技有限公司

Voice query QOS based on client-computed content metadata

A method includes receiving an automated speech recognition (ASR) request from a user device that includes a speech input captured by the user device and content metadata associated with the speech input. The content metadata is generated by the user device. The method also includes determining a priority score for the ASR request based on the content metadata associated with the speech input and caching the ASR request in a pre-processing backlog of pending ASR requests each having a corresponding priority score. The pending ASR requests in the pre-processing backlog are ranked in order of the priority scores. The method also includes providing, from the pre-processing backlog, one or more of the pending ASR requests to a backend-side ASR module, wherein pending ASR requests associated with higher priority scores are processed before pending ASR requests associated with lower priority scores.
Owner:GOOGLE LLC

system

We provide the system. [Solution] Means for obtaining user image information, A means for analyzing the aforementioned image information to infer the atmosphere and emotional state, means for generating musical information based on the aforementioned atmosphere and emotional state, Means for transmitting the generated music information to the user's terminal device, A system that includes this.
Owner:SOFTBANK GROUP CORP

Multilingual speech and semantic intelligent translation method and system applied to exhibition scene

The invention discloses a multilingual speech semantic intelligent translation method and system applied to an exhibition scene, and belongs to the technical field of machine translation, and the method comprises the following steps: S1, obtaining a multi-person question judgment result; s2, if the multi-person questioning judgment result is multi-person questioning, audio identification information is obtained through analysis, and otherwise, the audio identification information is directly obtained through analysis; s3, obtaining each storage question keyword, each contrast question keyword and a key matching weighting factor corresponding to each question keyword; s4, obtaining a same-group evaluation result, if the same-group evaluation result is the same group, analyzing to obtain the comprehensive matching similarity of the storage groups, otherwise, analyzing the comprehensive matching similarity of each parallel storage group; s5, obtaining a comprehensive matching judgment result, if the comprehensive matching judgment result is unqualified, performing secondary refining processing to obtain a question and answer, and otherwise, directly obtaining the question and answer; and S6, voice broadcasting is carried out, and accurate separation and language recognition of voice sources of different questioning users are achieved.
Owner:ZHEJIANG HUIZHAN ELF TECHNOLOGY CO LTD

Song searching method, device, searching system and computer readable storage medium

The application discloses a song search method and device, a search system and a computer readable storage medium, and belongs to the technical field of search. The embodiment receives a song search request, wherein the song search request comprises song search information; searches a first song from a first storage space corresponding to a first search engine based on the song search information through the first search engine, wherein the first storage space stores song information of all songs in a music platform; searches a second song from a second storage space corresponding to a second search engine based on the song search information through the second search engine, wherein the second storage space stores song information of songs meeting preset conditions in the music platform; and generates a search result corresponding to the song search request based on the first song and the second song, so that the stability of the search system and the accuracy of the search result can be improved.
Owner:HANGZHOU NETEASE CLOUD MUSIC TECH CO LTD

Information processing device, information processing method, and information processing program

To provide an information processing device, an information processing method, and an information processing program that are capable of increasing the convenience of users.SOLUTION: An information processing device includes a specification unit, an acquisition unit, and an image processing unit. The specification unit specifies a hair style that is estimated to be suitable for a face of a user, by using a learning model that is a model having learned a relationship between information indicating the characteristics of a face and a hair style evaluated to suit the face. The acquisition unit obtains a posted image including a specified hair image that is an image of the hair style specified by the specification unit. The image processing unit generates a composite image that is obtained by combining the image of the hairstyle specified by the specification unit and the face image of the user, based on the posted image obtained by the acquisition unit and the face image of the user.SELECTED DRAWING: Figure 8
Owner:LY CORP

Matching video content with podcast episodes

A system and method are provided for matching videos and podcast episodes. A data store comprising podcast episode identifiers is accessed. The podcast episode identifiers are associated with one or more podcast episode attributes. A video content item is identified. The video content item includes one or more video content item attributes. A matching podcast episode identifier that matches the video content item is determined based on the one or more podcast episode attributes and the one or more video content item attributes. A ranking of one of the video content items or the matching podcast episode identifiers is adjusted to reflect a correspondence between the video content item and the matching podcast episode identifier. Information associated with the matching podcast episode identifier is provided to a first user device.
Owner:GOOGLE LLC

system

We provide the system. [Solution] Means for acquiring audio information, A means of converting this audio information into text information, A means of analyzing the converted text information to understand the content of the inquiry or potential threat, A means of generating appropriate responses or warnings based on inquiries, A means for converting the generated text response into audio information, Means for making these operations compatible with multiple languages, A system that includes means for collecting and continuously learning from evaluation results of responses and warnings.
Owner:SOFTBANK GROUP CORP

Search method, device, equipment and readable storage medium

The application discloses a retrieval method, device and equipment and a readable storage medium. The method comprises the following steps: obtaining first feature information of to-be-retrieved content, and performing retrieval in a first feature library to obtain a first result set; when entries in the first result set are first-type entries, obtaining second feature information of the entries; performing retrieval in a second feature library containing second-type entry features according to the second feature information to obtain a second result set; and outputting a matched result sequence according to the first result set and the second result set. The application can output a matched list containing specific entries such as original content, improve the retrieval experience of users, and ensure the exposure rate of original content in a platform.
Owner:HANGZHOU NETEASE CLOUD MUSIC TECH CO LTD

System for the joint playback of playlists for vehicles and a method for its application

This document describes a system (100) for the collaborative playback of media playlists in vehicles. The system (100) comprises a server (102) that communicates with a vehicle infotainment unit (106). The infotainment unit (106) connects mobile devices (108), assigned to users (110), to the server (102) after user (110) authentication. The server (102) allows users (110) to select infotainment media to be played on the infotainment unit (106) using the mobile devices (108) and to create an initial collaborative playlist. This playlist contains the selected infotainment media, which are queued in a predefined configuration, as well as details about the respective users (110) who selected the infotainment media.The server (102) also plays the infotainment media associated with the created shared playlist when one of the users (110) has given their consent, and furthermore restricts the modification of the created shared playlist once the first created shared playlist has been selected and approved by the relevant users (110).
Owner:MERCEDES BENZ GROUP AG

Client-calculated content metadata-based voice inquiry service quality (QoS)

To deal with a traffic abrupt increase in a processing stack of a server base.SOLUTION: A method includes receiving an automated speech recognition (ASR) request from a user device that includes a speech input captured by the user device and content metadata associated with the speech input. The content metadata is generated by the user device. The method also includes determining a priority score for the ASR request based on the content metadata. The method includes caching the ASR request in a pre-processing backlog of pending ASR requests each having a corresponding priority score. The pending ASR requests in the pre-processing backlog are ranked in order of the priority scores. The method also includes providing, from the pre-processing backlog, one or more of the pending ASR requests to a backend-side ASR module. Pending ASR requests associated with higher priority scores are processed before pending ASR requests associated with lower priority scores.SELECTED DRAWING: Figure 1
Owner:GOOGLE LLC

Multimodal semantic analysis and image search

A system and method are provided for identifying and retrieving semantically similar images from a database. A visual language model is used to perform semantic analysis of the input query (406) and identify semantic concepts related to the input query. For the identified semantic concepts, a preliminary set of images is retrieved from the database (804). Related concepts are extracted from the images by comparing the images to a predefined label space using a tokenizer and identifying related concepts (806). A ranked list of related concepts is generated based on their frequency of occurrence in the set (810). By combining the input query with a specific related concept, a specific related concept is selected from the ranked list, and the preliminary set of images is narrowed down (812). Further semantic analysis is performed iteratively until a threshold condition is met (814), and an additional set of images semantically similar to the combined input query and the selection of a specific related concept is retrieved.
Owner:NEC LABORATORIES AMERICA INC

system

Provide a system. 【Solution means】 Means for receiving voice information and obtaining data, Voice recognition means for converting the data into character information, Analysis means for analyzing the converted character information and extracting a procedure flow, Generating means for automatically generating a procedure flow diagram based on the analysis result, Means for presenting operational issues and countermeasures based on the procedure flow diagram, Means for automatically generating materials based on the procedure flow diagram and countermeasures, Robot control means for updating an operation procedure through the generating means, A system including the above.
Owner:SOFTBANK GROUP CORP

system

The system according to this embodiment aims to generate and provide music to the user based on the atmosphere and emotions of a photograph. [Solution] The system according to the embodiment comprises a reception unit, an analysis unit, a generation unit, and a provision unit. The reception unit receives photos uploaded by the user. The analysis unit detects the characteristics of the photos received by the reception unit and analyzes the atmosphere, color tone, and emotion. The generation unit generates music based on the results analyzed by the analysis unit. The provision unit provides the music generated by the generation unit.
Owner:SOFTBANK GROUP CORP

Audio processing method and apparatus

An audio processing method and an electronic apparatus are provided. The audio processing method is applied to a conference system, and the conference system includes at least one audio capturing device. The audio processing method includes: receiving at least one segment of audio captured by the at least one audio capturing device; determining voices of a plurality of targets in the at least one segment of audio; and performing voice recognition on a voice of each of the plurality of targets, to obtain semantics corresponding to the voice of each target. Voice recognition is separately performed on voices of different targets, thereby improving accuracy of voice recognition.
Owner:HUAWEI TECH CO LTD +1

Internet news information capturing and audio abstract broadcasting method

The invention relates to the technical field of information capturing, in particular to an internet news information capturing and audio abstract broadcasting method. A cloud server firstly captures news data from each Internet information source through a web crawler tool, generates a corresponding news abstract audio, and obtains a corresponding content tag; the cloud server calculates a preference matching value of the current user for each content tag based on listening data of the current user in past first preset duration, and the higher the preference matching value is, the higher the preference degree of the current user for the news abstract audio of the corresponding content tag is; then, based on the content labels of the news summary audios and the preference matching values of the current user for all the content labels, a listening sequence of the current user for listening to the news summary audios is determined, and a subsequent user terminal installs the listening sequence to obtain and play all the news summary audios from the cloud server; therefore, the news audio which the user is interested in can be accurately identified and pushed.
Owner:HUNAN MANGO INTELLIGENT MEDIA TECH DEV CO LTD

Audio file metadata identification and completion method and system, storage medium and equipment

The invention relates to the technical field of multimedia data management, and discloses an audio file metadata identification and completion method and system, a storage medium and equipment, and the method comprises the steps: scanning an audio file, reading an embedded metadata tag of the audio file, and judging the metadata missing degree; extracting audio fingerprint characteristics of the audio file with the missing metadata, and analyzing a file name to obtain candidate metadata information; initiating a query request to a metadata database based on at least one of the audio fingerprint features and the candidate metadata information, and obtaining a returned candidate metadata set; carrying out credibility evaluation and data fusion on the candidate metadata sets of different sources, and determining an optimal metadata set; and writing the optimal metadata set into the metadata tag of the corresponding audio file, and synchronously updating the index of the local music library, thereby automatically complementing the audio file lacking metadata through the method, and improving the user experience and the product value of the audio playing equipment.
Owner:LINKPLAY TECHNOLOGY INC NANJING

Audio recording evidence generation and verification method and system, electronic device and storage medium

The application discloses a kind of generation and verification method, system, electronic equipment and storage medium of recording evidence, which comprises the following steps: obtaining original audio data from recording device;Original audio data is processed to embed digital watermark in original audio data, and target audio data is obtained;According to target audio data, generate main evidence file and secondary evidence file;According to main evidence file, evidence publicity is carried out;According to secondary evidence file, whether main evidence file is legal is verified.The present application can generate main evidence and secondary evidence according to audio data embedded with digital watermark, carry out evidence publicity through main evidence, and verify the legality of audio data through digital watermark, thereby improving the reliability of main evidence;When main evidence is questioned or may be copied, forged, tampered, secondary evidence can be used to verify main evidence to determine whether main evidence is legal, thereby improving the legality and reliability of recording evidence.
Owner:JINGCHEN SEMICON SHENZHEN CO LTD

Incomplete multi-mode sentiment analysis method based on retrieval enhanced mode completion and adaptive prompt learning

The invention discloses an incomplete multi-mode sentiment analysis method based on retrieval enhanced mode completion and adaptive prompt learning. The method comprises the following steps: obtaining to-be-processed text data, to-be-processed audio data and to-be-processed video data; and inputting the to-be-processed text data, the to-be-processed audio data and the to-be-processed video data into a trained sentiment analysis model based on retrieval enhanced mode completion and adaptive prompt learning to obtain a final sentiment classification result. The method has the characteristic of high accuracy.
Owner:GUANGDONG UNIV OF TECH

Navigation scene broadcast request management method and device, equipment and storage medium

The invention provides a broadcast request management method and device for a navigation scene, equipment and a storage medium, and relates to the technical field of artificial intelligence, in particular to the technical field of natural language processing and the like. The method comprises the following steps: receiving a first navigation broadcast request of a user, and obtaining a navigation audio file corresponding to the first navigation broadcast request; converting the file format of the navigation audio file into a text form to obtain a converted text; comparing an original text corresponding to the first navigation broadcast request with the converted text to obtain a distinguishing text; and generating a new rewriting rule according to the distinguishing text, and writing the new rewriting rule into a user-defined rule base, so that a subsequent navigation broadcast request is normalized according to rules in the user-defined rule base, and the user-defined rule base is used for storing and managing the generated rewriting rule. Through automatic navigation broadcast request processing and rewriting rule generation, dependence on manual intervention and a third-party software development kit is reduced.
Owner:BEIJING BAIDU NETCOM SCI & TECH CO LTD

Coordination of parallel processing of audio queries across multiple devices

The present disclosure is generally related to a data processing system to coordinate parallel processing of audio queries across multiple devices. A data processing system can receive an audio input signal detected the display device and parse the audio input signal to identify an entity. The data processing system can transmit a query command to the display device to cause a multimedia content application to perform a search for the entity. The data processing system can access at least one of an address database and a multimedia content provider to identify a reference address for the entity. The data processing system can provide the reference address for the entity to cause the display device to present a content selection interface. The content selection interface can include an element for the reference address, prior to completion of the search for the entity performed by the multimedia content application.
Owner:GOOGLE LLC

Electric toothbrush outputting melody by vibration and method of operation thereof

One embodiment of the present disclosure provides an electric toothbrush including: a toothbrush head; a vibration generating portion that generates vibration to the toothbrush head; and a processor that determines sound source data or vibration data corresponding to the sound source data as an object of output, and controls an operation of the vibration generating portion based on the vibration data corresponding to the sound source data.
Owner:LG HOUSEHOLD & HEALTH CARE LTD

Speech recognition based sentence correction method and apparatus, device, and storage medium

The present application relates to the technical field of sentence correction, and particularly relates to a sentence correction method and device based on speech recognition, equipment and a storage medium, wherein a comprehensive bundle search technology is adopted to recognize speech data, and a first candidate sentence with the highest bundle search score and a plurality of candidate sentences are extracted in terms of semantic features and pinyin features, and feature fusion is performed; the first candidate sentence is corrected by using the fused features after feature fusion, thereby reducing the negative influence caused by pronunciation problems of users and multi-pronunciation character problems of texts, and improving the accuracy and efficiency of sentence correction.
Owner:GUANGZHOU YANLI NETWORK TECH CO LTD +3

An internet news information capturing and audio abstract broadcasting method

The present application relates to the technical field of information capture, in particular to an internet news information capture and audio summary broadcasting method; a cloud server first captures news data from each internet information source through a network crawler tool, generates corresponding news summary audio, and obtains corresponding content tags; the cloud server calculates a preference matching value of the current user for each content tag based on the listening data of the current user in the past first preset time length, and the higher the preference matching value, the higher the preference degree of the current user for the news summary audio of the corresponding content tag; then, the listening sequence of the current user for listening to the news summary audio is determined based on the content tags of the news summary audio and the preference matching value of the current user for each content tag, and the subsequent user terminal installs the listening sequence to obtain and play each news summary audio from the cloud server, thereby realizing accurate identification and pushing of news audio of interest to the user.
Owner:HUNAN MANGO INTELLIGENT MEDIA TECH DEV CO LTD

Intelligent audio generation system, method, device and medium supporting multi-entity interaction

The application discloses a kind of intelligent audio generation system, method, equipment and medium supporting multi-entity interaction, it is related to audio equipment field.System includes: multiple NFC cards with unique identifier and being defined as story element attribute;Intelligent audio device is used to detect and read the unique identifier of multiple NFC cards in induction area by NFC card reading module, then combination generates combination request instruction and sends to cloud server;Cloud server is used to receive combination request instruction, according to multiple unique identifiers in story logic rule base matching determines corresponding story line logic, retrieves audio segment from audio segment library and generates ordered audio segment playing sequence and issues;Intelligent audio device receives audio segment playing sequence and plays audio content therein in order.The application can generate differentiated audio content with logical association, break the single solidified interaction limit, and enrich the audio content interaction experience of user.
Owner:SHENZHEN WELLDY TECH CO LTD

Audio processing method and device, equipment, storage medium and program product

The embodiment of the invention relates to an audio processing method and device, equipment, a storage medium and a program product. The method comprises the following steps: in response to a resource transfer message sent by a server, querying whether a target audio file corresponding to the resource transfer message exists in a local storage space; if the resource transfer message does not exist and the resource transfer message contains access address information corresponding to the target audio file, downloading the target audio file based on the access address information; if downloading fails, synthesizing a target audio file based on the resource transfer amount in the resource transfer message and the unit audio file in the local storage space; and based on the target audio file, performing voice broadcast processing after resource transfer is completed. Thus, the target audio file is obtained in combination with multiple modes, the broadcast delay is reduced, and the broadcast success rate and the broadcast efficiency are improved.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

Display device

A display device according to an embodiment of the present invention may include: a microphone; the controller is used for receiving a microphone starting command for controlling the microphone to be started from the remote control device; and a display that displays music information retrieved based on the audio recorded by the microphone from a time point at which the microphone turn-on command is received.
Owner:LG ELECTRONICS INC