Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

56results about "Audio data browsing/visualisation" patented technology

Managing access to digital assets

Managing access to digital content at a server system in communication with relevant gaming devices, including: receiving a user account identifier and performance information from a relevant gaming device; accessing a collectible database storing collectible records; comparing the received performance information with the performance information in the collectible records; accessing a user account database that stores user account records; adding the retrieved collectible identifier to the identified user account record; retrieving a music-related asset identifier from the identified collectible record; and sending a confirmation to the relevant gaming device that indicates the collectible asset has been collected and indicates the retrieved music-related asset identifier.
Owner:SONY GROUP CORP +1

Media content playback methods, devices, storage media and program products

This disclosure provides a media content playback method, device, storage medium, and program product. The media content playback method involves displaying a candidate media content list on the playback interface of the current media content when a preset condition is met during media content switching in a media content player. The candidate media content list displays multiple candidate media contents. In response to a selection instruction for a target media content in the candidate media content list, the target media content is determined as the media content to be played. This disclosure provides a candidate media content list for users to quickly browse multiple candidate media contents when the media content switching meets the preset condition, allowing users to conveniently and quickly select the desired target media content, avoiding frequent media content switching operations, and improving the efficiency of media content search and selection.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

Electronic apparatus and control method therefor

An electronic apparatus includes: a microphone; a camera; a display; a speaker; and a processor operatively connected to the camera, the display, and the speaker, where the processor displays, on the display, a live view image obtained through the camera, when a camera application is executed; obtain a characteristic text for the live view image based on ae piece of object information included in the live view image, displays the characteristic text together with the live view image, identifies a sound source including metadata which matches the characteristic text, from among sound sources; displays information about the sound source together with the live view image, and when a user manipulation for capturing a moving image is input, outputs a first sound source from among the sound source through the speaker and capture a moving image through the microphone and the camera.
Owner:SAMSUNG ELECTRONICS CO LTD

Quantification of music genre similarity

Techniques are disclosed for generating an attribute embedding for a catalog of items. A computer system can determine pairwise relationships between attributes of items in a digital catalog. The computer system can use the pairwise relationships to generate a graph including attribute nodes. Each attribute node can be related to each other attribute node of the graph according to the pairwise relationships. The computer system can also generate an attribute embedding based on the graph. The computer system can then generate a collection of items from the items in the digital catalog using the attribute embedding.
Owner:AMAZON TECH INC

System generated damage analysis using scene templates

Systems, methods, and computer-readable media, are disclosed in which a variety of data describing the condition of an object can be obtained and probabilistic likelihoods of causes and / or value of damages to the object can be calculated. In a variety of embodiments, data obtained from third-party systems can be utilized in these calculations. Any of a number of machine classifiers can be utilized to generate the probabilistic likelihoods and confidence metrics in the calculated liabilities. A variety of user interfaces for efficiently obtaining and visualizing the object, the surrounding geographic conditions, and / or the probabilistic likelihoods can further be utilized as appropriate. User interfaces can include a scene sketch tool application program interface and / or a liability tool user interface and various data sources can be used to dynamically generate diagrams of object damage and / or the environment in which the object was damaged.
Owner:ALLSTATE INSURANCE COMPANY

Information processing method and electronic equipment

The invention provides an information processing method and electronic equipment, and the method comprises the steps: obtaining the historical behavior information of a user for audio and video contents; based on the historical behavior information, generating a target image corresponding to a first target song and a target playing entrance corresponding to a second target song; and in response to a publishing instruction of a user, publishing publishing information including the target image and the target playing entrance. According to the method and the device, the historical behavior information of the user for the audio and video can be converted into the personalized and visual target image and the target playing entrance so as to be published, so that the published target image can be automatically generated according to the behavior information of the user for the audio and video in the past, the image generation function in a community is enriched, and the user experience is improved. And the publishing process is simplified, and the use threshold of the user is reduced, so that the publishing willingness of the user in the community is improved, and the content richness in the community is also improved. And meanwhile, the personalized degree of content generation is also improved.
Owner:HANGZHOU NETEASE CLOUD MUSIC TECH CO LTD

Audio playing method, device, medium and computing device

Embodiments of the present disclosure provide an audio playing method, device, medium and computing device, the method comprising: in response to a playing operation of a first audio, determining user information corresponding to a user triggering the playing operation; in response to determining that the user does not have complete audio playing permission of the first audio according to the user information, and the user satisfies complete audio audition condition of the first audio, playing the complete first audio. In the present disclosure, when the user satisfies the complete audio audition condition, the audio to which the playing operation of the user is directed is played completely, so that the user can listen to the played audio completely, thereby providing a more objective decision basis for the user to obtain complete audio playing permission, increasing user stickiness and improving user experience.
Owner:HANGZHOU NETEASE CLOUD MUSIC TECH CO LTD

Information processing apparatus, method for processing information, program, and information processing system

To improve convenience during positioning.SOLUTION: There is provided an information processing apparatus including: an acquisition unit for acquiring an image captured by an imaging device; a display control unit for causing a specific display object to be displayed on the image in association with a specific object captured in the image; and a positioning unit for performing measuring of the position of the imaging device on the basis of the image in which the specific object is captured.SELECTED DRAWING: Figure 4
Owner:NEC CORP

Media content playing method and device, and storage medium and program product

Provided in the embodiments of the present disclosure are a media content playing method and device, and a storage medium and a program product. The media content playing method comprises: when it is detected that the switching of media content in a media content player meets a preset condition, displaying a candidate media content list in a playback interface for current media content, wherein a plurality of pieces of candidate media content are displayed in the candidate media content list; and in response to a selection instruction for target media content in the candidate media content list, determining the target media content as media content to be played.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

Playing interface processing method and device, equipment, storage medium and program product

The invention provides a playing interface processing method and device, equipment, a storage medium and a program product, and relates to the technical field of computers. The method comprises the following steps: acquiring template trigger information, wherein the template trigger information comprises first target information acquired from a currently played audio and / or second target information input by a user; according to the template triggering information, detecting whether a target template corresponding to the template triggering information exists in a plurality of preset templates or not; responding to a target template corresponding to the template triggering information in the plurality of preset templates, and generating a first playing interface based on the target template; and in response to an interaction operation for the first playing interface, updating the first playing interface to obtain a second playing interface. The user operation fatigue can be reduced, and the user experience is improved.
Owner:HANGZHOU NETEASE CLOUD MUSIC TECH CO LTD

Music storage and extraction system based on big data

The invention discloses a music storage and extraction system based on big data, and belongs to the technical field of music data processing. The method actively supervises, processes and analyzes the negative effects on the storage device when different users download music storage, and supervises and digitalizes the targeted recommendation implementation effect of the music storage extraction scheme when different users download music storage. Performing supervision and digital processing on subsequent continuous adoption data of a music storage extraction scheme when different users download music storage, performing integration processing on recommendation implementation effect data and recommendation continuous effect data in different aspects in the earlier stage, and performing dynamic optimization management on the music storage extraction scheme; the method is used for solving the technical problems that in an existing scheme, when different users download music storage, the storage of a user side negatively affects supervision and analysis prompting effects are poor, and local supervision implementation and autonomous optimization management effects of different music storage extraction schemes are poor.
Owner:HEIHE UNIV

Method for generating diagnostic information of facility on basis of acoustic data

A method for generating diagnostic information of a facility on the basis of acoustic data according to the present invention is a method in which a facility diagnostic device generates diagnostic information of a facility on the basis of acoustic data, and comprises the steps of: acquiring raw data related to acoustic data generated from the facility; generating a spectrogram image on the basis of the raw data, wherein the spectrogram image is image information including a first axis in a time domain, a second axis in a frequency domain, and a color indicating the strength; generating first conversion data by normalizing the spectrogram image with respect to a frequency domain; generating second conversion data by axially synthesizing the first conversion data with respect to a synthesis time section; and generating third conversion data by extracting the second conversion data with respect to a frequency band of interest, wherein in the step of generating the second conversion data, second configuration information related to the synthesis time section is configured, and the second configuration information is determined according to characteristic information of the facility.
Owner:MOVIC LAB INC

A laparoscopic surgery video retrieval and visualization method and system

The present application relates to a kind of laparoscopic surgery video retrieval and visualization method and system, which is achieved by "video analysis-data organization-multi-mode retrieval" efficient surgery video retrieval.In the video analysis stage, the deep learning model is combined with the hidden Markov model of medical knowledge constraint, automatically analyzes the video content and divides the standardization process unit, reduces the consumption of computing resources and training time.In the data organization stage, the feature entity bidirectional mapping is proposed, which converts the unstructured video data into a hierarchical data structure with medical semantics, facilitating storage and retrieval.In the multi-mode retrieval stage, three retrieval modes of time axis orientation, instrument orientation and process unit orientation are designed, combined with multi-track display and medical semantic visualization coding to realize dynamic visualization.Through the interactive feedback of multi-mode retrieval method and time sequence feature dynamic visualization, a professional retrieval experience that meets the training and research needs of medical personnel is provided.
Owner:HUNAN UNIV

Music style identification and classification method and system

The invention relates to the technical field of music information retrieval, and discloses a music style identification and classification method and system, and the method comprises the steps: collecting target audio and multi-style reference audio data; performing music acoustic feature extraction and combination on the audio data, and constructing a target music acoustic feature vector; based on a preset standard feature vector matrix, searching and matching a most suitable target music style classifier type through an optimization algorithm; performing classification calculation on the target feature vector by using a matched classifier to obtain target music style classification data containing a predicted style label and a probability thereof; and finally, carrying out credibility evaluation on the classification result and generating a report. The system comprises an audio signal processing module, a feature processing and matching module and a music style classification execution module. According to the method, the robustness and discrimination capability of the feature layer are improved, and a reliable data foundation is laid for high-precision music style classification.
Owner:MINNAN NORMAL UNIV

Data processing method and device, equipment, medium and product

The invention relates to the technical field of computers, and discloses a data processing method and device, equipment, a medium and a product, and the method comprises the steps: determining a plurality of candidate songs according to the song listening behavior data of a user within a preset time period; according to the cover image of each candidate song, determining a color category corresponding to each candidate song; selecting a target color category from the color categories corresponding to the candidate songs; determining a target song contained in the target color category; and generating a target page corresponding to the target color category according to the cover image of each target song. The song listening behavior preference of the user and the color of the cover image are aggregated, the user color portrait with strong visual impact can be generated, the visual effect is good, and the displayed content is more visual.
Owner:HANGZHOU NETEASE CLOUD MUSIC TECH CO LTD

Lyrics labeling method, device and equipment and storage medium

The present disclosure provides a lyrics labeling method and device, equipment and storage medium, relates to the technical field of data processing, and particularly relates to the field of intelligent recommendation and multimedia application. The specific implementation scheme is as follows: a target background music and a corresponding target style thereof are determined; initial labeling information corresponding to preset lyrics is generated based on the target background music and the target style; target labeling information of the preset lyrics is determined based on the initial labeling information corresponding to the preset lyrics; and target lyrics of the target background music are generated based on the preset lyrics and the corresponding target labeling information. According to the technical scheme of the present disclosure, the accuracy and efficiency of labeling can be ensured.
Owner:BEIJING BAIDU NETCOM SCI & TECH CO LTD

Multimodal abstract generation method and system based on keyword prediction enhancement

The present application relates to the technical field of multi-modal abstract generation, in particular to a multi-modal abstract generation method and system based on keyword prediction enhancement, comprising the following steps: 1) obtaining multi-modal abstract data, wherein the multi-modal abstract data comprises text, images corresponding to the text, and text keywords, the text keywords are obtained by calculating entities in the intersection of the original text and the reference abstract, and provide reliable training data for subsequent visual key information extraction; the present application effectively uses a small model as an auxiliary model to extract keywords to assist a large model in performing a multi-modal abstract generation task, more effectively guides the large model to generate a multi-modal abstract by using this framework, can lock core information in the text, thereby obtaining more robust and stronger fact consistency results, effectively solves problems such as focus shift, visual redundancy, and semantic gap between different modalities, and improves the quality of abstract generation.
Owner:FUZHOU LIANCHUANG ZHIYUN INFORMATION TECH CO LTD

A song customization method and device, electronic equipment and storage medium

Embodiments of the present application provide a song customization method and device, electronic equipment and storage medium, which realize customization and playing of personalized songs for different users and improve user experience. The song customization method comprises: obtaining associated account information of a client and a corresponding customized song push time; comparing the customized song push time with a current date, and if it is determined that the comparison result is consistent, sending customized song information corresponding to the associated account information to the client for display by the client, wherein the customized song is synthesized according to the associated account information and an original song.
Owner:HANGZHOU NETEASE CLOUD MUSIC TECH CO LTD

A song claiming method, a claiming device, and a storage medium

The embodiment of the application discloses a song claiming method and device and a storage medium, and is used for the technical field of song claiming. The method comprises the following steps: obtaining target song creator information added by an already-registered song creator when uploading a song on a music platform, the target song creator information being information of a target song creator who has not registered on the music platform, and the target song creator information comprising a target song creator identifier; verifying the identity of the target song creator according to the target song creator information; if the identity verification of the target song creator is passed, triggering a registration application of the target song creator to register on the music platform, and guiding the target song creator to claim a song corresponding to the target song creator identifier in the music platform; and when the registration application is passed, determining that the target song creator successfully claims the song. The method can effectively reduce the time for claiming a song when a song creator has not registered on the music platform.
Owner:TENCENT MUSIC ENTERTAINMENT TECH (SHENZHEN) CO LTD

Method of enabling digital music content to be downloaded to and used on a portable wireless computing device

The invention enables digital music content to be downloaded to and used on a portable wireless computing device. An application running on the wireless device has been automatically adapted to parameters associated with the wireless device without end-user input (e.g. the application has been configured in dependence on the device OS and firmware, related bugs, screen size, pixel number, security models, connection handling, memory etc., This application enables an end-user to browse and search music content on a remote server using a wireless network; to download music content from that remote server using the wireless network and to playback and manage that downloaded music content. The application also includes a digital rights management system that enables unlimited legal downloads of different music tracks to the device and also enables any of those tracks stored on the device to be played so long as a subscription service has not terminated.
Owner:TIKTOK PTE LTD

Music-driven image generation method and system based on cross-modal alignment and computer device

The invention relates to the technical field of artistic artificial intelligence, in particular to a music-driven image generation method and system based on cross-modal alignment and a computer device, and the method comprises the steps: S1, obtaining a music audio clip and a text title and sentiment classification corresponding to the music audio clip; s2, performing feature extraction and normalization processing on the music audio clips to obtain audio embedding features; s3, the audio embedding features are mapped to a pseudo text embedding space, and audio pseudo text emotion features are obtained through a cross-modal alignment network; s4, fusing the audio pseudo-text emotion features with the text title features to obtain joint condition features; and S5, inputting the joint condition features into a diffusion model, and generating an image consistent with the music theme and the emotion features in the music audio clip. According to the method, the modal difference between the music and the image is effectively reduced, the semantic consistency and emotion expression of the generated result are improved, and the method is suitable for scenes such as music visualization and intelligent creation.
Owner:HANGZHOU DIANZI UNIV

AI music creation system based on multiple modes

The invention discloses an AI music creation system based on multiple modes, and the system comprises a user control layer which is used for receiving the input of a user, and transmitting a control signal to a controllability engine and an accidental engine; and the controllability engine is connected with a rule constraint domain, and the rule constraint domain comprises a music theory rule base and a culture compatible filter and is used for performing rule constraint on music generation based on the music theory rule base and the culture compatible filter. According to the AI music creation system based on multiple modes, through a closed-loop feedback mechanism of the artistry evaluation module and the strategy adjustment module, a music generation strategy can be dynamically adjusted in real time, the mechanization and monotonicity problems of music generation are effectively avoided, the artistry of music works and the user satisfaction degree are remarkably improved, and then the user experience is improved. According to the method, the dimension collaborative optimizer is introduced, consistency analysis is carried out on music dimensions such as chord proceeding, melody development and rhythm design, dimension conflicts are effectively prevented, and the generated music works are more harmonious in structure.
Owner:XINJIANG UNIVERSITY

An audio-visual cross-modal interaction design method for flat embroidery

The application discloses an audio-visual cross-modal interaction design method for flat embroidery needle methods, which comprises the following steps: collecting different flat embroidery needle method graphs, vectorizing and sample augmenting each graph; using a K-means clustering algorithm to cluster graphs with similar features into clusters, screening representative graphs of each cluster to construct a visual modal information library; screening perceptual vocabulary capable of representing the visual modal information of the graphs, including attribute layer, perception layer and association layer vocabulary; scoring the matching degree of perceptual vocabulary at each level and each representative graph to construct the mapping relationship of "visual modal information-perceptual vocabulary" of each representative graph; finding out music bars related thereto, scoring the matching degree of each music bar and the representative graph, thereby constructing the "visual modal information-music bar" mapping relationship based on perceptual vocabulary; obtaining an audio-visual cross-modal information parameter table containing each representative graph and the corresponding music bar parameters based on the mapping relationship, adjusting the music bar parameters to output new music bars, evaluating and optimizing the matching degree of each representative graph and the corresponding new music bar, and outputting the optimized audio-visual cross-modal information library. The application produces a new interactive experience mode through scientific description of different flat embroidery needle methods to meet the needs of people for all-around perception and in-depth experience of embroidery technology.
Owner:NANJING UNIV OF AERONAUTICS & ASTRONAUTICS

A tinnitus sound therapy music playing system with tinnitus detection function

The application discloses a tinnitus sound therapy music playing system with a tinnitus detection function, relates to the technical field of tinnitus sound therapy music treatment, and comprises a tinnitus data acquisition unit, a tinnitus detection unit, a music playing library and a comprehensive matching unit. The tinnitus data acquisition unit, the tinnitus detection unit, the music playing library and the comprehensive matching unit are arranged to construct a perfect tinnitus sound therapy music playing system. The tinnitus data acquisition unit is used to collect tinnitus sound data of simulated patients to generate corresponding data transmission channels. Meanwhile, a positioning encryption module is installed to accurately position the collected simulated tinnitus data, tinnitus models and generated music data, and to simultaneously perform encryption processing, so that the required viewing links for accurate positioning during visualization are facilitated, the overall treatment cycle is shortened, and the intelligent level of tinnitus sound therapy data management is improved through internet cloud management and control.
Owner:PEKING UNION MEDICAL COLLEGE HOSPITAL

Audio recommendation system

An audio recommendation system adds a “recents” tab in a memory accessible to a messaging system including any audio tracks (songs or sounds) encountered by applications in order to allow quick access to any recently played songs or sounds provided in a message to / from another user or encountered during activities of the user. The displayed “recents” tab enables the user to revisit songs or sounds that the user may wish to use later in another message or to include in the user's music playlist. The “recents” tab also enables the user to browse an audio history in received messages without needing to explicitly save the music or sounds upon receipt. The system determines the source of the encountered sound or song and stores the source of the sound or song in a playlist associated with the recents tab with identifying information for the sound or song.
Owner:SNAP INC

Lyric display method and device, storage medium and program product

The invention discloses a lyric display method and device, equipment and a storage medium, and relates to the technical field of computers. The method comprises the following steps: acquiring a lyric file of a first song; according to the lyric file of the first song, obtaining a lyric text of the first song, the lyric text of the first song including a text of a first language type; in the playing process of the first song, generating a lyric translation text of the first song according to the lyric text of the first song through a first artificial intelligence AI model, the lyric translation text of the first song being a text of a second language type; and synchronously displaying the lyric translation text of the first song according to the playing progress of the first song. According to the method, real-time translation of the lyric text is realized locally on the terminal equipment, and storage resources and computing resources of the server do not need to be called, so that the resources and the flow of the server are saved.
Owner:GUANGZHOU KUGOU COMP TECH CO LTD

Information processing device, information processing method, and program

An information processing apparatus includes a control unit configured to perform control to display related content that is associated with music piece content and includes a plurality of pages on which a character or an image is displayed, and display a page guide indicating a page being displayed in a whole of the related content.
Owner:SONY GROUP CORP