Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

65 results about "Audio watermark" patented technology

An audio watermark is a unique electronic identifier embedded in an audio signal, typically used to identify ownership of copyright. It is similar to a watermark on a photograph. Watermarking is the process of embedding information into a signal (e.g. audio, video or pictures) in a way that is difficult to remove. If the signal is copied, then the information is also carried in the copy. Watermarking has become increasingly important to enable copyright protection and ownership verification.

Audio watermark generation method and device and computer storage medium

The invention provides an audio watermark generation method, audio watermark generation equipment and a computer storage medium. The audio watermark generation method comprises the following steps: acquiring target audio data; extracting amplitude spectrum features and phase spectrum features of the target audio data; acquiring watermark adding frame information of the amplitude spectrum features; selecting one watermark information code from a watermark codebook library, and performing frame-level copying according to the frame number of the amplitude spectrum characteristics to obtain first watermark data; performing frame selection on the first watermark data according to the watermark adding frame information to obtain second watermark data; fusing the amplitude spectrum feature with the second watermark data to obtain a watermark amplitude spectrum feature; and fusing the watermark amplitude spectrum feature and the phase spectrum feature to obtain watermark audio data. According to the audio watermark generation method, the watermark adding selector is designed, the frame number and the frame number needing to be added with the watermark are selected through the frame selection logic, and it is guaranteed that after the audio watermark is generated, the audio listening feeling is not affected.
Owner:IFLYTEK CO LTD

Audio watermark processing method and device, equipment and medium

The invention relates to the technical field of artificial intelligence and audio processing, can be applied to the field of intelligent medical treatment and financial science and technology, and discloses an audio watermark processing method and device, equipment and a medium, and the method comprises the steps: carrying out the coding of continuously inputted audio blocks, and obtaining an audio feature vector; acquiring a digital watermark, and performing time modulation on the digital watermark according to the audio lengths of the blocks by using a multi-layer perceptron to obtain a watermark implicit vector; embedding the watermark implicit vector into the audio feature vector to obtain a target audio vector; performing quantization processing on the target audio vector to generate a target audio spectrum embedded with a digital watermark; and decoding the target audio frequency spectrum to obtain a target audio. According to the method, the continuously input audio blocks are coded, the method can adapt to a streaming scene, the audio data are processed segment by segment, the watermark embedding requirement of the continuous audio stream is met, the audio can be compressed, the data volume is effectively reduced, the transmission time consumption is reduced, and the low-delay communication requirement is met.
Owner:PING AN TECH (SHENZHEN) CO LTD

Audio watermark embedding method, audio watermark extracting method and model training method

The invention discloses an audio watermark embedding method, an audio watermark extracting method and a model training method, and belongs to the technical field of artificial intelligence. The audio watermark embedding method comprises the following steps: acquiring watermark data and an original audio clip in which a watermark needs to be embedded in original audio data; inputting the original audio clip and the watermark data into an encoder, and determining embedding coefficients of at least two time frames in the original audio clip through the encoder; based on the embedding coefficient, the watermark data is embedded into the original audio clip, a watermark-containing audio clip corresponding to the original audio clip is obtained, and the embedding coefficient of the time frame is in positive correlation with the energy of the time frame; and replacing the original audio clips in the original audio data with the audio clips containing the watermarks to obtain the audio data containing the watermarks.
Owner:VIVO MOBILE COMM CO LTD

Training method and device of reversible neural network for audio watermark processing

The invention discloses a training method and device of a reversible neural network for audio watermark processing. The method comprises the steps of obtaining a sample data set; repeatedly executing the following process until the target loss value is smaller than a preset threshold value, stopping iteration, and obtaining a target reversible neural network: adding watermark information to each piece of sample audio data in the target training set according to an encoder of the initial reversible neural network, and obtaining a plurality of pieces of first sample audio data embedded with the watermark information; performing simulation attack on the plurality of first sample audio data according to a weighted attack strategy to obtain a plurality of second sample audio data; decoding the plurality of second sample audio data according to a decoder of the initial reversible neural network to obtain restored watermark information of each second sample audio data; and determining a target loss value according to the loss values corresponding to all the sample audio data, and adjusting network parameters of the initial reversible neural network according to the target loss value.
Owner:CHINA TELECOM ARTIFICIAL INTELLIGENCE TECHNOLOGY (BEIJING) CO LTD

Synthetic speech robust watermark embedding and extracting method based on multistage residual error

The invention provides a synthetic speech robust watermark embedding and extracting method based on multistage residual errors, and relates to the technical field of audio watermarks. According to the watermark embedding and extracting method provided by the invention, the convolutional neural network is combined with a coordinate attention mechanism, so that the controllability of a watermark embedding region is improved, the watermark information can be accurately recovered in the extracting process, and the traceability and the effectiveness of the watermark are improved. An audio signal is processed in a frequency domain, and time-frequency domain conversion is carried out through short-time Fourier transform and inverse transform thereof, so that the concealment and stability of watermark embedding are ensured. Besides, a synchronous code is added in the embedding process so as to be used for quickly positioning the initial position of the watermark information, and accurate watermark extraction and time sequence recovery of anti-flip attack are facilitated.
Owner:HEFEI UNIV OF TECH

Methods and apparatus to perform audio watermarking and watermark detection and extraction

Methods and apparatus to audio watermarking and watermark detection and extracted are described herein. An example method includes receiving a media content signal, sampling the media content signal to generate samples, storing the samples in a buffer, determining a first sequence of samples in the buffer, determining a second sequence of samples in the buffer, wherein the second sequence of samples is of substantially equal length as the first sequence of samples, calculating an average of the first sequence of samples and the second sequence of samples to generate an average sequence of samples, extracting an identifier from the average sequence of samples, and storing the identifier in a tangible memory.
Owner:THE NIELSEN CO (US) LLC

Robust audio watermarking method with sound source separation resistance

The invention provides a robust audio watermarking method with sound source separation resistance. The robust audio watermarking method comprises the steps of preprocessing an input audio, and embedding and extracting watermarks based on a reversible neural network. According to the method, a structure based on a reversible neural network is adopted, the watermark information and the audio spectrum are deeply fused, the resistance of the watermark to sound source separation, compression and other signal processing attacks is effectively improved by introducing a disturbance enhancement training mechanism, and compared with a traditional time domain or frequency domain embedding method, the robustness is higher; according to the method, watermark embedding is carried out on an audio amplitude spectrum, phase information of an original audio is fully reserved, and high-quality audio reconstruction is realized in combination with a perception fidelity strategy and a window compensation mechanism; according to the method, the symmetric reversible neural network structure is designed, it is ensured that information in the embedding and extracting processes is symmetric and can be restored, watermark recovery can be completed without accessing the original audio, and the good blind detection capability is achieved.
Owner:UNIV OF ELECTRONICS SCI & TECH OF CHINA ZHONGSHAN INST

Audio watermark generation method, device and equipment and computer storage medium

The invention provides an audio watermark generation method, an audio watermark generation device, audio watermark generation equipment and a computer storage medium. The audio watermark generation method comprises the following steps: extracting a first potential vector representation of target audio data; inputting the target audio data into an audio watermark embedding network of an audio watermark generation network, adding watermark data, and obtaining watermark audio data; extracting a second potential vector representation of the watermark audio data; extracting third potential vector representation of other audio data; using the first potential vector representation, the second potential vector representation and the third potential vector representation to construct a triple loss; and the audio watermark generation network is updated by using the triple loss, and the audio watermark embedding network is used for generating the audio watermark after being updated. According to the audio watermark generation method, the audio features before and after the watermark is added are aligned in the potential space, so that the network parameters of the audio watermark embedding network are updated, and the audio watermark generation effect is improved.
Owner:IFLYTEK CO LTD

Decoding audio watermarks using time shifts

Described herein is a system for performing watermark detection using multiple time shifts to increase a resolution of watermark detection. Instead of decoding blocks of successive audio frames (e.g., 10 ms of audio) using a single watermark decoder, watermark verification can be performed by decoding overlapping frame shifts using multiple decoders in parallel, thereby increasing a chance that one of the watermark decoders will be synchronized with the embedded audio watermark. For example, watermark verification may split audio data into parallel streams and decode using two decoders (e.g., 2× shifts-per-frame), four decoders (e.g., 4× shifts-per-frame), or the like. Increasing resolution by performing overlapping detection increases an accuracy of the watermark detection without changing the embedded audio watermark.
Owner:AMAZON TECH INC

Method of inserting audio-watermark specialized for music usage and NFT and providing music source

A method of inserting an audio watermark specialized for music usage and non-fungible token (NFT) and providing a music source includes receiving a music source download request from a user who accesses a music source providing system; authenticating whether the user is a legitimate user; receiving a music source usage purpose from the user when the user is the legitimate user; inserting an audio watermark including different types of data information into the music source according to the music source usage purpose; and performing a music source download requested by the user.
Owner:KEISER INC

Hotword suppression

PendingUS20260188318A1Audio watermarkSpeech sound
A method includes adding, by a first computing device, a first audio watermark to first speech data corresponding to playback of a first utterance including a hotword used to invoke an attention of a second computing device. The method includes outputting, by the first computing device, the playback of the first utterance corresponding to the watermarked first speech data. The second computing device is configured to receive the watermarked first speech data and determine to cease processing of the watermarked first speech data.
Owner:GOOGLE LLC

Recorded media hotword trigger suppression

Methods, systems, and user devices for suppressing hotword triggering when a hotword in a recorded media is detected are disclosed, including computer programs encoded on computer storage media. In an aspect, a method includes receiving, at data processing hardware, audio data corresponding to playback of a media content item, the audio data including an audio watermark and an utterance of a command preceded by a hotword; determining, by the data processing hardware, that the received audio data includes the hotword; processing, by the data processing hardware, the audio data to: identify the audio watermark included in the audio data; and determine a corresponding bitstream of the audio watermark; and based on the determined corresponding bitstream of the audio watermark, determining, by the data processing hardware, to bypass performing the command preceded by the hotword without accessing an audio watermark database to identify a matching audio watermark.
Owner:GOOGLE LLC

Audio watermark embedding method and device and conference system

PendingCN121687081ASpeech analysisPattern recognitionWatermark robustness
The invention provides an audio watermark embedding method and device and a conference system, and relates to the technical field of digital watermarking. And obtaining an original audio collected from an environment where the terminal is located in the first time period, wherein the original audio comprises an audio played by the terminal in the first time period. The playing audio is configured to be embedded with a watermark based on the first watermark embedding parameter value by using a first watermark embedding algorithm. And carrying out watermark extraction on the original audio according to the first watermark embedding algorithm. And if the watermark extraction from the original audio fails, embedding the watermark in the audio to be played of the terminal based on the second watermark embedding parameter value by using the first watermark embedding algorithm to obtain a new played audio. Under the condition that the watermark cannot be successfully extracted from the original audio, the watermark embedding parameter value is adjusted, so that the watermark embedded in the audio played later by the terminal can adapt to the environment and is successfully extracted, the adaptive environment adjustment of the audio watermark is realized, and the robustness of the audio watermark is improved.
Owner:HUAWEI CLOUD COMPUTING TECHNOLOGIES CO LTD

Systems and methods for audio watermarking

PCT designated stageWO2026107213A1Speech analysisPattern recognitionAudio watermark
Described is a neural network based system for use in audio watermarking, comprising: a neural network based watermark generator configured for embedding, based on a learnable embedding table, a predefined watermark into an input audio signal, thereby obtaining a watermarked audio signal; and a neural network based watermark detector comprising: a detection head that is configured for determining presence or absence of a watermark in a target audio signal; and a decoding head that is configured for decoding the watermark from the target audio signal, wherein the embedding table is shared between the watermark generator and the watermark detector, such that the decoding of the watermark from the target audio signal is based on the embedding table that has been used for the embedding.
Owner:DOLBY LABORATORIES LICENSING CORP

Audio watermark generation and detection method

The invention discloses an audio watermark generation and detection method. The method comprises the following steps: acquiring an audio signal corresponding to an audio to be processed; the amplitude of the confrontation disturbance signal is smaller than that of a value set corresponding to a psychological acoustic masking threshold value of the audio signal, the psychological acoustic masking threshold value is used for quantitatively representing an added disturbance amplitude upper limit corresponding to each frequency point in the audio signal, and the confrontation disturbance signal is used for carrying watermark information of the audio to be processed; and embedding the adversarial disturbance signal into the audio signal. The technical problem that the audio watermarking technology adopted in the related technology is mostly based on a static or dominant embedding mode and is easily removed by an attacker through a signal processing method is solved.
Owner:CHINA TELECOM ARTIFICIAL INTELLIGENCE TECHNOLOGY (BEIJING) CO LTD

Robust audio watermarking method based on adaptive quantization strategy and feature classification

The invention discloses a robust audio watermarking method based on an adaptive quantization strategy and feature classification, and belongs to the technical field of digital audio copyright protection. The method comprises the following steps: framing an audio signal, extracting a logarithmic mean feature (DWT-CLM) of discrete wavelet transform, and combining a zero-crossing rate, a variance and energy to form a frame feature vector; a Sigmoid classifier is used to discriminate frame characteristics, and a fixed or variable quantization step size is adaptively selected to embed watermark bits into approximate components; and during extraction, the watermark is accurately extracted through the same feature analysis and classifier discrimination recovery quantization mode. Experiments show that the algorithm has better inaudible property and robustness, can effectively resist attacks such as MP3 compression, resampling, low-pass filtering and re-recording, and is suitable for digital audio copyright protection scenes.
Owner:XINYANG NORMAL UNIVERSITY

Audio watermark embedding and extracting method for resisting desynchronization attack

The invention relates to an audio watermark embedding and extracting method for resisting a desynchronization attack, which comprises the following steps of: bearing an embedded bit by using three sections of intermediate-frequency average energy at an embedded end, implementing a triangular modulation strategy on the three sections of average energy according to the embedded bit, and introducing buffer compensation at the embedded end; uniformly scaling all frequency domain coefficients according to a proportion for the buffer expansion interval by taking a segment as a unit, carrying out inverse transformation on the whole frame to obtain frames containing watermarks, and splicing the frames containing the watermarks into audio containing the watermarks; judgment is only carried out on a target frequency band at an extraction side, clipping detection is carried out through symmetric consistency discrimination, rapid and reliable resynchronization in a clipping scene is realized in combination with double-end sliding window search, and blind detection is realized on the premise of not depending on original audio. Compared with the prior art, the method has the advantages that the balance among the watermark capacity, the robustness and the hearing feeling during desynchronization attack resistance is realized, the robustness is improved, and the like.
Owner:SHANGHAI UNIV

An audio watermark adding, analyzing method, device and medium

Embodiments of the present application disclose an audio watermark adding method, comprising: a playing terminal acquiring a first audio in real time; the playing terminal embedding an audio watermark in the first audio, the audio watermark being associated with the playing terminal; and the playing terminal playing the first audio with the embedded audio watermark. Embodiments of the present application also provide an audio watermark analysis method and device and a medium, in a scenario of playing an audio in real time, a playing terminal adds an audio watermark in an audio stream in real time, so that a later device can determine the playing terminal according to the audio watermark when analyzing the watermark, and tracing is facilitated after the first audio is transcribed.
Owner:HUAWEI TECH CO LTD

Method for processing an audio watermark and audio watermark generating device

This invention provides a method for processing audio watermarks and an apparatus for generating audio watermarks. The insertion position of a reference code in an initial watermark sequence is determined based on the signal power of the main audio signal to generate an extended watermark sequence. The main audio signal and the extended watermark sequence are then synthesized to generate an audio signal with an embedded watermark. This overcomes noise interference.
Owner:ACER INC

Method and apparatus for training reversible neural network for audio watermarking processing

The application discloses a method and device for training a reversible neural network for audio watermark processing. The method comprises: obtaining a sample data set; repeatedly performing the following process until the target loss value is less than the preset threshold, stopping iteration, and obtaining a target reversible neural network: adding watermark information to each sample audio data in the target training set according to the encoder of the initial reversible neural network to obtain a plurality of first sample audio data embedded with watermark information; performing simulated attacks on the plurality of first sample audio data according to a weighted attack strategy to obtain a plurality of second sample audio data; decoding the plurality of second sample audio data according to the decoder of the initial reversible neural network to obtain restored watermark information of each second sample audio data; determining the target loss value according to the loss value corresponding to all sample audio data, and adjusting the network parameters of the initial reversible neural network according to the target loss value.
Owner:CHINA TELECOM ARTIFICIAL INTELLIGENCE TECHNOLOGY (BEIJING) CO LTD

Timely addition of human-perceptible audio to mask an audio watermark

A method and system for adding overtly human-perceptible supplemental audio content into a media stream to help mask audio effects of an audio watermark in the media stream. A method involves receiving a media stream that defines a sequence of audio content presentable by a content presentation device, modifying the media stream to produce a modified media stream that defines the sequence of audio content, and outputting the modified media stream for presentation by the content presentation device. The modified media stream includes an audio watermark that is machine-detectable to trigger an interactive event. Further, the act of modifying the media stream involves adding into the media stream supplemental audio content coincident with the audio watermark, to help mask the audio watermark in the modified media stream during presentation of the modified media stream by the content presentation device.
Owner:THE NIELSEN CO (US) LLC

Audio watermarking for people monitoring

Disclosed example people monitoring methods include detecting a first watermark in a first audio signal obtained from an acoustic sensor, the first watermark identifying media presented by a monitored media device, determining whether a second watermark, different from the first watermark, is embedded in the first audio signal obtained from the acoustic sensor, the second watermark identifying at least one of a mobile device or a user of the mobile device, classifying the second watermark as a media watermark or a people monitoring watermark based on a characteristic of the second watermark, and when the second watermark is determined to be embedded in the first audio signal, reporting at least one of the second watermark or information decoded from the second watermark to identify at least one of the mobile device or the user of the mobile device as being exposed to the media presented by the monitored media device.
Owner:THE NIELSEN CO (US) LLC

A multi-bit audio watermarking method based on phase distribution and efficient bit mapping

The application provides a multi-bit audio watermarking method based on phase distribution and high-efficiency bit mapping, and belongs to the field of audio information hiding. The application divides multiple independent phase subintervals based on the natural distribution characteristics of phases, and can realize one-to-many mapping from single feature to multi-bit watermarking by constructing unique phase features in the phase subintervals corresponding to watermarking information. In addition, the application introduces an optimization strategy, and can realize automatic phase redistribution. The application can significantly reduce the number of subspaces / feature modes required for multi-bit embedding, and reduce the design difficulty of the multi-bit watermarking algorithm. In addition, since an automatic phase redistribution algorithm based on optimization constraints is adopted, the algorithm can more effectively balance the inaudibility and robustness.
Owner:TIANJIN POLYTECHNIC UNIV

Audio watermark embedding method and device, electronic equipment and storage medium

ActiveCN121331148ASpeech analysisWatermark robustnessAlgorithm
The invention provides an audio watermark embedding method and device, electronic equipment and a storage medium, and belongs to the technical field of audio processing, and the method comprises the steps: obtaining the amplitude spectrum characteristics of an original audio signal, and determining a watermark embedding energy mask; inputting the amplitude spectrum features into a plurality of mask prediction network models to obtain watermark embedding risk masks; determining a target watermark embedding risk mask according to the watermark embedding risk mask to generate a frame-level embedding mask; and embedding the watermark information into the feature component corresponding to the target audio frame to obtain a watermark-containing amplitude spectrum feature, and performing inverse time-frequency transformation on the watermark-containing amplitude spectrum feature to obtain a watermark-containing audio signal. According to the method, the audio energy and the model prediction risk are combined, the audio frame which is most suitable for embedding the watermark is adaptively screened out, modification in hearing sensitive frames such as silent frames and transient frames is avoided, the problem that the audio quality is reduced due to undifferentiated embedding of the watermark is solved, and on the premise that the watermark robustness is not sacrificed, the watermark embedding efficiency is improved. And the imperceptibility of the watermark and the auditory quality of the audio containing the watermark are obviously improved.
Owner:IFLYTEK CO LTD +1

Strong-robustness lossless high-capacity audio watermark embedding method based on deep learning

The invention discloses a high-robustness lossless high-capacity audio watermark embedding method based on deep learning, and relates to the technical field of digital watermarking. The method comprises the following steps of audio-watermark information preprocessing, watermark information embedding and watermark embedding detection and judgment. According to the method, the original audio and the watermark information to be embedded are correspondingly preprocessed, so that the quality of the original audio and the watermark information is effectively improved, the subsequent embedding process is more efficient and accurate, then the preprocessed audio and watermark information are input into the audio embedding network to output the watermark audio, and the watermark embedding efficiency is improved. According to the method, high-capacity watermark information is embedded, audio distortion in the embedding process is effectively avoided, the quality of the embedded audio is guaranteed, finally, the stability and the anti-jamming capability of the embedded watermark can be effectively detected by introducing adversarial training attack simulation, the audio distortion is reduced to the maximum extent while the watermark capacity is guaranteed, and the watermark quality is improved. And the anti-interference capability of the watermark is improved.
Owner:GUANGZHOU SHUOGU TECHNOLOGY CO LTD

Hotword suppression

A method includes adding, by a first computing device, a first audio watermark to first speech data corresponding to playback of a first utterance including a hotword used to invoke an attention of a second computing device. The method includes outputting, by the first computing device, the playback of the first utterance corresponding to the watermarked first speech data. The second computing device is configured to receive the watermarked first speech data and determine to cease processing of the watermarked first speech data.
Owner:GOOGLE LLC

Audio watermark embedding method and apparatus, electronic device, and storage medium

Embodiments of the present application provide an audio watermark embedding method and device, electronic equipment and storage medium, belonging to the technical field of speech processing. The method comprises: performing attribute division on watermark attribute data to obtain a watermark attribute block, and performing content division on watermark content data to obtain a watermark content block; performing encoding processing on the watermark attribute block, merging the encoded attribute code sequence block and the watermark content block to obtain a target watermark data block; and embedding the target watermark data block into a target spectrum according to an encoding rate to obtain a target audio. Through encoding the watermark attribute block, the embodiments of the present application can enable the watermark to be identified even when the position of the watermark in the audio signal is shifted; and embedding the target watermark data block into the target spectrum according to the encoding rate can repeatedly embed the audio watermark into the audio, and when part of the audio watermark is damaged or lost, complete watermark information can still be extracted from the remaining audio, thereby improving the effect of audio watermark embedding.
Owner:PING AN TECH (SHENZHEN) CO LTD

Anti-copying audio watermark embedding method, extraction method and device

The present invention relates to the field of audio watermark technology, and in particular to an anti-copying audio watermark embedding method, extraction method and device. The anti-copying audio watermark embedding and extraction method first pre-processes the audio data and uses fast Fourier transform to convert the original audio signal from the time domain to the frequency domain; then constructs the watermark information and embeds the watermark bits into specific frequency points based on the frequency domain; finally, uses inverse fast Fourier transform to restore the frequency domain signal to the time domain signal, superimposes the watermark audio and the original audio signal in proportion, and outputs an audio file with a watermark. The present invention embeds the watermark information into the frequency domain of the audio and uses the frequency domain characteristics to enhance the robustness of the watermark, thereby realizing self-synchronous extraction of the watermark. It can effectively solve the time domain offset problem caused by copying distortion in traditional methods, and at the same time ensures the high accuracy of watermark extraction through a robust frequency domain analysis method.
Owner:HEFEI HIGH DIMENSIONAL DATA TECH CO LTD +1

Audio watermark embedding method and apparatus, and conference system

PCT designated stageWO2026056323A1Speech analysisAudio watermarkEngineering
An audio watermark embedding method (600) and apparatus, and a conference system, relating to the technical field of digital watermarking. The method (600) comprises: acquiring first original audio collected, during a first time period, from an environment where a first terminal is located, the first original audio comprising first playback audio played back by the first terminal during the first time period, and the first playback audio being configured to use a first watermark embedding algorithm to embed a first watermark on the basis of a first watermark embedding parameter value (601); performing watermark extraction on the first original audio on the basis of the first watermark embedding algorithm (602); and if the extraction of the first watermark from the first original audio fails, using the first watermark embedding algorithm to embed, on the basis of a second watermark embedding parameter value, the first watermark into first audio to be played back that is to be played back by the first terminal, so as to obtain second playback audio, the second watermark embedding parameter value differing from the first watermark embedding parameter value, and the second playback audio being configured to be played back by the first terminal during a second time period (603).
Owner:HUAWEI CLOUD COMPUTING TECHNOLOGIES CO LTD