Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

9results about How to "Guaranteed naturalness" patented technology

Real-time call translation method and device

The invention provides a real-time call translation method and device, relates to the technical field of data processing, and is applied to a Bluetooth headset, and the Bluetooth headset establishes communication connection with a mobile terminal through a Bluetooth hands-free protocol and establishes communication connection with a translation application running in the mobile terminal through a Bluetooth low-power-consumption general attribute protocol. The method comprises the following steps: acquiring an original sound to be translated, and performing audio compression coding on the original sound to be translated to obtain a compressed audio; transmitting the compressed audio to the translation application through a Bluetooth low-power-consumption general attribute protocol; receiving a translated audio returned by the translation application through a Bluetooth low-power universal attribute protocol, wherein the translated audio is an audio generated after the translation application translates the compressed audio; and carrying out sound mixing processing on the original sound to be translated and the translated audio to obtain a mixed audio, and outputting the mixed audio.
Owner:IFLYTEK CO LTD

Image enhancement method and application thereof

The application discloses an image enhancement method and application thereof, the image enhancement method enhances the brightness constant processing on the collected image through a constructed camera response model, obtains an enhanced image, converts the enhanced image into a gray matrix image and then carries out denoising processing. The application also provides an edge extraction method based on the image enhancement method, uses a new SED model based on dynamic feature fusion to extract edges, that is, carries out semantic segmentation on the denoised gray image, converts the denoised gray image into a binary image, normalizes the amplitude scale of the multi-layer features of the binary image, carries out dynamic feature fusion, obtains the required edge features, and thus realizes edge extraction. The image enhancement method and the edge extraction method are both upgraded, and the application prospect is wide, and the image enhancement method and the edge extraction method can especially meet the demand of large batch and high precision product size detection, and reduce the artificial pressure.
Owner:MINDU INNOVATION LAB

Active defense method, device and equipment for diffusion speech conversion, and medium

ActiveCN121811851BImprove active defense capabilitiesachieve global optimizationEngineeringAcoustics
The application discloses an active defense method, device and equipment for diffusion speech conversion and a medium, and relates to the technical field of speech security. The active defense method comprises the following steps: obtaining source speech and reference speech, introducing a constrained protection disturbance into the reference speech, constructing protected speech, and ensuring that the protection disturbance satisfies a disturbance imperceptibility constraint. The source speech and the protected speech are input into a speech conversion model with the reference speech as a speaker condition, content information analysis, speaker condition constraint acoustic generation and waveform synthesis are completed by the speech conversion model, and a protection generated speech is output. A joint loss function is constructed based on speaker embedding representations of the protection generated speech and the reference speech, gradients of a current disturbance position and a predicted midpoint position are calculated according to the joint loss function and are fused to obtain an update direction, the protection disturbance is iteratively optimized, and projection or clipping is performed to satisfy the disturbance imperceptibility constraint. Finally, the protected speech is output.
Owner:HUAQIAO UNIVERSITY

Working face dynamic machine-following panoramic video stitching method and device

PendingCN122069333AGuaranteed visual qualityGuaranteed naturalnessTelevision system detailsColor television detailsCorrection algorithmComputer graphics (images)
The invention provides a working face dynamic machine-following panoramic video stitching method and device. The method comprises the following steps: calling an original video stream of a target camera; wherein the target cameras are a plurality of cameras which are determined according to the position information of the coal mining machine and are within a preset range with the coal mining machine as the center; mapping the plurality of original video streams to a global coordinate system of the same working face according to the position information through a dynamic coordinate system conversion algorithm to obtain a machine-following panoramic video stream; performing local correction on the machine-following panoramic video stream through a local deformation correction algorithm to obtain a corrected video stream; fusing the overlapping areas of the adjacent cameras in the machine-following panoramic video stream through a multi-band fusion algorithm to obtain a fused video stream; the dynamic panoramic video which takes the coal mining machine as the center and continuously changes along with the movement of the coal mining machine is generated, and the remote control operation efficiency is remarkably improved. And the visual quality and naturalness of the dynamic panoramic video are ensured.
Owner:YANKUANG ENERGY GRP CO LTD +2

Method for training animal dependent visual clue positioning reward ability by using animal behavior training and testing system

The invention provides a method for training the visual clue-dependent positioning reward capability of animals by using an animal behavior training and testing system, and the system realizes that one set of system covers multiple types of experiments through the opening and closing logic of an access door of an adjusting arm, the reward region reward putting rule and the free conversion of visual clues. A visual clue-reward association learning scene is constructed, and animals are guided to establish memory association between clues and reward positions. A three-stage visual clue navigation and memory ability testing method is established, the memory retention ability associated with clue rewards is evaluated in the association stage, the spatial memory ability is independently evaluated in the blank stage, and the regulation and control mechanism of visual clues on animal spatial memory is analyzed in the non-association stage.
Owner:BEIJING INST OF TECH

A method and apparatus for bit-flipping attacks targeting large language models

This invention discloses a bit-flipping attack method and apparatus targeting large language models, relating to the field of large language model security technology. The method includes: generating an attack dataset containing questions using a large language model; inputting the attack dataset into the target large language model for forward propagation, outputting text data; constructing a perplexity loss function based on the text data; using a part-of-speech tagger to filter keyword elements in the text data, obtaining processed keyword elements; constructing a keyword element loss function based on the processed keyword elements; integrating the perplexity loss function and the keyword element loss function to obtain a total loss function; calculating the gradient value of each parameter based on the total loss function; and using a progressive bit search method based on the gradient values ​​to search for vulnerable bits in the target large language model to complete the bit-flipping attack. This invention can effectively reduce the accuracy of the output while maintaining the naturalness of the target large language model's output.
Owner:ZHEJIANG UNIV

An intelligent audio noise reduction system based on an attention mechanism

PendingCN122598673Aavoid distortion problemsGuaranteed naturalness
The application relates to the technical field of artificial intelligence, in particular to an intelligent audio noise reduction system based on an attention mechanism, which comprises an audio input end, a feature extraction module, a noise reduction module and an audio output end; the audio input end is used for acquiring a noisy audio signal and outputting the noisy audio signal; the feature extraction module is used for performing multi-scale time-frequency feature extraction and feature fusion on the noisy audio signal; the noise reduction module comprises a time sequence modeling unit, an attention unit and a noise estimation and suppression unit. The application can effectively distinguish speech and noise, avoid speech distortion, intelligently adjust attention weights, pertinently retain speech dominant area features and suppress noise dominant areas, iteratively converge to improve noise reduction effect and output high-quality noise reduction audio.
Owner:SHANGHAI MAIJUN TECHNOLOGY CO LTD

Head-mounted display equipment control method and device, equipment and storage medium

PendingCN121957402Aachieve naturalnessGuaranteed naturalnessInput/output for user-computer interactionImage data processingComputer hardwareMedicine
The embodiment of the invention provides a head-mounted display equipment control method and device, equipment and a storage medium, and belongs to the field of head-mounted display equipment. The method comprises the steps that under the condition that the head-mounted display device displays virtual content, in response to the change of the head posture of a wearer of the head-mounted display device, the current three-axis posture and the target anti-shake axial direction of the head-mounted display device are obtained, the target anti-shake axial direction is determined according to the current use scene of the head-mounted display device, and the target anti-shake axial direction is determined according to the current use scene of the head-mounted display device; the target anti-shake axial direction comprises at least one of a transverse rolling axis, a pitching axis and a yaw axis of the head-mounted display equipment; compensating the attitude of the target anti-shake axial direction in the current three-axis attitude to obtain a target three-axis attitude of the head-mounted display equipment; and according to the target three-axis attitude, controlling the head-mounted display device to adjust the pose and / or perspective deformation of the currently displayed virtual content. According to the technical scheme of the embodiment of the invention, the flexibility of attitude compensation of the head-mounted display equipment is improved.
Owner:ZHUHAI MOJIE TECH CO LTD