Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

518 results about "Acoustic source localization" patented technology

Acoustic source localization is the task of locating a sound source given measurements of the sound field. The sound field can be described using physical quantities like sound pressure and particle velocity. By measuring these properties it is possible to obtain a source direction.

Intelligent microphone pickup and speech enhancement method, system and device

The invention relates to an intelligent microphone pickup and speech enhancement method, system and device, and the method comprises the steps: obtaining an audio signal collected by a dual-silicon microphone array, and carrying out the cross-correlation analysis of the audio signal, and obtaining a corresponding human voice signal correlation feature; performing sound source positioning analysis on the audio signal based on the human voice signal correlation feature to obtain a corresponding target human voice signal; performing frequency characteristic analysis on the target human voice signal, and performing segmented dynamic gain processing on the signal according to a preset frequency band range to obtain a corresponding voice enhancement signal; performing noise component adaptive filtering processing on the target human voice signal to obtain a corresponding noise suppression parameter; and inputting the speech enhancement signal and the noise suppression parameter into a preset echo cancellation model for joint optimization to obtain a corresponding output human voice signal. According to the invention, the coupling problem of multiple acoustic interferences can be effectively solved.
Owner:SHENZHEN SHIDU DIGITAL TECH CO LTD

Microphone array sound source localization method and system based on cross-correlation-beam forming closed-loop optimization

The invention relates to a microphone array sound source positioning method and system based on cross-correlation-beam forming closed-loop optimization, and belongs to the technical field of sound source positioning. The method comprises the following steps: collecting multichannel sound signals through a microphone array and preprocessing the multichannel sound signals to extract time-frequency features and suppress noise interference; time delay information among the microphones is estimated by adopting a generalized cross-correlation phase transformation algorithm, and an optimization strategy is introduced to improve estimation stability and anti-interference performance; enhancing the target sound source signal in combination with a minimum variance undistorted response beam forming algorithm and an adaptive Kalman filtering mechanism; constructing a closed-loop feedback optimization mechanism based on the beam output signal to realize feedback adjustment; and adopting a hybrid network architecture, taking the beam output signal amplitude spectrum as input, and outputting the frequency spectrum or mask of the obtained target sound source signal. The method has the advantages of high calculation efficiency, high positioning precision and strong anti-interference capability, and is suitable for real-time acoustic signal processing in a complex environment.
Owner:CHONGQING UNIV OF POSTS & TELECOMM

Multi-element microphone array sound source localization method based on multistage signal preprocessing and subspace spectrum optimization

The invention relates to a multi-element microphone array sound source localization method based on multistage signal preprocessing and subspace spectrum optimization, and belongs to the technical field of acoustic detection. Aiming at the problems of poor noise immunity, weak multi-sound-source resolution capability and low calculation efficiency of the existing sound source positioning technology, a triple signal preprocessing and subspace collaborative optimization scheme is provided; firstly, incoherent noise is suppressed through phase coherent filtering, a signal is reconstructed through principal component analysis, and phase deviation is calibrated through fundamental frequency; then constructing a guiding matrix and decomposing a noise subspace, and extracting a coarse positioning result; and finally, high-precision angle optimization is realized based on a chaos initialization differential evolution algorithm, and the efficiency is improved by combining a dynamic search range and an early stop mechanism. According to the method, the anti-interference capability in a low signal-to-noise ratio environment is remarkably enhanced, the problems of missing detection and false detection during dense distribution of multiple sound sources are effectively solved, meanwhile, the positioning precision and the real-time performance are considered, and the method is suitable for acoustic fault detection of complex scenes such as power transmission line inspection.
Owner:CHONGQING UNIV

Immersed tunnel leakage voiceprint recognition system based on deep learning

The invention discloses an immersed tunnel leakage voiceprint recognition system based on deep learning. The system comprises an acoustic signal acquisition module, a signal preprocessing module, a voiceprint feature extraction module, a deep learning recognition module and a leakage positioning and early warning module. According to the invention, the leakage water flow sound and the structural damage elastic wave signal of the immersed tunnel are comprehensively considered, the multi-dimensional voiceprint feature extraction and deep learning model are fused, and the adaptive signal noise reduction and sound source positioning technology is combined, so that the high-precision monitoring and real-time early warning of the leakage of the immersed tunnel are realized; the accuracy and real-time performance of immersed tunnel leakage nondestructive monitoring are improved, and powerful support is provided for follow-up risk prediction and early repair.
Owner:TIANJIN PORT ENG INST LTD OF CCCC FIRST HARBOR ENG +2

Submarine cable fault positioning system based on underwater beacons

The invention relates to the technical field of submarine cable fault positioning, in particular to a submarine cable fault positioning system based on underwater beacons. The method comprises the steps that an underwater acoustic beacon array unit monitors and collects submarine cable fault sound wave signals in real time based on underwater acoustic beacons; the fault signal processing analysis unit extracts submarine cable fault sound wave signal parameters based on a non-point source space sound source waveform analysis model, and generates multi-parameter fault sound source feature vectors; the sound source positioning analysis algorithm unit performs three-dimensional inversion positioning based on the multi-parameter fault sound source feature vector and the arrival time difference of the submarine cable fault sound wave signal to the underwater acoustic beacon, and obtains the spatial position information of the submarine cable fault point; and the data communication alarm management unit sends the spatial position information of the submarine cable fault point to a shore-based operation and maintenance system. The submarine cable fault positioning method is used for realizing a submarine cable fault positioning technology for quickly and accurately positioning and identifying spatial extension characteristics in a complex submarine environment.
Owner:HUANENG RUDONG BAXIANJIAO OFFSHORE WIND POWER GENERATION CO LTD +2

Single-microphone multi-array pickup method and system based on sound attenuation simulation

The invention provides a single-microphone multi-array pickup method and system based on sound attenuation simulation, and the method comprises the steps: obtaining an original audio signal recorded by a single microphone and sound source direction information, and carrying out the sound attenuation simulation based on the sound source direction information and a preset virtual microphone array orientation parameter, calculating the included angle between the sound source direction and each virtual microphone; generating an audio intensity attenuation coefficient of each virtual microphone according to the included angle; and acting the audio intensity attenuation coefficient on the original audio signal to generate multi-channel audio data simulating the multi-microphone array. By adopting the method, the response difference of different microphone positions to the sound source can be simulated, the pickup result of the multi-microphone array is obtained, and high hardware cost and complex installation flow caused by actual deployment of a complex microphone array are avoided; and low-cost and high-diversity training data sources are provided for sound source localization, noise suppression and other models based on deep learning.
Owner:BEIJING YUANZHI DIGITAL INFORMATION TECHNOLOGY CO LTD

Sound beam forming pointing control method based on crowd recognition and related equipment thereof

The invention relates to the technical field of sound beam control, and provides a sound beam forming pointing control method based on crowd recognition and related equipment thereof. The method comprises the following steps: performing time-space synchronous acquisition and calibration on multi-modal sensor data to obtain a time-space aligned original data set, and performing layered signal enhancement on the original data set to obtain a visual feature tensor and an audio feature tensor; dense crowd detection and tracking are carried out according to the visual feature tensor to obtain a target crowd space-time trajectory matrix, and sound source localization and beam weight calculation are carried out in combination with the audio feature tensor to obtain a beam forming weight vector; performing multi-constraint dynamic optimization in combination with environmental parameters in the multi-modal sensor data to obtain an adaptive control parameter set, and finally performing real-time beam forming in combination with a loudspeaker sequence to obtain a directional sound field output signal. According to the invention, target detection, sound source localization and dynamic adaptive control are combined on the multi-source data, so that the beam direction is dynamically adjusted.
Owner:SHENZHEN JINGJING TECH CO LTD

Fabricated building detection method and system

The invention relates to the technical field of building structure health monitoring, in particular to an assembly type building detection method and system, and the method comprises the steps: collecting a strong interference signal generated by an external impact source of a building, and obtaining the sound source positioning information of the strong interference signal, converging the sound source positioning information of the strong interference signal on a preset three-dimensional geometric model aligned with a coordinate system to generate a real-time interference source distribution map; in the conventional detection mode, after a to-be-distinguished target acoustic signal is detected, obtaining a theoretical acoustic signal arrival time sequence based on sound source positioning information of at least one strong interference signal in the real-time interference source distribution diagram; comparing the actual acoustic signal arrival time sequence of the target acoustic signal with the theoretical acoustic signal arrival time sequence to obtain a comparison result; and judging whether the target acoustic signal is damaged in the building or not according to a comparison result. Real damage events from the interior of the structure can be effectively separated and confirmed, and the accuracy and reliability of detection are improved.
Owner:GUANGDONG HUIHE ENG TESTING CO LTD

Examination room multi-source data fusion abnormal behavior intelligent analysis method and system

The invention provides an examination room multi-source data fusion abnormal behavior intelligent analysis method and system, and relates to the technical field of artificial intelligence, and the method comprises the steps: collecting examination room video and audio data, and carrying out the multi-view skeleton point fusion and sound source positioning to extract features; constructing a hidden Markov model and kernel density estimation to carry out abnormal behavior identification; calculating a seat correlation degree based on a spatial weight coefficient and wavelet decomposition to identify multi-person cooperative cheating; and early warning information is pushed in real time. According to the invention, the cheating behavior identification accuracy is improved, and effective detection of multi-person cooperative cheating behaviors is realized.
Owner:ATA ONLINE (BEIJING) EDUCATION TECH LTD

Implementation method and device of multi-channel voiceprint recognition system

The invention relates to the technical field of voice recognition, in particular to an implementation method and device of a multi-channel voiceprint recognition system, and the implementation method comprises the steps of multi-channel data acquisition and synchronization, signal preprocessing and enhancement, feature extraction and fusion, model training, real-time deployment and adaptive optimization. Compared with the problems that a traditional multichannel voiceprint recognition system depends on a fixed beam forming algorithm and an independent clock synchronization module, the synchronization error is large, manual parameter adjustment is needed for noise suppression, and generalization is poor, hardware-level clock synchronization is achieved through a PTP protocol, and the accuracy of noise suppression is improved. The method combines an end-to-end neural network to automatically learn noise distribution and a sound source space position, dynamically generates a beam forming weight, can improve the voice quality in a complex noise scene without manual intervention, remarkably reduces the interference of a synchronization error on sound source positioning, and enables the precision and stability of far-field voice enhancement to reach a new level.
Owner:MINAMI ACOUSTICS LTD

All-weather sound source positioning system and method based on multi-sensor fusion

The invention discloses an all-weather sound source positioning system and method based on multi-sensor fusion, and the method comprises the steps: firstly, obtaining acoustic sensing data, millimeter wave radar data and infrared thermal imaging data, and enabling each data to have a collection timestamp; then, cross-modal time alignment processing is carried out on the multi-source data to map the multi-source data to a unified time reference, and multi-modal fusion features under a unified time axis are obtained; for each time point in the time-aligned multi-modal fusion features, according to the confidence coefficient of the data of each sensor, adaptive weighted fusion is carried out on the data of different modals, and fused common feature representation is generated; and finally, time sequence modeling and joint reasoning are carried out based on the common feature representations of a plurality of continuous time points, and continuous position information of the sound source in the three-dimensional space is regressed. According to the method, high-precision positioning of three-dimensional positions of a plurality of sound sources is realized, so that the accuracy and the stability of sound source positioning are improved in a complex environment and an all-weather condition.
Owner:HANGZHOU DIANZI UNIV

Road traffic noise intelligent monitoring and three-dimensional sound field reconstruction system

The invention relates to the technical field of environmental noise monitoring, and discloses a road traffic noise intelligent monitoring and three-dimensional sound field reconstruction system, which comprises an acoustic sensor module, a data collection and storage module, a data analysis and evaluation module, a three-dimensional sound field construction and display module and a traffic flow feature library. According to the system, a differential geometry principle is adopted, a sound field is regarded as a Riemannian manifold with a local microstructure, and accurate description of an irregular sound field is realized through a curvature self-adaptive sound field manifold construction technology; introducing a covariant derivative in Riemannian geometry, and constructing a sound propagation model adapted to a complex road environment; and realizing hierarchical decomposition and reconstruction of the sound field by using a multi-scale analysis theory. According to the method, the sound source positioning precision and the calculation efficiency are improved, seamless analysis from microcosmic to macroscopic is realized, an innovative solution is provided for traffic and noise collaborative management, and intelligent traffic and environmental noise management are effectively supported.
Owner:SHAANXI XIEHUA TECHNOLOGY CO LTD

Real-time duplex translation method based on multi-channel parallel processing and corresponding product

The invention relates to the field of real-time translation, and provides a real-time duplex translation method based on multichannel parallel processing and a corresponding product, and the method comprises the steps: collecting multipath voice signals of at least two user groups in real time through a group of audio collection modules; dynamically adjusting beam forming parameters of each audio acquisition module in one group of audio acquisition modules based on a sound source positioning result, and feeding back the beam forming parameters to the corresponding audio acquisition modules; monitoring the voice activity of each audio acquisition module corresponding to each audio channel; when it is monitored that the voice activity of any audio channel reaches a preset condition, automatically activating the translation processing flow of the audio channel and keeping the monitoring state of the other audio channels; a parallel processing mechanism is adopted for voice signals of the activated audio channels, and meanwhile real-time translation of the currently activated audio channels and voice activity monitoring of the other audio channels are executed; and transmitting a translation result of the current speaking user to other users participating in dialogue in the user group to realize synchronous coordination of multichannel data.
Owner:MEIG SMART TECH CO LTD +1

Sound source direction finding positioning method and system oriented to complex environment

The invention relates to the technical field of voice processing, in particular to a sound source direction-finding positioning method and system for a complex environment, and the method comprises the steps: serializing the sound intensity of each microphone, and calculating a maximum receiving interval according to the distance between adjacent microphones, the sound transmission speed and the sampling frequency; the method comprises the following steps: determining a short-time audio similarity through moving a sound source signal sequence, extracting a sound source synchronization sequence and a sound source same-frequency sequence, and further evaluating a tone quality deviation degree; optionally selecting a reference microphone, and comparing the reference microphone with a sound source synchronization sequence of other microphones to obtain a time delay data length and a sound source similar sequence; based on the arrangement values of the same position elements of the sound source similar sequence, a sound intensity distribution sequence is obtained, and the sound intensity interference degree of each microphone is obtained; and determining a fixed beam direction and calculating the distance between the reference microphone and the sound source. The invention aims to improve the accuracy of sound source positioning.
Owner:SUZHOU AUDITORYWORKS CO LTD

Space projection regularization method for sound source localization and sound field reconstruction and related device

PendingCN120761972APosition fixationEquivalent source methodSound sources
The invention discloses a spatial projection regularization method for sound source localization and sound field reconstruction and a related device, and belongs to the technical field of sound source localization. Specifically, according to the method, an equivalent source method model is utilized, and an external radiation sound field of a sound source to be measured is simulated through a series of virtual sources. Then, according to the first sound pressure transfer matrix from the virtual source arrangement surface to the holographic measuring point plane and the sound pressure of the holographic measuring point, a space projection regularization algorithm is used for solving the inverse problem, and therefore the equivalent source intensity is obtained; and then, reconstructing an external radiation sound field of the sound source to be detected by using the second sound pressure transfer matrix and the equivalent source intensity. And finally, by taking the maximum sound pressure amplitude as a standard, screening field points in an external radiation sound field so as to determine the position of the sound source to be detected. According to the space projection regularization algorithm, truncation processing of the vector space is carried out with the projection size of the measurement information in the vector space as the criterion, real information is reserved, and noise interference is restrained to the maximum extent.
Owner:CHONGQING UNIV

Vehicle exterior voice control method and device, storage medium and electronic equipment

The invention provides an out-of-vehicle voice control method and device, a storage medium and electronic equipment, and relates to the technical field of artificial intelligence, and the method comprises the steps: firstly obtaining audio signals collected by a microphone array on a vehicle for different areas outside the vehicle; the audio signals are preprocessed; performing sound source positioning and sound source separation according to the preprocessed audio signal to obtain a multi-channel audio signal; then performing wake-up word recognition based on the multi-channel audio signal, and determining a target channel where the wake-up word is located; and executing a corresponding vehicle voice control command according to the audio signal corresponding to the target channel. According to the method and the device, the audio signals are acquired through the microphone arrays in different areas around the vehicle, and the multichannel audio signals are obtained through preprocessing, sound source positioning and separation, so that accurate wake-up word recognition and voice control command execution are realized, outdoor complex environmental noise can be effectively shielded, and the accuracy of out-of-vehicle voice control is greatly improved.
Owner:XIAOMI EV TECH CO LTD +2

Device for Acoustic Source Localization

Acoustic signals from an acoustic event are captured via sensing nodes of sensor group(s) that comprise a group of sensing nodes at a location comprising spatial boundaries. Each of the sensing nodes comprise a sensor area. Each of the sensor group(s) is based on: range limits of each of the sensing nodes; shared sensing areas of the sensing nodes; and intersections between the sensor area for each of the sensing nodes and the spatial boundaries. Solutions(s) are generated by processing the acoustic signals. The solution(s) indicate the location or trajectory of the acoustic event. A strength of solution compliance value for at least one of the solution(s) is determined. A refined solution is generated employing: sensor contributions of sensing nodes; and the strength of solution compliance value with the spatial boundaries and at least one of the solution(s). A report is created comprising the location or trajectory of the acoustic event.
Owner:DATABUOY CORP

Gas leakage detection method, device, system, equipment, medium and program product

The invention provides a gas leakage detection method, device, system and equipment, a medium and a program product, and relates to the technical field of equipment detection. The method comprises the following steps: acquiring a sound source positioning coordinate of a sound source positioning point; the sound source positioning coordinates are used for representing the relative position of the sound source positioning point relative to the sound source positioning device; converting the sound source positioning coordinate into a sound source image coordinate in a target image coordinate system; based on the sound source image coordinates and a plurality of to-be-detected areas of a to-be-detected device, determining a gas leakage detection result of the to-be-detected device; the plurality of to-be-detected areas are image areas under the target image coordinate system; wherein the plurality of to-be-detected areas are obtained by carrying out image segmentation on a scene image of a scene where the to-be-detected device is located, and the plurality of to-be-detected areas are all areas where the to-be-detected device may generate gas leakage. According to the invention, accurate positioning of gas leakage points can be realized, so that the accuracy of gas leakage detection is improved.
Owner:ZHEJIANG TIDAL POWER TECH CO LTD

Transformer abnormal sound source positioning method and system

The invention relates to a transformer abnormal sound source positioning method and system, and belongs to the technical field of power equipment state evaluation, and the method comprises the steps: constructing a Bayesian neural network embedded with a voiceprint physical mechanism, and enabling a sound wave propagation equation to serve as a physical constraint to be embedded into the Bayesian neural network; reconstructing the voiceprint signal by adopting a compressed sensing technology to obtain a reconstructed voiceprint field; designing a multi-task objective function including data fitting, physical constraint and positioning loss, and optimizing data fitting, physical constraint and positioning precision to obtain a trained Bayesian neural network; based on a gradient sound source inversion positioning algorithm and the trained Bayesian neural network, sparse regularization is combined to obtain a prediction result of accurate positioning; and a prediction result is visualized to a three-dimensional model of the transformer, and the position of an abnormal sound source is visually displayed. According to the method, the limitation of a traditional method in a complex environment is overcome, and high-precision and high-robustness transformer abnormal sound source positioning is realized.
Owner:CHINA ELECTRIC POWER RESEARCH INSTITUTE CO LTD +2

Sound source localization method based on multi-frequency separation and Newton optimization deconvolution

The invention discloses a sound source localization method based on multi-frequency separation and Newton optimization deconvolution, relates to the technical field of array acoustic signal processing, and is used for solving the problem that weak sound sources and multiple sound sources are difficult to identify. According to the method, dominant frequency is extracted through multichannel frequency domain analysis, a cross-spectrum matrix is constructed in combination with a near-field propagation model and a guide vector, delay summation beam forming is executed to obtain sound source preliminary distribution, then the distribution is regarded as a convolution result, a maximum likelihood model is introduced, and a two-stage deconvolution strategy of coarse estimation and Newton method fine optimization is adopted to obtain a high-resolution sound source. According to the multi-sound-source positioning method, subgrid-level analysis of sound source positions and amplitudes is achieved, finally, all frequency results are fused, continuous sound source images are smoothly output through a two-dimensional Gaussian kernel, the resolution and real-time performance of multi-sound-source positioning are remarkably improved, and the multi-sound-source positioning method is suitable for high-precision acoustic imaging in a complex sound field.
Owner:STATE GRID JIANGXI ELECTRIC POWER CO LTD

Sound source localization rehabilitation training method, system, equipment and medium

The invention provides a sound source localization rehabilitation training method, system and device and a medium. The sound source localization rehabilitation training method comprises self-adaptive training and scene simulation training, the self-adaptive training comprises the steps of providing a training task in the period, and the training task in the period comprises the step of selecting training difficulty; according to the training difficulty, sound stimulation is presented in different spatial orientations, and the training difficulty can be adjusted according to a certain logic; the training difficulty of the next period of training is adaptively adjusted according to the correct rate p of one training group of the subject until the preset training amount is reached, and the difficulty setting includes but is not limited to scene setting, analog input information amount and signal-to-noise ratio adjustment. According to the scheme, a hearing-impaired person can perform rehabilitation training in a home environment, the need of manual intervention is reduced, and the training convenience and flexibility are improved. Through adaptive training, on the basis of ensuring scientificity, accuracy and effect of a training plan, the positioning ability of the hearing-impaired person is enhanced, and the integration of the hearing-impaired person into the society is accelerated.
Owner:BEIJING JISHENG TECH CO LTD

Array decoupling sound source localization method and device based on deep learning, and readable medium

The invention discloses a formation decoupling sound source localization method and device based on deep learning and a readable medium, and the method comprises the steps: constructing and training a sound source localization model, and obtaining a trained sound source localization model; acquiring a first sound source signal received by a first microphone and a second sound source signal received by a second microphone in the microphone array; calculating generalized cross-correlation frequency domain representation between the first sound source signal and the second sound source signal based on the first sound source signal and the second sound source signal; obtaining a frequency domain feature based on the guide vector between the first microphone and the second microphone and the generalized cross-correlation frequency domain representation; determining an input feature based on the frequency domain feature, and inputting the input feature into a trained sound source localization model to obtain a candidate sound source angle and a confidence coefficient corresponding to the candidate sound source angle; and post-processing the candidate sound source angle to obtain a sound source positioning result. The invention solves the problems that the existing sound source localization method based on deep learning is large in calculated parameter quantity and cannot be applied to embedded equipment and the like.
Owner:YEALINK (XIAMEN) NETWORK TECHNOLOGY CO LTD

Education robot voice signal processing method

The invention discloses a voice signal processing method for an education robot, relates to the technical field of voice signal processing, and aims to solve the problems of multi-person overlapping language and environmental noise interference in a classroom, multi-channel data is acquired by relying on a microphone array and a camera, and an environmental model is constructed in the step 1 to determine a noise baseline and student distribution; in the second step, overlapped voices are detected, and sound source localization is carried out in combination with the time difference of arrival and mouth shape data; in the third step, directional gain is executed in the target direction, and a deep network is used for separating aliasing voice; and in the step 4, the separated voice is input into a children customized recognition engine to complete high-precision recognition and interaction in combination with confidence evaluation. The recognition accuracy and the interaction efficiency can be remarkably improved under the complex scenes of classroom reverberation and simultaneous speaking of multiple persons, meanwhile, the noise change is tracked through the global environment model so that the education robot can keep stable recognition performance in diversified teaching interaction, and the teaching effect is remarkably enhanced.
Owner:北京爱宾果科技有限公司

Sound source localization method and system based on fusion confidence

The invention relates to the technical field of sound source localization, and provides a sound source localization method and system based on fusion confidence. The method comprises the following steps: when a preset wake-up word is detected, starting sound source positioning processing, calculating a horizontal azimuth angle according to a detected current voice signal, and determining a sound source positioning area; calculating a matched existing user voiceprint, and confirming a current user corresponding to the current voice signal; calculating a user preference fusion coefficient corresponding to the current voice signal, calculating a fusion confidence coefficient of the current voice signal, dynamically adjusting a scanning range of a holder camera, capturing an image frame of a current user at a fixed interval in a scanning process, executing face detection, obtaining a to-be-processed face image, performing face feature extraction, and obtaining a to-be-processed face image; and carrying out visual identity collaborative confirmation, stopping rotation of the holder camera when an identity consistency condition is satisfied, and locking the current direction as a sound source positioning position. According to the invention, the hardware complexity is reduced, and the positioning precision is effectively improved.
Owner:CHINA UNICOM ONLINE INFORMATION TECHNOLOGY CO LTD

Robotic dog monitoring system for bird sound source positioning and tracking

The invention discloses a robot dog monitoring system for bird sound source positioning and tracking, and the system comprises a sound collection module which is used for collecting a multi-channel sound signal, and carrying out the preprocessing of the multi-channel sound signal; the edge calculation module is used for processing based on the preprocessed multi-channel sound signals to obtain a sound source positioning result; the robot dog motion control module is used for receiving the sound source positioning result, performing path planning in combination with a control algorithm, driving a robot dog to autonomously move to approach a sound source, and integrating a sound source array sensor to realize environment perception and navigation; the image recognition module is used for constructing a bird recognition model and deploying the bird recognition model in an image recognition sensor to obtain a bird recognition result; and the remote monitoring module is used for receiving and displaying the sound source positioning result, the motion trail of the robot dog and the bird recognition result in real time. According to the invention, an efficient and intelligent solution is provided for bird monitoring.
Owner:NORTHEAST FORESTRY UNIV +2

ResNet sound source localization method based on attention mechanism

The invention discloses a ResNet sound source localization method based on an attention mechanism, and relates to the technical field of sound source localization. The method comprises the following steps that sound source signals are obtained, the sound source signals comprise the sound source signal of each microphone in a microphone array, the sound source signals are converted into a frequency domain through short-time Fourier transform, phase components of the microphone sound source signals are extracted and obtained in the frequency domain, the phase components between every two microphones are subtracted to obtain a phase difference, and the phase difference is obtained; feeding the phase difference image into a trained residual network based on an improved attention mechanism, and outputting a sound source angle prediction probability by the network; the residual network comprises ResNet-34 Stage1-4, a first output part, an SC-SEAM module, ResNet-34 Stage5, an SC-SEAM module, a second output part, a full connection layer and an output layer which are connected in sequence, the residual network is an improved neural network structure, the sound source positioning precision is higher, and the robustness is better.
Owner:SUZHOU ACOUSTIC IND TECH RES INST CO LTD

Pipeline leakage external detection system and detection method

The invention relates to a pipeline leakage external detection system and a detection method. The invention relates to the technical field of pipeline leakage detection, and discloses a pipeline leakage detection system which collects leakage sound signals and normal pipeline environment sound signals through an acoustic sensor array, processes and analyzes the collected signals through a signal processing module, positions leakage positions through a sound source positioning technology and displays leakage information through images. And leakage information is transmitted to a big data cloud platform through a communication module, and alarm information is generated, so that related management personnel can perform remote monitoring, and workers can conveniently check and repair leakage positions.
Owner:HARBIN ENG UNIV +1

Acoustic emission source positioning method and system based on hybrid model

The invention provides a sound emission source positioning method and system based on a hybrid model, and relates to the technical field of sound source localization, and the method comprises the steps: calibrating the angle range and sound velocity value of each sound velocity partition; calculating a theoretical receiving time difference between the reference acoustic emission receiver probe and other acoustic emission receiver probes for receiving acoustic emission sources at different grid points; acquiring an actual receiving time difference; calculating a plurality of error values between the actual receiving time difference and theoretical receiving time differences corresponding to different grid point acoustic emission sources; the preset proportion error value is reserved, the reserved error value is clustered, and the position of a first prediction acoustic emission source of the target to be detected is positioned; inputting the current actual receiving time difference into the random forest model, and outputting a second predicted acoustic emission source position of the to-be-detected target; and verifying the comprehensive confidence of the first predicted acoustic emission source position, weighting the first predicted acoustic emission source position and the second predicted acoustic emission source position, and outputting the target predicted acoustic emission source position of the target to be detected.
Owner:NANJING FIBERGLASS RES & DESIGN INST CO LTD +2

Reconfigurable microphone array system and control method thereof

The invention relates to the technical field of voice recognition and sound source localization, and discloses a reconfigurable microphone array system and a control method thereof.The reconfigurable microphone array system comprises a flexible substrate used for bearing a microphone array, a microphone array module comprises at least two fixed microphones and four movable microphones, and the movable microphones are connected with the flexible substrate through magnetic attraction sliding blocks; the motor driving module is used for controlling the movable microphone to move along a preset track; the multi-mode audio processor integrates a sound source localization algorithm and a 3A audio processing algorithm and is used for analyzing a user instruction, generating a motor control signal and processing audio data; the closed-loop feedback module comprises a plurality of travel switches, detects the position of the motor in real time and feeds back the position to the audio processor; the man-machine interaction module is used for selecting an array form mode; according to the system, by adjusting the spatial layout of the movable microphones, the dynamic switching of the array form among a spherical form, a circular form and a linear form is realized. According to the invention, the flexibility and adaptability of the microphone array are improved.
Owner:SHANDONG INSPUR SCI RES INST CO LTD

Conference picture display method and device based on audio positioning

The invention belongs to the technical field of video conferences, and discloses a conference picture display method and device based on audio positioning, and the method comprises the steps: obtaining conference video data collected by a camera, and extracting a current image frame; detecting a portrait in the current image frame and a corresponding portrait position; judging whether audio data collected by the microphone array is received or not; if yes, sound source positioning is carried out according to the audio data, and a sound source orientation is obtained; determining the position of a spokesman in the portrait position of the current image frame according to the sound source orientation; and generating a focusing instruction according to the spokesman position and sending the focusing instruction to the camera. According to the method and the device, automatic frame selection and focusing of the spokesman can be realized, manual operation is not needed, and timeliness and accuracy of focusing the spokesman in the conference picture are greatly improved.
Owner:YEALINK (XIAMEN) NETWORK TECHNOLOGY CO LTD