Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

363 results about "Silence" patented technology

Silence is the absence of ambient audible sound, the emission of sounds of such low intensity that they do not draw attention to themselves, or the state of having ceased to produce sounds; this latter sense can be extended to apply to the cessation or absence of any form of communication, whether through speech or other medium.

Mute floor

The utility model discloses a mute floor which comprises a surface treatment layer, a protection layer, a decoration layer, a first carrier, a first buffer layer and a second carrier. The surface treatment layer is arranged on one side of the protection layer, the decoration layer is arranged on the other side of the protection layer, the first carrier is arranged between the decoration layer and the first buffer layer, the first buffer layer is arranged between the first carrier and the second carrier, and the first buffer layer is of a porous structure and / or a net structure. A buffer groove is formed in at least one side face of the first buffer layer and extends to the two ends of the side face. By arranging the floor with the first buffer layer with the porous and / or honeycomb structure, noise is reflected and refracted for multiple times in the porous and / or honeycomb structure, so that sound is attenuated, a sound wave propagation path is prolonged, and the effect of more effectively reducing noise propagation is achieved.
Owner:CHANGZHOU BEMATE HOME TECH CO LTD

Distributed audio group mute and recovery method and system, storage medium and equipment

The invention relates to the technical field of audio control, and discloses a distributed audio group mute and recovery method and system, a storage medium and a device, and the method comprises the steps: determining a master device in a distributed audio network, and enabling the master device to find and register a plurality of slave devices to establish an audio group; receiving a group mute or recovery instruction sent by a user through the master device, and performing validity verification on the instruction; the master device allocates a uniform timestamp to the verified instruction, and sends a synchronous control command to all slave devices in the group based on the timestamp; the slave equipment executes mute or recovery operation at the moment corresponding to the timestamp, synchronous response of the equipment in the group is realized, all the equipment is ensured to accurately and synchronously execute the mute / recovery operation through unified management of the master equipment and the adoption of global timestamp synchronization and adaptive delay prediction technologies, and an intelligent state monitoring and exception recovery mechanism is provided, so that the service life of the slave equipment is prolonged. Finally, efficient and reliable multi-room audio group control with excellent user experience is realized.
Owner:LINKPLAY TECHNOLOGY INC NANJING

Audio processing method and device

The embodiment of the invention discloses an audio processing method and device, and the method comprises the steps: carrying out the traversal of a segmented sentence list which is determined through the interception of an audio stream based on a mute detection result, and determining a current segmented sentence and an endpoint detection result of the current segmented sentence, determining a start frame of the current segmented sentence as a start frame of a latest to-be-identified sentence when the current segmented sentence has a sentence start point, determining an end frame of the current segmented sentence as an end frame of the latest to-be-identified sentence when the current segmented sentence has a sentence end point, and determining the latest to-be-identified sentence according to the start frame and the end frame of the latest to-be-identified sentence. According to the method and the device, the segmented sentences in the audio stream can be continuously updated, the to-be-recognized sentences can be determined according to the endpoint detection results of the segmented sentences, continuous long speech recognition is realized, and the complete content of the to-be-recognized sentences from the start frame to the end frame is recognized, so that the speech recognition efficiency is improved. The accuracy of long speech recognition can be improved.
Owner:ZHEJIANG FUTURE ELF ARTIFICIAL INTELLIGENCE TECH CO LTD

Audio recognition and noise reduction processing method and system for power grid regulation and control workbench environment

The invention discloses an audio recognition and noise reduction processing method and system for a power grid regulation and control workbench environment, and the method comprises the steps: obtaining pure audio signals of all power grid regulation and control workers in different working scenes and different background noises in a dispatching desk, carrying out the cross-domain fusion of two features of a time domain and a frequency domain, and taking the two features as a training set to train a speaker model; voice signals are collected in real time, and pure noise and silence in the voice signals are removed; according to the noise level estimation and the noise power, carrying out weighting processing on the voice signal after the pure noise and the mute are removed; recognizing and separating human voice fragments by adopting a continuous down-sampling and resampling algorithm with multi-resolution characteristics; judging whether the human voice fragment contains a target speaker or not through a speaker model; and generating a plurality of voice segments based on the human voice segments, and extracting target speaking human voice through clustering and a speaker model. The method adapts to environmental changes of regulation and control dispatching desks at all levels, environmental noise can be remarkably reduced, and accurate recognition and extraction of human voices are achieved.
Owner:STATE GRID FUJIAN ELECTRIC POWER CO LTD +3

Speech enhancement and high-precision recognition method and system in complex environment

PendingCN121641016ASpeech recognitionSpectral density estimationNerve network
The invention provides a voice enhancement and high-precision recognition method and system in a complex environment, and relates to the technical field of voice processing, and the method comprises the steps: collecting a time domain signal in an off-road parking sentry box environment for preprocessing, detecting a mute segment signal in a standard time domain signal for noise power spectral density estimation, and obtaining a noise power spectral density value; a reverberation parameter is obtained by combining voice onset information and noise spatial correlation estimation, prediction is performed by using a deep neural network model, voice masking is applied to microphone array signals to perform enhancement processing, adaptive feature extraction is performed on time domain enhanced voice signals, and a voice signal is obtained. And performing high-precision recognition on the voice adaptive feature sequence based on an acoustic model and a language model, and outputting a target recognition text. The technical problems of poor voice signal quality and low recognition accuracy in a complex noise environment in the prior art are solved. The technical effects of improving the voice signal quality and the recognition accuracy and realizing clear, accurate and real-time voice interaction are achieved.
Owner:INTELLIGENT INTER CONNECTION TECH CO LTD

OTA upgrading method and system for Bluetooth sound equipment

The invention relates to the technical field of wireless communication, in particular to an OTA upgrading method and system for Bluetooth sound equipment. The method comprises the following steps: acquiring to-be-upgraded firmware data and currently played audio data, and acquiring power amplifier current feedback data of a sound speaker by sending a double-frequency detection signal containing 21kHz and 22kHz; acquiring equipment transmission capability confirmation data according to a silent response mode of the sound equipment to the preset handshake signal; and calculating a nonlinear distortion coefficient of the loudspeaker according to the intensity of the 1kHz difference frequency component in the power amplifier current feedback data, and obtaining real-time channel evaluation data based on current waveform deviation analysis. According to the method, the nonlinear distortion degree of the loudspeaker is accurately evaluated through the 1kHz difference frequency component generated by the 21kHz and 22kHz dual-frequency detection signals, so that the reliability of data transmission is more accurately predicted.
Owner:深圳中成圣达科技有限公司

Audio processing method, device and equipment for calling for help for old people and medium

The invention relates to the technical field of elderly care, solves the problems that in the prior art, due to the fact that noisy and intermittent elderly voice is lack of robust sentence bound recognition and interruption repair, misinformation and missing information of calling-for-help triggering are prone to occurring, and time delay is high, and provides an audio processing method and device for elderly calling-for-help, equipment and a medium. The method comprises the following steps: preprocessing a voice signal of the elderly to obtain a preprocessed voice signal; performing statement boundary recognition and labeling on the preprocessed voice signal to obtain a voice fragment set; carrying out interruption segment and silent interval detection on the voice segment set to obtain an initial voice segment sequence, and carrying out intelligent recombination and time sequence repair on the initial voice segment sequence to obtain a target voice segment sequence; and carrying out emergency call recognition on the target voice segment sequence to obtain an emergency call judgment result. According to the invention, the accuracy of distress call identification of the elderly is improved.
Owner:NINGBO SIMSHINE INTELLIGENT TECH CO LTD

Conversational AI low-delay response control method and system based on semantic analysis

The invention relates to the technical field of voice interaction, in particular to a dialogue type AI low-delay response control method and system based on semantic analysis. The method comprises the following steps: firstly, generating a driving intensity correlation factor of a current time window according to the change of CAN bus data; further monitoring the output of the NLU, analyzing the acoustic intonation raising characteristics, and obtaining an interaction suppression coefficient in combination with the driving intensity correlation factor; further acquiring a time pressure coefficient based on the mute duration after the user stops sounding; further comparing the interaction suppression coefficient and the time pressure coefficient of the current time window to obtain a response judgment value; further analyzing the change trend of the interaction inhibition coefficient, and updating an environment improvement flag bit; and finally, according to the driving intensity correlation factor, the environment improvement flag bit and the response decision value of the current time window, carrying out multi-level response decision control, and carrying out selective reset of an interaction state, thereby solving the problems of wrong truncation and response delay of voice interaction under dynamic driving.
Owner:SHANGHAI SHENGWANG TECH CO LTD

Audio identification method and device based on marking and backtracking correction

The invention relates to an audio recognition method and device based on marking and backtracking correction, and the method comprises the steps: segmenting a target audio, and enabling adjacent segments after segmentation to have a partial overlapping region; if a slice point exists in a non-mute segment of the target audio and the acoustic feature similarity of a preset number of frames before and after the slice point is smaller than a set similarity threshold value, the slice point is marked as a cut-off risk point, and the cut-off risk point is used for indicating that continuous semantics before and after the slice point has a cut-off risk; after a text corresponding to the audio of each slice is recognized through the speech recognition model, the language confidence degree of an overlapping area where the truncation risk point is located is determined, and the language confidence degree is used for indicating context language logic of the overlapping area; and if it is determined that the language confidence is lower than a set confidence threshold, correcting the recognition texts of the slices before and after the truncation risk point. According to the invention, the accuracy of audio recognition is improved.
Owner:FIBOCOM WIRELESS

Information processing system

An information processing system includes: a sensor unit that performs detection of contact of a pen-shaped input device with an operation target surface and presence of the pen-shaped input device within a detection limit distance on the operation target surface; a sound output unit that outputs a sound corresponding to input sound data; and a sound output control unit that inputs silent sound data to the sound output unit in response to the sensor unit detecting a first state in which the pen-shaped input device is detected within the detection limit distance, and inputs audible sound data to the sound output unit in response to the sensor unit detecting that, after the first state, a second state in which the pen-shaped input device is in contact with the operation target surface has been reached.
Owner:LENOVO (SINGAPORE) PTE LTD

Soft decision audio decoding system

ActiveCN115472169BSpeech analysisCode conversionNoiseSoft output Viterbi algorithm
The present invention provides a soft decision audio decoding system for maintaining audio continuity in a digital wireless audio receiver, which infers the likelihood of errors in a received digital signal based on generated hard bits and soft bits. The soft audio decoder can utilize the soft bits to determine whether the digital signal should be decoded or muted. The soft bits can be generated based on a detection point and detected noise power or by using a soft output Viterbi algorithm. The value of the soft bits can indicate the confidence in the strength of the hard bit generation. The soft decision audio decoding system can infer errors and decode perceptually acceptable audio without error detection as in conventional systems and has low latency and improved granularity.
Owner:SHURE ACQUISITION HLDG INC

Apparatus and methods for efficient muting and transmission of immersive audio bitstreams

An apparatus comprising means configured to: obtain, from a further apparatus, at least one bitstream comprising at least one component or stream of an immersive audio bitstream; identify at least one of the at least one component or stream of the immersive audio bitstream to be muted; transmit, to a further apparatus, a mute request indicating the identified at least one component or stream to be muted; obtain, from the further apparatus, the at least one bitstream, wherein the at least one bitstream is modified with respect to the identified at least one component or stream of the at least one immersive audio bitstream to be muted, wherein the modification in the at least one bitstream, is at least one of: the at least one component or stream having reduced energy; the at least one component or stream not being received; the at least one component or stream comprises a silence descriptor indicator for indicating the at least one component or stream is to be muted; the at least one component or stream having a reduced bit rate; and at least one other component or stream having an increased bit rate.
Owner:NOKIA TECHNOLOGIES OY

Content aware audio processing

PendingUS20260018179A1Speech analysisComfort noiseNoise
Content aware audio processing includes receiving, by a digital signal processor, a frame of audio data. In response to detecting that the frame of audio data is a silent frame, the digital signal processor selects a light graph from a plurality of graphs including the light graph and a full graph. Comfort noise is generated that corresponds to the silent frame. The comfort noise frame is processed through the light graph in place of the silent frame. The light graph is dedicated for processing comfort noise frames.
Owner:ADVANCED MICRO DEVICES INC

Intelligent Muting Of Participant Audio In Communication Sessions

An input audio signal associated with a communication session is received. A determination is made that a participant associated with the input audio signal is not audibly speaking within the input audio signal. In response to determining that the participant is not audibly speaking, an audio feed corresponding to the input audio signal is muted by rendering the audio feed not audible to at least one other participant device connected to the communication session.
Owner:ZOOM COMMUNICATIONS INC

Lightweight sound-absorbing mute cotton and preparation method thereof

The invention provides lightweight sound-absorbing mute cotton and a preparation method thereof, the lightweight sound-absorbing mute cotton is of a bionic gradient structure and is divided into three layers, the bottom layer dissipates high-frequency sound waves, the middle layer optimizes intermediate-frequency resonance, and the upper layer enhances low-frequency sound wave scattering. The preparation method comprises the following steps: respectively preparing bottom-layer slurry, middle-layer slurry and upper-layer slurry; pre-cooling the bottom of a mold, and carrying out synchronous three-layer co-injection molding on the three kinds of slurry: carrying out one-way freezing to form a longitudinal gradient aperture; freezing and drying to obtain a three-layer gradient structure material; immersing in a graphene oxide quantum dot ethanol solution for vacuum treatment; heating, adding ascorbic acid, keeping for 2 hours, and applying 0.5% pre-strain to the material; photo-thermal curing is performed; and performing hot-pressing shaping to obtain the product. The gradient structure design is adopted, and the aperture, the fiber distribution and the filler type are regulated and controlled, so that acoustic impedance gradient change is realized, sound wave reflection is reduced, and the sound absorption efficiency is improved.
Owner:NANTONG HENGJIA HOME TECH CO LTD

Connected wire box with sound insulation function

The utility model discloses a conjoined junction box with a sound insulation function. The conjoined junction box comprises a first junction box, a second junction box and a sound insulation body, the sound insulation body is arranged between the first junction box and the second junction box; the shells of the first junction box and the second junction box are also wrapped with a third noise reduction layer; and a fixing piece is arranged outside the one-piece wire box. According to the invention, by arranging the one-piece junction box in the household wall, the problems of repeated construction and noise generated in the construction process when the junction box is independently arranged in the household wall can be avoided, the damage to the structure of the household wall in the construction process is reduced, and in addition, the sound insulation body is arranged between the first junction box and the second junction box, so that the sound insulation effect is improved. In the process of installing the junction box on the household wall body, the mute effect in the installation process of the junction box at the same position between household splitting can be achieved, and the obvious noise reduction effect in the use process of the junction box in the household wall can be achieved.
Owner:SHANDONG DAWEI INT ARCHITECTURE DESIGN CO LTD

Intelligent sentence segmentation active speech detection method and device based on multi-state temporal modeling

This application discloses an intelligent method and apparatus for detecting active speech with sentence segmentation based on multi-state temporal modeling. The method includes: receiving audio signals from at least one channel; extracting acoustic feature sequences from the audio signals using a target speech recognition model corresponding to the number of channels; determining the probability distribution of each speech frame corresponding to the acoustic feature sequences belonging to different speech activity states, obtaining a state sequence corresponding to each channel, wherein the speech activity state includes at least one of the following: initial silence state, speech state, intra-turn pause silence state, and inter-turn sentence segmentation silence state; and determining the time of sentence segmentation in the audio signal based on the state sequence. This application solves the technical problem of erroneous sentence segmentation in speech activity detection based on a fixed silence threshold in related technologies.
Owner:CHINA TELECOM ARTIFICIAL INTELLIGENCE TECHNOLOGY (BEIJING) CO LTD

Automated segmentation and transcription of unlabeled audio speech corpus

A method includes obtaining initial transcription for input natural speech; performing segmentation of initial transcription into text portions, based on punctuation marks in initial transcription; determining segment-level timestamps for text portions based on the input natural speech; performing audio segmentation on input natural speech, by cutting input natural speech based on segment-level timestamps, to obtain audio chunks; generating transcription portions for each of the audio chunks; merging transcription portions to form re-transcription; determining word-level timestamps for re-transcription, by aligning input natural speech against re-transcription; calculating silence time periods, each corresponding to silence between each two adjacent words of input natural speech, based on word-level timestamps; performing a final segmentation on input natural speech and re-transcription, based on silence time periods, to generate final audio segments and corresponding final transcription portions. The final audio segments and corresponding final transcription portions may be included in training dataset for training a model.
Owner:ORACLE INT CORP

Low-noise intelligent oxygen cabin

The utility model discloses a low-noise intelligent oxygen cabin. The low-noise intelligent oxygen cabin comprises a cabin body, a cabin door, an inner isolation plate, an integrated control part, a monitoring sensor group and an operation part, the cabin body comprises a hollow shell made of sound insulation materials. The cabin door is made of sound insulation glass and can be installed at an access port of the cabin body in an openable mode. A sound attenuation and noise reduction isolation layer made of a sound absorption material is arranged in the inner isolation plate, and the inner isolation plate is detachably connected to the connecting step of the cabin body in a sealed mode; the monitoring sensor group monitors pressure, temperature and oxygen-enriched concentration data in real time and feeds back the pressure, temperature and oxygen-enriched concentration data to the integrated control part, and the integrated control part controls the oxygen generation module, the pressurization module and the temperature control module to work or start and stop, so that the silent oxygen chamber is at the set temperature, pressure and oxygen-enriched concentration. The utility model has the advantages of low noise, silence environment, immersion treatment of personnel in the cabin, constant-temperature, constant-pressure and constant-oxygen medical environment of the silence oxygen chamber, and good medical effect.
Owner:OXYGEN HEALTH TECHNOLOGY (JIANGSU) CO LTD

Determination device

To predict whether there is no sound in a predetermined space based upon whether onomatopoeia indicative of silence that a human perceives in a quiet environment is generated.SOLUTION: A determination device comprises: an input part which inputs predetermined sound data generated by converting a predetermined sound into numeric data through frequency analysis to a determination model outputting whether a specific frequency of 1,800 Hz or more to 2,200 Hz or less indicative of induced otoacoustic emissions is included in addition to the frequency of sound data once the sound data is input; and a control part which displays, on a display unit, the generation of onomatopoeia indicative of silence that a human perceives in a quiet environment when the determination model receiving the predetermined sound data generates an output showing that the specified frequency is detected, but displays, on the display unit, no generation of onomatopoeia when the determination model generates an output showing no detection of the specific frequency.SELECTED DRAWING: Figure 1
Owner:TOYOTA JIDOSHA KK

A method, device, equipment and storage medium for voice endpoint detection

The present application provides a method, apparatus, device and storage medium for voice endpoint detection, wherein the voice endpoint detection method can determine whether the audio frame contained in the audio data to be detected is a silence frame, a noise frame or a voice frame, that is, the present application can detect the relatively accurate attributes of the audio frame contained in the audio data, and perform detection of the voice front-end point and the voice back-end point on this basis, so as to obtain relatively accurate detection results. On the basis of realizing voice endpoint detection, the present application can obtain the recognition text of the voice segment, and can determine the semantic scene of the recognition text based on the semantics of the recognition text, and then can set a suitable post-silence timeout threshold based on the semantic scene of the recognition text, thereby triggering a post-silence timeout event based on the suitable post-silence timeout threshold to improve the user experience.
Owner:UNIV OF SCI & TECH OF CHINA +1

Automatic rail alignment system and method based on AdobeAudition (AU) extension program

The invention relates to an AdobeAudition (AU) extension program-based automatic track matching system and method, and relates to the technical field of audio processing, a picture book analysis module identifies and analyzes a role line sequence and an interval parameter of a line book through a regular expression and a role tag; the audio matching and loading module is used for matching role audios according to a filename rule or a metadata label and loading the role audios to an initial track; the silence detection and marking module adds a mark at a non-silence initial position by adopting framing energy calculation; the line-audio verification module is used for realizing audio and line matching degree analysis through voice recognition to text and font pinyin double-path similarity calculation, and providing a visual calibration interface; and the automatic execution module supports a full-automatic sequential splicing mode or a semi-active sentence-by-sentence splicing mode. The method comprises the steps of line book analysis, audio matching and loading, mute detection marking, verification and calibration and dual-mode execution. The problems of low manual rail alignment efficiency and poor precision in the audio book making process are solved.
Owner:晏海钦

Multi-person voice processing system and separation method based on voiceprint recognition and clustering

The invention discloses a multi-person voice processing system and separation method based on voiceprint recognition and clustering, and belongs to the technical field of multi-person voice separation processing, and the system comprises a voice activity detection unit, an audio cache unit, a voiceprint processing unit and a data processing unit. According to the method, voice activity detection, voiceprint technology, speaker separation technology and caching technology are fused, voice activity detection is performed on a section of voice, non-human voice such as silence, noise and music in the audio is removed, only human voice is reserved, the caching technology is used for caching the audio data, and when the cached data reach a preset threshold value, the voice data are cached. The method comprises the following steps: caching audio data, segmenting the cached audio data according to a certain frame shift and frame length, then extracting speaker features in sub-audio segments by using a voiceprint recognition technology, finally clustering the features, recognizing the audio segment of each person, and outputting corresponding time information.
Owner:HEFEI KEXIN ZHILIAN TECHNOLOGY CO LTD

A role separation-based sales voice dialogue speaker segmentation and labeling method

The present application relates to the technical field of speech processing, in particular to a sales speech dialogue speaker segmentation and marking method based on role separation, comprising the following steps: generating a time-stamped speech segment sequence, synchronously extracting a voiceprint feature vector of each speech segment, starting a silence intention analysis engine, predicting the attribution role of the current silence according to the semantic content of the speech segments before and after the silence, generating a virtual voiceprint feature vector and inserting it into the speech segment sequence; performing role clustering on the voiceprint feature vector, splitting to generate a new speaker cluster; when detecting that there is speech overlap or audio loss in the speech segment, triggering mixed speech separation and generation compensation. The present application improves the abnormal recognition capability in the scene of sales role transformation and sound line disguise, reduces the risk of sales and customer speech confusion, and ensures the coherent and accurate role marking of each speech segment.
Owner:GUANGDONG INFORMATION NETWORK CO LTD

Real-time communication audio equipment detection method based on multiple threads

The invention discloses a multi-thread-based real-time communication audio equipment detection method, which realizes comprehensive coverage of various system environments and driver versions by parallelly starting a plurality of threads corresponding to bottom audio acquisition interfaces APIs of different operating systems, and introduces a mechanism for automatically adjusting the sampling rate, the sound channel number and the sampling format. The problem of silence or abnormal noise caused by mismatching of sampling parameters is effectively solved, and the effectiveness of audio data acquisition is improved. Meanwhile, a neural network voice activity detection model and a comprehensive scoring strategy of time domain statistical indexes are combined, so that the system can accurately recognize effective voice in a complex noise environment, and misjudgment and missed judgment of traditional time domain threshold detection are avoided. The dynamic incremental detection duration strategy balances the detection speed and accuracy, ensures that a user can quickly obtain a preliminary detection result, performs full analysis for abnormal conditions, and improves the detection robustness.
Owner:BEIJING ANXIN ZHITONG TECH CO LTD

Speech punctuation detection method, apparatus, device, storage medium, and program product

The application discloses a speech punctuation detection method, device, equipment, storage medium and program product. The method comprises the following steps: obtaining first speech basic data of a first speech and second speech basic data of a second speech; determining a second silence duration threshold according to the first speech basic data and the second speech basic data; determining a first punctuation detection result of the second speech according to a silence duration of the second speech and the second silence duration threshold, wherein the first punctuation detection result is used for indicating whether the second speech needs to be punctuated. The second speech basic data of the speech that needs to be punctuated and the first speech basic data of the speech that has been punctuated are used to dynamically adjust the silence duration threshold according to the needs, and then the speech punctuation detection is performed according to the second speech basic data and the adjusted silence duration threshold, so that the speech punctuation detection result is obtained, and the accuracy of the speech punctuation detection is improved.
Owner:MASHANG CONSUMER FINANCE CO LTD

Audio data processing method, device, electronic device and storage medium

The present disclosure relates to an audio data processing method, device, electronic device, and storage medium. The method includes: obtaining first audio data to be processed and a storage status of an ear-return buffer area, wherein the ear-return buffer area is used to cache audio data; when the storage status satisfies a preset status, obtaining second audio data corresponding to the first audio data from the stored audio data; generating target audio data based on the first audio data and the second audio data; and processing the target audio data. Since the present disclosure generates target audio data based on the first audio data and the second audio data when the ear-return buffer area satisfies a preset status, and then processes the target audio data, there is no need to insert silent frames or discard audio data, thereby avoiding the introduction of ear-return noise.
Owner:BEIJING DAJIA INTERNET INFORMATION TECH CO LTD

An adjustable frequency response microphone with integrated XLR and USB outputs

This invention discloses an adjustable frequency response microphone integrating XLR and USB outputs, comprising a microphone body, a dynamic microphone driver, a passive filter, a pass-through / preamp circuit, an XLR XLR interface, a USB module, and a function control circuit. The passive filter employs an LC circuit structure, supporting three frequency modes: low-frequency cutoff, mid-frequency boost, and flat response, effectively filtering out ambient noise and enhancing vocal quality. The pass-through / preamp circuit offers both pass-through and preamp modes, with adjustable preamp gain to accommodate dynamic microphone drivers of varying sensitivities, and can be externally or internally powered by phantom power. The USB module integrates a Type-C interface and an audio chip, supporting analog-to-digital / digital-to-analog conversion, real-time headphone monitoring, and external power supply. The function control circuit integrates an encoder and a touch panel, enabling adjustments such as volume, mixing, and mute, and enhances the interactive experience with an RGB lighting module. This invention combines the advantages of professional XLR output with convenient USB output, making it suitable for professional recording, live streaming, gaming, and other scenarios.
Owner:ZHAOQING HEJIA ELECTRONICS CO LTD

Lyrics alignment method based on automatic speech recognition and electronic device

This invention relates to the fields of audio signal processing and artificial intelligence, specifically to a lyrics alignment method and electronic device based on automatic speech recognition. The lyrics alignment method based on automatic speech recognition in this application uses a pre-trained automatic speech recognition model to recognize audio samples and generate recognized text. The edit distance sequence between the recognized text and the lyrics text is calculated, and spoken segments are determined by analyzing the distance change pattern through a sliding window. These spoken segments are then replaced with silent segments, effectively eliminating spoken segments and avoiding interference from non-lyric content in the alignment process, thus laying the foundation for subsequent alignment.
Owner:TIANJIN UNIV

A kind of adjusting device with seat belt tensioning prompt function

ActiveCN224690054USeat beltSound production
The utility model discloses an adjusting device with safety belt tensioning prompt function relates to children seat regulator technical field, including the casing that is composed of shell and backplate, the shell and backplate are fixedly connected through screw between, the button is slid to be provided with on the shell, be equipped with the sound production module between the button and shell, be equipped with the sliding rod below the button, be equipped with the red green display module for showing the state of adjusting device on the sliding rod, the utility model discloses the sound production module is integrated between the button and shell, and the displacement of linkage sliding rod realizes the sound production / mute switching, reaches the effect whether the safety belt is tensioned of auditory and visual double prompt, has improved the use safety and reliability under the scene of dim light or sight obstruction significantly.
Owner:YANGZHOU LETTAS BABY PROD CO LTD