Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

5 results about "Normal speech" patented technology

Normal speech between two people typically has a range of 50 to 60 decibels. When two people are speaking in a public place with background noise, normal speech is louder, around five extra decibels.

A deep learning-based teaching quality evaluation method and system

The present application belongs to the technical field of intelligent teaching, and particularly relates to a teaching quality evaluation method and system based on deep learning. The method comprises extracting feature data to be evaluated from normal speech, evaluating the feature data to be evaluated by using a speech evaluation model to generate an evaluation result, and determining the quality grade of the normal speech according to the evaluation result, wherein the feature extraction from noise speech and noise speech comprises extracting amplitude information and frequency information from a sound production section. The present application performs screening on teaching speech to identify abnormal sound sections; through voiceprint comparison, the teaching speech in which the abnormal sound sections that can match the pre-stored voiceprint are classified as noise speech, and the teaching speech that cannot be matched is classified as noise speech, so as to distinguish the noise speech originating from the background environment from the noise speech originating from the teaching subject, overcome the evaluation error problem caused by regarding the two as noise without distinction, and lay a data foundation for subsequent evaluation.
Owner:CNSCI SOFT EDUCATIONAL TECH (BEIJING) CORP

A false trigger suppression method for a speech recognition system

PendingCN122417020ACarrier signalAcoustics
This invention relates to the field of speech recognition technology, specifically to a method for suppressing false triggering in a speech recognition system. This method acquires and segments candidate wake-up speech signals, performs audible acoustic analysis and high-frequency carrier anomaly analysis on each speech segment to be verified, and calculates a command validity score by combining cross-segment consistency, audible injection risk, and target secondary discrimination threshold. Based on this, false triggering speech signals are suppressed. This invention combines candidate recognition, segment verification, audible acoustic analysis, and high-frequency carrier anomaly analysis, and comprehensively considers cross-segment physical consistency, audible injection risk, and target secondary discrimination threshold to generate a command validity score, thereby improving the accuracy, security, and reliability of false triggering recognition and normal speech command response.
Owner:BEIJING ZHONGWANG BOCAI TECHNOLOGY CO LTD

Voice processing method and electronic device

Embodiments of the present application disclose a speech processing method and an electronic device, and relate to the technical field of information processing. The method comprises: obtaining original speech; preprocessing the original speech to obtain a plurality of original speech features; in the case where the plurality of original speech features comprise a first speech feature, extracting a voiceprint feature from the first speech feature to obtain a first voiceprint feature, the speech type corresponding to the first speech feature being a whisper speech type; determining a target voiceprint feature corresponding to the first voiceprint feature from a target voiceprint feature library; and performing whisper speech conversion on the first speech feature based on the target voiceprint feature to obtain converted normal speech. According to the present application, the same tone as the real tone of the user or the tone specified by the user can be restored in the whisper speech conversion process, thereby improving the user experience.
Owner:HONOR DEVICE CO LTD

How to automatically switch between mesh calls and 5G data network calls

This application relates to a method for automatically switching between mesh calls and 5G data network calls. Specifically a Bluetooth device constructs a mesh network to form a mesh group One of the nodes is used as the master node, and an online loop is created on the AP P. All Bluetooth devices communicate with the application in a two-way dynamic heartbeat data mode, and the cloud synchronizes information in the virtual network formed by mapping the nodes; The voice data between online nodes is communicated in mesh mode; when the nodes of Bluetooth devices are offline or reconnected, the node device automatically switches to the data network mode through the application and performs network communication using the data network mode via the cloud. This automatic switching method between mesh calls and 5G data network calls monitors the online status of the device in real time and automatically switches the voice data transmission to the cloud network when the device / node goes offline, thereby ensuring the normal voice data transmission and reception of non - communicating nodes and ensuring the real - time nature, continuity and completeness of voice data transmission. ​
Owner:FUKA AIKOSHI INTELLIGENT TECHNOLOGY CO LTD

Information processing device, information processing method, computer program, learning device, remote conference system, and support device

Provided is an information processing device that perform processing related to speech conversion of a speech that is not normally uttered and does not include pitch information such as a whisper or a faint speech. The information processing device includes a speech-to-unit encoder that generates an acoustic unit from a speech waveform, and a unit-to-speech decoder that reconstructs a speech waveform from an acoustic unit. The unit-to-speech decoder is subjected to preliminary learning by self-supervised learning of a Masked Language Model type using a normal speech and a whisper without a text label of a specific speaker to generate an acoustic unit common to the normal speech and the whisper, the acoustic unit being a latent expression in which a difference between the normal speech and the whisper is absorbed.
Owner:SONY GROUP CORP