Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

253 results about "Voice analysis" patented technology

Voice analysis is the study of speech sounds for purposes other than linguistic content, such as in speech recognition. Such studies include mostly medical analysis of the voice (phoniatrics), but also speaker identification. More controversially, some believe that the truthfulness or emotional state of speakers can be determined using voice stress analysis or layered voice analysis.

Intelligent nursing record generation method and device

The invention discloses an intelligent nursing record generation method and device, and the method comprises the steps: obtaining nursing voice data, and carrying out the noise reduction and sound enhancement processing to obtain a preprocessed voice stream; a speech recognition technology fusing ECAPA-TDNN and an x vector is adopted to recognize medical terminologies, and a text transcription result is generated; recognizing the voice of the target nurse from the preprocessed voice stream and the transcription result in combination with a speaker-independent model and a speaker condition model to obtain a voice transcription text of the voice; performing voice analysis and structured processing on the text by using a multi-modal self-supervised learning model and a hierarchical CNN-BiLSTM framework to obtain a structured text with clinical semantic features; key clinical information is extracted in combination with the medical ontology knowledge base, and a nursing record element set is obtained; and based on the nursing record element set, automatically generating a nursing record meeting the specification. According to the invention, the automatic generation from the nursing voice to the standardized nursing record is realized, and the efficiency and accuracy of the medical nursing record are improved.
Owner:THE FIRST AFFILIATED HOSPITAL ZHEJIANG UNIV COLLEGE OF MEDICINE

Man-machine physical twin system based on multi-modal video analysis and adaptive mapping and control method

The invention belongs to the technical field of man-machine interaction and robot control, and particularly relates to a man-machine physical twin system based on multi-modal video analysis and adaptive mapping and a control method, and the system comprises a multi-modal collection unit which is used for collecting original video streams of human body actions, expressions and voices, and compensating and repairing dynamic shielding; the semantic analysis unit comprises an action analysis module used for extracting a human skeleton key point set object interaction track from the original video stream, and an expression / voice analysis module used for outputting a facial action unit coefficient and a voice text with an emotion label; the self-adaptive mapping unit is used for establishing human body-robot joint kinematics mapping, distributing control bandwidth in real time according to task types and degrading non-key modes based on network states to guarantee action continuity, and non-contact action capture is achieved through a pure vision scheme by eliminating dependence on a wearable sensor; and the problems of failure and the like of a traditional visual scheme under shielding and illumination variation are solved.
Owner:AIMI (BEIJING) ROBOT CO LTD

Power supply service management method and system based on multi-source data

The invention relates to the field of power supply work order service management, in particular to a power supply service management method and system based on multi-source data. The method comprises the following steps: collecting a current fault repair work order based on a power supply service platform, carrying out deep work order text semantic analysis and standard fault instance modeling, and constructing a global fault instance map; recognizing a client real-time repair call audio stream, carrying out time window voice analysis one by one, and carrying out intelligent fault work order filling, thereby obtaining a real-time audio intelligent fault work order; performing fault demand decomposition on the real-time audio intelligent fault work order and the global fault instance map, performing power supply service resource dynamic configuration regulation and control, and constructing a service resource regulation and control engine; and performing global fault work order real-time processing based on the service resource regulation and control engine, performing overdue work order early warning, and generating a service timeliness early warning strategy. Through dynamic resource allocation and preventive work order creation, the power supply service efficiency, the power grid reliability and the intelligent level are improved.
Owner:ZHANGZHOU POWER SUPPLY COMPANY STATE GRID FUJIANELECTRIC POWER +1

Smart home control method and device for realizing user interaction based on smart mirror

The invention relates to the technical field of smart home, and discloses a smart home control method and device for realizing user interaction based on a smart mirror, and the method comprises the steps: confirming a plurality of user registration instructions, obtaining a registration voice stream based on the user registration instructions, carrying out the feature extraction of the registration voice stream, obtaining reference voiceprint features, summarizing the reference voiceprint features, and carrying out the collection of the reference voiceprint features. The method comprises the following steps: obtaining a plurality of reference voiceprint features, confirming to receive a pre-confirmed voice wake-up word, carrying out voice collection based on the voice wake-up word to obtain a user voice stream, obtaining a voiceprint label by using the user voice stream and the plurality of reference voiceprint features, carrying out voice analysis on the user voice stream to obtain a control instruction, and sending the control instruction to a server. And performing permission verification by using the voiceprint tag and the control instruction to obtain a permission verification result, and obtaining response feedback based on the permission verification result. According to the invention, the problem that the personalized intelligent service cannot be provided because the current user identity cannot be distinguished and the control authority cannot be verified according to the user identity in the smart home interaction scene can be solved.
Owner:DONGGUAN LAIMSEN TECH BUILDING MATERIAL CO LTD

Service compliance dynamic evaluation system based on real-time voice recognition

The invention relates to a service compliance dynamic evaluation system based on real-time voice recognition, and belongs to the cross technical field of artificial intelligence and service compliance management. The system comprises a voice signal enhancement acquisition unit, a voice analysis and risk initial judgment unit, a dynamic evaluation iteration unit and a compliance risk disposal unit. The voice signal enhancement acquisition unit processes low-quality voice and generates a voice data set; the voice analysis and risk initial judgment unit is used for transferring voice in real time, extracting a dialogue intention and identifying violation contents and interactive emotions; the dynamic evaluation iteration unit outputs compliance scores and violation details and is in butt joint with a manual labeling platform; and the compliance risk disposal unit generates a risk label, carries out linkage training and appeal review, and forms a compliance risk disposal scheme. According to the invention, full quality inspection of service scenes in multiple industries such as finance, operators and the like is realized, the voice transfer accuracy and compliance rectification rate are improved, the verbal skill violation complaint rate is reduced, and dynamic monitoring and closed-loop disposal of service compliance are achieved.
Owner:SHANGHAI RONGDA DIGITAL TECH CO LTD

Alzheimer disease recognition method and system based on voice features

The invention relates to the technical field of speech analysis, in particular to an Alzheimer's disease recognition method and system based on speech features, and the method comprises the following steps: extracting frame-level parameters to construct a sequence, recognizing sparse and fractured sections, generating an abnormal trend, and completing speech feature recognition. According to the method, multiple parameters such as the mean value of amplitude absolute values, the maximum difference value and the minimum difference value are serialized and integrated, a double analysis mechanism for the sparsity and the jump of the voice amplitude fluctuation is formed by combining multi-section continuous ratio comparison and mutation trend positioning, the overlapping degree of trend indexes in adjacent frame sections is calculated, and sites in a trend structure are extracted; a trend structure line of the time sequence is established, directional change and point location density of the trend structure line are extracted, quantitative classification of abnormal trends in the frame sequence is completed, cross-scale feature coupling recognition from voice micro fluctuation to time sequence trends is achieved, the discrimination degree and accuracy of voice features of the Alzheimer's disease are improved, and the recognition accuracy of the Alzheimer's disease is improved. And the stability and the discrimination efficiency of the identification result are obviously improved.
Owner:WUXI NO 2 PEOPLES HOSPITAL

Service control method based on voice analysis and large language model

The invention discloses a service control method based on voice analysis and a large language model, relates to the technical field of voice interaction, and solves the technical problems that in the prior art, an enterprise service system needs a user to manually fill in form fields item by item, so that operation is tedious and low in efficiency, and a process is interrupted due to lack of an intelligent completion mechanism due to lack of necessary parameters. According to the invention, a full-closed-loop control chain from the voice instruction to the service operation is constructed, the intention of the user voice input is analyzed through the fine-tuning large language model, and the service interface is dynamically matched, so that the system function operation is directly driven by the voice, and a two-way cooperation mechanism of voice input-system operation-result feedback is formed; and meanwhile, a dynamic parameter completion technology is adopted, the interface field needing to be filled is automatically verified in the parameter extraction stage, missing parameters are completed through context completion or voice guidance, the operation bottleneck of traditional form item-by-item filling is eliminated, missing of the field needing to be filled is avoided, and the service execution efficiency and the process robustness are remarkably improved.
Owner:CEEC ANHUI ELECTRICAL POWER CONSTR NO 1 CO

Vehicle-mounted image generation device, in-vehicle infotainment system and vehicle

The invention relates to the technical field of vehicles, and discloses a vehicle-mounted image generation device, a vehicle machine system and a vehicle. The method comprises the following steps: at least acquiring first point cloud data and first image data through a data fusion module, and converting the two data into a bird's-eye view feature map through a preset conversion algorithm; performing semantic analysis on the user voice through a preset semantic analysis model by using a voice analysis module; generating a vehicle-mounted application image at least applied to head-up display and / or screen display according to the bird's-eye view feature map and the semantic analysis result by using an image generation engine through a preset compression model; and terminating the generation process of the vehicle-mounted application image in a preset time period after the automatic emergency braking signal is received through the emergency interruption module, and at least applying the graphic processing computing power of the generation process to a vehicle collision avoidance decision. Therefore, the problem that in an existing vehicle-mounted image generation technology, due to the fact that the computing power is too large, vehicle decision making is not timely under the emergency situation can be solved, and the driving safety of a user can be guaranteed.
Owner:CHINA FAW CO LTD +1

Vocal music training method and system based on artificial intelligence

The invention relates to the technical field of voice analysis, and particularly discloses a vocal music training method and system based on artificial intelligence, and the method comprises the steps: collecting and analyzing the body posture parameters of a target vocal music training person, carrying out the preprocessing, detecting the posture abnormality, forming a first abnormal training set, capturing and analyzing the pronunciation feature parameters, and recognizing the pronunciation abnormality, thereby obtaining a vocal music training result; and integrating the first abnormal training set and the second abnormal training set to generate a total abnormal parameter training set, matching a correction set, and visually displaying the body posture abnormal degree and the abnormal reason and the pronunciation abnormal degree and the reason, so that the target vocal music training personnel can visually see own posture and pronunciation problems, and self-adjustment and correction are facilitated.
Owner:SHANGLUO UNIV

AI psychological counseling method and system integrating multi-modal sentiment analysis and own intelligence

The invention discloses an AI psychological counseling method and system based on multi-modal sentiment analysis and body intelligence fusion. The method comprises the steps that text, voice and image feature vectors of a patient are extracted; aligning feature dimensions through linear projection, adding space / time sequence position codes, inhibiting noise in combination with cross-modal attention and a gating weighting mechanism, hierarchically fusing text, image and voice features, and generating fusion features; based on facial actions, postures and voice analysis, identifying seven-estrus states of traditional Chinese medicine, outputting emotion probability prediction distribution of a patient, and performing element-by-element product on the emotion probability prediction distribution and the prediction distribution to obtain a video generation vector; a video frame is generated by using StyleGAN-V, the consistency of content and emotion is constrained through emotion driving loss, and a personalized virtual human image is configured. The system comprises a data acquisition module, a feature extraction module, a multi-modal fusion module, an emotion recognition module and a video generation module. According to the method, a closed loop of'emotion analysis-emotion quantification-video generation 'is constructed, and the accuracy and universality of remote psychological consultation are improved.
Owner:杨雪飞

Spoken English pronunciation quality evaluation method based on multi-mode speech feature analysis

ActiveCN121528247ASpeech analysisFeature extractionModal voice
The invention belongs to the technical field of speech analysis, and discloses a spoken English pronunciation quality evaluation method based on multi-modal speech feature analysis, which comprises the following steps: acquiring a spoken English speech signal of a target user, and performing multi-domain decomposition on the speech signal to obtain multi-modal speech features; performing time-frequency domain corresponding relation analysis and feature extraction on the multi-mode speech features to obtain a pronunciation detail feature sequence; performing multi-scale matching on the pronunciation detail feature sequence and a preset standard pronunciation template, constructing a multi-dimensional representation model based on a multi-scale matching result, and calculating a fine-grained quality score of a phoneme unit in each multi-dimensional representation model in combination with rhythm and rhythm parameters in the multi-modal speech features, the rhythm coherence score and the overall fluency score are fused to generate a comprehensive pronunciation quality evaluation result and a visual diagnosis report of pronunciation deviation; the oral English pronunciation evaluation method realizes comprehensive and refined evaluation of oral English pronunciation, and provides a scientific guidance basis for personalized language learning.
Owner:ZHANG ZHOU HALTH VOCATIONAL COLLEGE

Video generation method and system based on voice analysis, and storage medium

The invention relates to the technical field of artificial intelligence and multimedia, and particularly discloses a video generation method based on voice analysis, and the method comprises the following steps: analyzing an input voice, and extracting a multi-modal voice feature; inputting the multi-modal speech features into a pre-trained scene association model, and outputting a scene label set; on the basis of dialect features in the input voice, region categories are recognized through a dialect classifier, and a corresponding visual element library is loaded from a culture database according to the region categories; selecting a scene template according to scene type labels in the scene label set, and selecting a character action template in combination with emotion type labels and interaction object relation labels; and calculating time sequence distribution of the video elements based on the speech speed change parameters, and rendering the scene template, the character action template and the rhythm of the input speech according to the time sequence distribution through a time sequence alignment algorithm to generate a target video. The method can improve the accuracy of the video generated based on the voice.
Owner:SHENZHEN SMART INSURANCE TECH CO LTD

Robot control method based on voice analysis

The invention relates to the field of voice analysis, in particular to a robot control method based on voice analysis, and the method comprises the steps: obtaining image information in a space region where a target robot is located, determining semantic tags, determining a potential association semantic tag group, and screening out a semantic guiding corpus group for the target robot; when voice control data is received, instruction fuzzy parameters of the voice control data are determined so as to judge instruction fuzzy tendency, optimization is carried out on the voice control data with the instruction fuzzy tendency, specifically, confidence centralized clusters and semantic discrete clusters are determined, semantic expansion is carried out on the semantic discrete clusters, and then an expansion text is obtained; and screening the expanded text based on the semantic oriented corpus group to obtain a confidence instruction text. According to the method, the semantic oriented corpus group is constructed in combination with the image information, the analysis of the voice control data with the instruction fuzzy tendency is guided, the analysis accuracy of the voice control data under the semantic fuzzy condition is improved, and the control instruction recognition precision is ensured.
Owner:厦门工学院

Gas cylinder filling method, early warning method, checking method, system and medium

The invention relates to a gas cylinder filling method, an early warning method, an assessment method, a system and a medium in the technical field of industrial safety production. The gas cylinder filling safety assessment method based on voice analysis comprises the following steps: acquiring acoustic data of an operator; and analyzing the acoustic data and calculating to obtain a knowledge proficiency score Spro. And acoustic data are extracted, and a psychological quality score Semo is obtained through calculation. Collecting motion posture data of the gas cylinder and calculating to obtain an operation stability score Ssta; and carrying out weighted fusion on the Stro, the Semo and the Ssta to obtain an assessment score Stotal. And setting a score line based on the assessment requirement, and if the Stotal is greater than or equal to the score line, determining that the operator passes the assessment. According to the invention, through innovative multi-mode assessment, the ability of the operator is changed from feeling-based ability to data-based ability, the final result can be known, and the specific weak link can be known according to the specific condition of each link, so that the psychological quality and emergency response ability of the operator can be effectively assessed.
Owner:ANHUI SPECIAL EQUIP INSPECTION INST

Transparent HUD interaction system based on event-frame-electro-oculogram three-source fusion

The invention discloses a transparent HUD interaction system based on event-frame-electro-oculogram three-source fusion. The transparent HUD interaction system comprises an event camera, an electro-oculogram sensor and a frame type camera. The voice analysis module is used for collecting a voice instruction, extracting a specified interface element in the voice instruction as a voice interaction target and generating a voice confidence coefficient; the three-source fusion module is used for fusing the eye movement event flow, the eye electric signal and the eye image frame to obtain a fixation vector; the prediction module is used for predicting a gaze vector and gaze confidence at the next moment according to the gaze vector sequence in the recent preset time period; the interaction target confirmation module is used for acquiring a sight line interaction target, and confirming the interaction target when the sight line interaction target is consistent with the voice interaction target and the product of the voice confidence coefficient and the gaze confidence coefficient is greater than or equal to a preset gating threshold value; and the transparent display module is used for projecting the interactive interface on a vehicle-mounted transparent interface and displaying the real-time interactive information. The method is higher in recognition accuracy.
Owner:SOUTHEAST UNIV

Natural sound interference stripping algorithm and application thereof in noise monitoring

The invention relates to the field of speech analysis, and discloses a natural sound interference stripping algorithm and an application thereof in noise monitoring, which are used for efficiently and accurately stripping interference components in a target speech signal in a complex natural sound environment. Comprising the following steps: synchronously acquiring an environmental sound signal by using a spatially distributed microphone array, generating an original mixed voice signal with spatial coherence, performing real-time fractional order differential processing, strengthening resonance characteristics, realizing spatial separation of a voice guided wave mode and an interference radiation mode through real-time phase analysis, and extracting an interference radiation mode component. And reconstructing a time domain interference waveform, and outputting a pure target voice signal according to the original mixed voice signal and the reconstructed interference waveform. The system has the functions of residual interference sound energy recovery, equipment vibration suppression, environment sudden change monitoring, system resetting, bimodal output and the like, the stability, reliability and practicability of the system are effectively improved, and pure voice can be accurately extracted in a complex environment.
Owner:JIANGSU ENVIRONMENTAL MONITORING CENT

Self-adaptive voice semantic communication method based on hierarchical time sequence importance

The invention relates to a self-adaptive voice semantic communication method based on hierarchy time sequence importance, which belongs to the field of voice analysis, and comprises the following steps: a sending end performs priority division on a discrete feature matrix according to hierarchy and time sequence attributes of voice features, and screens the features according to importance scores in a feature matrix packaging and selecting link; the method comprises the following steps: interacting with wireless communication through a channel adaptive scheduling module, introducing a channel feedback mechanism, and dynamically adjusting transmission power and a resource allocation strategy by sensing current channel state information; and after completing signal demodulation, a receiving end inputs the acquired sparse feature flow into a voice restoration module, and globally reconstructs the received features by using voice priori knowledge of deep pre-training. According to the method, a hierarchical speech feature extraction technology based on a discrete codebook and a generative semantic repair technology are combined, so that the speech word error rate in a severe channel environment is reduced to the maximum extent, and the semantic intelligibility of a receiving end is improved.
Owner:UESTC (SHENZHEN) ADVANCED RES INST

Driving emotion early warning method and system based on voice analysis

The invention provides a driving emotion early warning method and system based on voice analysis, relates to the technical field of voice emotion recognition, and effectively solves the interference problem of vehicle-mounted dynamic strong noise by deploying a microphone array and a noise sensor in a cockpit and combining multi-channel noise reduction algorithms such as beam forming and linear beam minimum variance. A high-quality voice instruction is extracted from a source; on the basis, the driving intention of the clear voice signal is accurately recognized by using the deep neural network, so that the accuracy is remarkably improved; more importantly, according to the method, the instability Ls and the high interactive friction probability Pf are recognized innovatively through calculation instructions, real-time quantitative evaluation on the man-machine interaction fluency is achieved, when the interactive friction probability Pf exceeds a preset threshold value, the system can automatically trigger self-adaptive strategy adjustment, and therefore the situation that a driver is distracted due to recognition errors is avoided, and user experience is improved. And the interaction experience and the driving safety are greatly enhanced.
Owner:FUJIAN LUOYUAN COUNTY SENIOR VOCATIONAL HIGH SCHOOL

AIBox embedded order sending control method and system for intelligent operation and maintenance

PendingCN120412576ASpeech recognitionInference methodsMel-frequency cepstrumEmbedded technology
The invention relates to the field of embedded technology, and provides an AIBox embedded order sending control method and system for intelligent operation and maintenance, and the method comprises the steps: responding to a repair request initiated by a user through a voice gateway; the AIBox receives and preprocesses the digital voice signal, extracts audio features by using a Mel frequency cepstrum coefficient technology, and performs sequence alignment on the extracted audio features through a dynamic time warping algorithm to convert the audio features into a text; key information fields in the text are extracted through the NLP technology, and a repair work order is automatically generated and verified; and the priority is automatically distributed according to the emergency degree, an automatic distribution mechanism is introduced, the repair work order is distributed to the corresponding maintenance personnel, and the repair work order is closed and archived according to the completion feedback of the maintenance personnel. According to the invention, seamless connection between AI voice analysis and an event and repair service process is realized, an AI intelligent voice algorithm module is packaged in AIBox and is connected with a voice data interface of an intelligent operation and maintenance repair call, voice content is automatically identified, and a repair work order is generated and distributed.
Owner:SHANGHAI LINGANG JINGHONG SECURITY TECH DEV CO LTD

System and method for voice analysis

The invention provides a system for analysing vocal signals from an individual, said system comprising an acoustic sensor configured to have a measured frequency response in the ultrasonic range; mean
Owner:THE SEC OF STATE FOR DEFENCE IN HER BRITANNIC MAJESTYS GOVERNMENT OF THE UK OF GREAT BRITAIN & NORTHERN IRELAND

Intent inference in audiovisual communication sessions

In one aspect, a user's intent can be inferred based on voice analysis during a communications session, and prompts can be presented, or other actions taken, at least partly in response to the inferred intent. For example, a network microphone device (NMD) having one or more microphones can capture voice input and transmit the voice input to remote computing device(s) for a communication session (e.g., a videoconference). The NMD can analyze the voice input to detect one or more utterances. Based on the utterance(s), the NMD can cause a user prompt to be displayed via a display device communicatively coupled to the NMD. The particular prompt can depend at least in part on one or more context parameters associated with the communication session (e.g., a microphone state of one or more users, a screen share state of one or more users, or a recording status of the session, etc.).
Owner:SONOS INC

System

An object of a system according to an embodiment is to enable an elderly person to easily access necessary information and services.SOLUTION: A system includes a voice input unit, an analysis unit, a guidance unit, and a connection unit. The voice input unit acquires voice. The analysis unit analyzes the voice acquired by the voice input unit. The guidance unit provides appropriate information based on the content analyzed by the analysis unit. The connection unit connects to a necessary service based on the information provided by the guide unit.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

System

An object of the system according to the embodiment is to efficiently create minutes of a conference, provide opinions, and deal with foreign languages.SOLUTION: A system according to an embodiment includes a voice analysis unit, a minutes generation unit, an advice providing unit, and a foreign language handling unit. The voice analysis unit analyzes voice data during a conference. The minutes creating section creates minutes based on the voice data analyzed by the voice analyzing section. The advice providing unit provides advice or an opinion on the basis of the minutes generated by the minutes generation unit. The foreign language correspondence unit translates the minutes generated by the minutes generation unit into a foreign language.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

Mental health appointment face inquiry system and method based on multi-role cooperation

The invention discloses a psychological health appointment face inquiry system and method based on multi-role collaboration. The system comprises a visitor module, an appointment person module, a face inquiry person module, an administrator module and a data management module. The visitor module is provided with an emotion recognition sub-module which is used for analyzing the emotion tendency and the emergency degree described by the appointment text; the face inquiry module is integrated with a cross-modal biological feedback sub-module, physiological indexes, eye movement parameters and acoustic characteristics are collected in real time through intelligent wearable equipment, an eye tracker and voice analysis, and a comprehensive emotional state index is generated through a data fusion processor. And a polling teacher carries out consultation based on the biofeedback data, and an administrator monitors system operation. According to the method, multi-role efficient collaboration is realized, and the accuracy, the response speed and the service quality of psychological health services are remarkably improved through intelligent emotion recognition, multi-modal biofeedback and adaptive process scheduling.
Owner:JIANGSU ZHUODUN INFORMATION TECH CO LTD

system

PendingJP2026105311ADialog systemVoice analysis
We provide the system. [Solution] A receiving method that accepts language selection, A generation means for generating learning content based on a generative AI model, An acoustic conversion means that presents the generated learning content as an audio output, A voice analysis method that converts user voice information into text data, A management system that continues the dialogue based on user responses and provides feedback, A dialogue system that includes this.
Owner:SOFTBANK GROUP CORP

Methods and system for distributing information via multiple forms of delivery services

A content distribution facilitation system is described comprising configured servers and a network interface configured to interface with a plurality of terminals in a client server relationship and optionally with a cloud-based storage system. A request from a first source for content comprising content criteria is received, the content criteria comprising content subject matter. At least a portion of the content request content criteria is transmitted to a selected content contributor. If recorded content is received from the first content contributor, the first source is provided with access to the received recorded content. The recorded content may be transmitted via one or more networks to one or more destination devices. Optionally, a voice analysis and / or facial recognition engine are utilized to determine if the recorded content is from the first content contributor.
Owner:GREENFLY

A breath-speech pause signal analysis system and method

The present application relates to the cross technical field of speech signal processing, respiratory physiological monitoring and artificial intelligence, and specifically discloses a respiratory-speech pause signal analysis system and method. The present application synchronously collects speech and respiratory signals, extracts language-adapted pause parameters and physiological indexes, performs time sequence alignment and correlation analysis, and then fuses a prediction model and a true-false recognition model for intelligent analysis, thereby solving the problems that traditional heart-lung function detection relies on professional equipment and cannot be remotely and non-contactly monitored, and that existing speech analysis lacks a physiological coupling mechanism, leading to the inability to identify synthetic speech and insufficient cross-language adaptation, and realizing the dual ability improvement of non-contact respiratory physiological state evaluation and speech fraud recognition.
Owner:ZHONGDE NUOHAO (BEIJING) EDUCATION TECH CO LTD

system

We provide the system. [Solution] A voice analysis method for acquiring voice data in real time and identifying the characteristics of the speaker, A learning method for analyzing past conversation history and learning conversation patterns, A means of providing and displaying interactive games to participants, A means of managing information to collect and organize family events and health information, A means of distributing information to notify user terminals of organized information. Includes system.
Owner:SOFTBANK GROUP CORP

Hearing device comprising an own voice estimator

Disclosed herein are embodiments of a hearing device including at least one first, outward-facing, input transducer configured to pick up first sounds from the environment of a user and a second, inward-facing, input transducer configured to pick up a second sounds at the eardrum of the user. The hearing device can further include a directional system including a) an own voice beamformer configured to provide an estimate of the user's own voice in dependence of the at least one first and the second electric input signals and configurable own voice beamformer weights; and b) an own voice analyzer configured to analyze at least one of the at least one first and said second electric input signals, or to analyze a signal or signals originating therefrom, and to provide an own voice beamformer weight control signal.
Owner:OTICON