Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

721 results about "S Voice" patented technology

S Voice is an intelligent personal assistant and knowledge navigator which is only available as a built-in application for the Samsung Galaxy S III, S III Mini (including NFC Variant), S4, S4 Mini, S4 Active, S5, S5 Mini, S II Plus, Note II, Note 3, Note 4, Note 10.1, Note 8.0, Stellar, Mega, Grand, Avant, Core, Ace 3, Tab 3 7.0, Tab 3 8.0, Tab 3 10.1, Galaxy Camera, and other 2013 or later Samsung Android devices. The application uses a natural language user interface to answer questions, make recommendations, and perform actions by delegating requests to a set of Web services. It is based on the Vlingo personal assistant.

Multi-mode-based AI digital human intelligent interaction method, system and equipment

The invention relates to the technical field of computer vision and human-computer interaction, and discloses an AI digital human intelligent interaction method, system and equipment based on multiple modalities, and the method comprises the steps: pre-awakening a digital human when a human face is detected, and further thoroughly awakening the digital human based on recognized preset voice information or preset gesture information; voice and video information of a user in the interaction process is obtained, a keyword extraction result, a gesture recognition result and an emotional state tag are generated, a pre-constructed knowledge base is utilized to retrieve related information, a big language generation model module is combined to generate an answer text, and the answer text is input into a preset voice synthesis model to generate emotional voice output. And based on the current emotional state label of the user, driving the digital human animation to be output in an emotional manner. According to the method and the system, the digital human for understanding the emotion of the user, generating personalized answers, providing voices with rich emotions and displaying natural expressions and actions can be created, better interaction with the user can be realized, and more humanized and effective services can be provided.
Owner:BEI JING WAN JIE SHU JU KE JI YOU XIAN ZE REN GONG SI WU HAN FEN GONG SI +1

Dynamic self-adaptive multi-modal sentiment analysis fusion method and system

The invention provides a dynamic self-adaptive multi-modal sentiment analysis fusion method and system, and relates to the technical field of multi-modal sentiment analysis. The method comprises the following steps: synchronously acquiring voice, text, facial expression and limb movement data of a target user to form a multi-modal data set; the method comprises the following steps: firstly, extracting emotional characteristics of each mode, and constructing a cross-mode correlation model to capture a collaborative and complementary relationship among different modes; and calculating a real-time confidence score and a complementarity index of each modal based on the weight matrix of the cross-modal correlation model. Then, according to the scores and the indexes, a weighted average or maximum entropy algorithm is dynamically selected to fuse multi-modal emotion features, and a comprehensive emotion feature vector is generated; and finally, inputting the vector into a pre-training deep learning model, and outputting an emotional state category of the user. According to the method, the user emotion can be accurately and comprehensively captured, efficient emotion recognition and classification are realized, and the robustness and flexibility of an emotion analysis system in a complex scene are improved.
Owner:HUNAN OPEN UNIV (HUNAN PROVINCIAL CADRE EDUCATION & TRAINING ONLINE COLLEGE)

Vehicle state monitoring and early warning method and system

The invention belongs to the technical field of vehicle state detection and early warning, and particularly relates to a vehicle state monitoring and early warning method and system.The monitoring and early warning method comprises the steps that a sensor obtains real-time data, an anomaly detection algorithm is applied, and an abnormal event is marked; calculating an information priority according to the abnormal severity and the driving scene, and distributing the information priority to a high-priority queue; multi-mode early warning is generated, and high-frequency sound and vibration are used during high-speed driving; extracting a voice prompt, and generating voice waveform data through a voice synthesis module; voice input of a driver is recognized, intention is analyzed, and abnormal information feedback is provided; according to the scene and the abnormal state, interactive output is optimized, and detailed information is displayed during low-speed congestion; dynamically adjusting the interface of the instrument panel, and amplifying the key area in case of abnormal severity; and integrating the image data and the voice data, generating a multi-mode signal, and transmitting the multi-mode signal after rendering processing. The vehicle abnormal information can be timely and accurately transmitted to a driver, and the driving safety is effectively improved.
Owner:XIAN HUODA NETWORK TECH CO LTD

Voice conversation method and system, electronic equipment, storage medium and program product

The embodiment of the invention provides a voice conversation method and system, electronic equipment, a storage medium and a program product. One method is implemented as follows: a voice dialogue stream (which can be an inquiry voice stream input by a user for a target commodity) of a user is subjected to streaming response, a plurality of dialogue text segments are generated in a segmented manner in the streaming response process, and the dialogue text segments are output in real time after being generated and stored in a first cache queue; and when it is detected that the reply text fragment is stored in the first cache queue, the stored reply text fragment is converted into a reply voice fragment in real time and played. The streaming response may be performed by an agent. Visibly, according to the response of the voice dialogue stream of the user, the dialogue is played while being generated, the waiting time from text to voice synthesis can be greatly shortened, and therefore the response time can be shortened.
Owner:ZHEJIANG TMALL TECH CO LTD

Dynamic health monitoring system integrating artificial intelligence and wearable device and method thereof

The invention provides a dynamic health monitoring system integrating artificial intelligence and wearable equipment and a method thereof, and relates to the field of health monitoring, and the method comprises the steps: collecting physiological parameters and environmental data of a user, eliminating time migration of multi-source data through a timestamp alignment technology, extracting dynamic coupling features of the physiological parameters and the environmental data, and obtaining a dynamic health monitoring result; and generating the health state characterization guided by the environment. The method comprises the following steps: synchronously acquiring user voice input information, and extracting and analyzing acoustic features to generate emotional state change features after performing Mel spectrum conversion; then, cross-modal semantic association between emotions and health states is established through a modal soft constraint mechanism, and personalized guidance suggestions are generated and displayed, so that the dynamic perception ability of health state assessment and the scene adaptability of suggestion generation are improved; therefore, the problems of environmental factor splitting, subjective and objective data isolation and insufficient dynamic nonlinear relation modeling in a traditional method are solved.
Owner:JIANGXI YANGNING TECHNOLOGY CO LTD

Real-time automatic online voice translation system and method for telephone conversations

An automatic online voice translation system and method, comprising a virtual translator, a SIP gateway, and a storage, wherein the SIP gateway receives a call from a caller, when the virtual translator is not activated, the SIP gateway establishes a communication between the caller and the agent by streaming in real-time caller's voice directly to the agent and agent's voice directly to the caller, and when the virtual translator is activated, the SIP gateway creates audio files based on the received voice streams of both the caller and the agent, and sends the created audio files to the virtual translator for translating the caller's voice into a language understood by the agent, before transmitting the voice to the agent. In another embodiment, a contact center platform is used for applying features such as call recording, Interactive Voice Response (IVR) service, and conferencing service to the call.
Owner:SESTEK SES & ILETISIM BILGISAYAR TEKNOLOJILERI TIC & SAN AS

Dialogue interaction state recognition method and system, electronic equipment and storage medium

The embodiment of the invention provides a dialogue interaction state recognition method and system, electronic equipment and a storage medium, and relates to the technical field of voice interaction, and the method comprises the steps: extracting time sequence features and semantic features from obtained voice information; the time sequence feature represents the rhythm change condition of the user in the dialogue process, and the semantic feature represents the voice content and the voice intensity of the user in the dialogue process; recognizing the current dialogue interaction state of the user according to the time sequence features and the semantic features; the dialogue interaction state comprises one of a silent state, an interrupted state and a hesitant state. Therefore, the conversation interaction state of the user is recognized by combining the time sequence features and the semantic features, voice signal level analysis is considered, and changes of voice rhythm, voice content and voice intensity are also considered, so that the conversation interaction state of the user is accurately recognized, and the recognition precision of complex interaction states such as hesitation and interruption is improved.
Owner:SHANGHAI XULU INFORMATION TECHNOLOGY CO LTD

AI-based Intelligent Voice Customer Service Response Method and System for Water Services

The present invention relates to the technical field of intelligent voice customer service response, and a water service intelligent voice customer service response method and system based on AI, including: obtaining a plurality of users, and performing the following operations on each user among the plurality of users: obtaining the user's voice, performing voice signal processing on the user's voice to obtain user voice features, retrieving matching voice features, detecting the feature similarity between the user voice features and the matching voice features, constructing a water service knowledge graph, obtaining inquiry information voice and encrypted user data, obtaining updated voice, using the updated voice as the user's voice, returning to the step of performing voice signal processing on the user's voice, obtaining the number of inquiries of the inquiry information voice, updating the water service voice database to obtain an optimized water service voice database, and completing the water service intelligent voice customer service response based on AI based on the encrypted user data set and the optimized water service voice database. The present invention can improve the efficiency and quality of water service customer service and enhance the user experience.
Owner:SEQUOIA LIBRA TECH GRP CO LTD

Natural voice utility asset annotation system

Systems and methods for annotating, also known as tagging, visually identifiable objects using natural language (i.e. spoken voice) are provided. More specifically, but not exclusively, this disclosure relates to systems and methods for tagging identified objects related to underground utilities and assets, and communication systems. Once tagged, data representing identified objects can be stored, transmitted, and / or mapped. In an exemplary embodiment, a utility service worker or other personnel may walk or drive around an area of interest using utility locating equipment or systems to collect data. As data is being collected, the user may visually identify utility related or other items using a laser pointer, and then use their voice to tag items of interest. A headset including a microphone may be provided for capturing the user's voice. Hardware and / or software may be provided for processing tagged items, and relating them to a specific location, and / or utility asset.
Owner:SEESCAN INC

User intention classification method and device based on multi-model hierarchy, vehicle and medium

The invention relates to the technical field of intelligent recognition, in particular to a multi-model hierarchy-based user intention classification method and device, a vehicle and a medium, and the method comprises the steps: obtaining a voice instruction of a user; decoding and converting the voice instruction into text data, inputting the text data into a pre-constructed intention classification model to obtain a matching result, and recognizing the user intention through a plurality of model layers in the intention classification model based on a confidence interval of the matching result to obtain a user intention recognition result; the target control instruction is determined according to the user intention recognition result, the vehicle is controlled to execute the corresponding operation according to the target control instruction, and the operation result is fed back, so that the problems of low user intention recognition speed and low user intention recognition accuracy in related technologies are solved, the user intention recognition speed is increased, and the user intention recognition efficiency is improved. And the accuracy of user intention recognition is improved.
Owner:BEIJING AUTOMOBILE RES GENERAL INST

Real-time voice interaction and adaptive content generation system based on RTC and AIGC

The invention belongs to the technical field of artificial intelligence, and particularly relates to a real-time voice interaction and adaptive content generation system based on RTC and AI GC, which comprises the steps of capturing voice input of a user through a microphone of equipment, and converting the voice input into text information by using a voice recognition technology; the recognized text input is transmitted to a natural language understanding module, semantic analysis is carried out, and user intention and information are recognized; according to the intention and demand of the user, the AI GC technology is used for dynamically generating personalized content; and the generated content is optimized in real time according to the feedback in interaction, the demand change of the user is adapted, the generated text content is converted into voice to be output and provided for the user, and the effects of meeting the demands of different types of users and promoting real-time interactive propagation and automatic content generation are achieved.
Owner:SHENZHEN UASCENT TECH CO LTD

Voice control ultrasonic wave adjustment setting method based on CTC-Attention mixed architecture

The invention relates to a voice control ultrasonic wave adjustment setting method based on a CTC-Attention mixed architecture, and the method comprises the following steps: S1, receiving a voice instruction of a doctor through a voice receiving device, and carrying out the preprocessing of an original voice signal; s2, constructing an acoustic model of a CTC-Attention hybrid architecture, and training the acoustic model to obtain a trained acoustic model; s3, inputting the preprocessed voice signal into the trained acoustic model, and performing voice recognition through CTC path probability alignment and Attention context dependence joint optimization; s4, outputting an identification result and decoding the identification result into a text instruction; and S5, analyzing the operation intention according to the text instruction, and controlling the ultrasonic instrument to execute a corresponding state adjustment operation. According to the method, the voice signal of the operator is received, the voice signal is recognized and analyzed, the obtained result is optimized and then converted into the character content with the meaning, then the ultrasonic energy with the intensity corresponding to the character content is generated by the ultrasonic instrument based on the character content, and the ultrasonic instrument is effectively used for medical operation.
Owner:SHUGUANG HOSPITAL AFFILIATED WITH SHANGHAI UNIV OF T C M

Self-media content streaming matching method and system based on dynamic semantic analysis

The invention provides a self-media content streaming matching method and system based on dynamic semantic analysis, and the method comprises the steps: carrying out the dynamic semantic analysis of the text content of a voice stream of a creator, generating a semantic theme sequence, aligning the semantic theme sequence with an emotion fluctuation curve of the voice stream of the creator, and generating an emotion semantic incidence matrix; dividing a mapping relationship between an emotional intensity numerical range in the emotional fluctuation curve and a semantic topic type in the semantic topic sequence; performing behavior association fitting on the mapping relationship based on user historical feedback behavior data, and generating a dynamic association rule between the emotion intensity numerical range and the user feedback behavior; and adjusting a preset self-media content delivery strategy in real time, and generating an adjusted delivery strategy so as to carry out self-media content delivery stream matching. According to the method, the emotion semantic association matrix is constructed, accurate space-time matching of the content theme and the emotion expression is realized, and the target of adaptively optimizing the flow casting matching according to the real-time emotion state of the creator is achieved.
Owner:SHANGHAI YUXING CULTURAL COMMUNICATION CO LTD

Remote virtual visitation and information exchange

The present invention relates to a communication system and method designed for patients in clinical settings who face challenges using standard communication devices. This system includes a mobile application and associated devices that establish a secure, asynchronous communication channel between clinicians, patients, and their support networks. Authorized users can record and send voice messages and other media, which are automatically played on the patient's device without requiring active operation. The system is tailored for ease of use, accommodating patients with severe physical or cognitive limitations through hands-free operation and simple commands. It adapts to the patient's condition, pausing messages during rest or medical procedures and resuming when appropriate. The system also features cross-linguistic messaging, translating and synthesizing voice messages across languages while maintaining the speaker's voice. Additional features include support for rich media, automatic transcription, and AI-driven content analysis. Secure, encrypted channels ensure privacy, complying with healthcare regulations. This invention enhances emotional connection and support for isolated patients, acting as a virtual visitation tool.
Owner:VOICELOVE LLC

Game assisting method and device

The invention provides a game assisting method and device.The method comprises the steps that in the process that a user plays a game, voice of the user is recognized, the voice is converted into a text, and a game image corresponding to the text is matched and recorded as a target game image; retrieving a preset game strategy knowledge base by utilizing the target game image to obtain target game strategy information; the target game image, the text and the target game strategy information are submitted to the large-scale multi-mode model for analysis, the instructive text is generated, the instructive text is synthesized into the voice and played, and a game assisting scheme which does not affect the game experience of the user in the game playing process of the user and efficiently and effectively guides the user is provided.
Owner:HAIMA CLOUD TIANJIN INFORMATION TECH CO LTD

Intelligent customer service voice interaction method and system based on voice recognition

The invention relates to the technical field of voice recognition, in particular to an intelligent customer service voice interaction method and system based on voice recognition, and the method comprises the steps: collecting voice signals of a user in a voice instruction issuing process in a voice interaction scene in real time, and text data generated when the user interacts in a historical period, and forming a historical text set; dividing the voice signal into a plurality of signal segments; obtaining each instruction signal segment; obtaining a phoneme sequence corresponding to each instruction signal segment and all candidate texts corresponding to the phoneme sequence; calculating a matching degree and a semantic association degree of each vocabulary in each candidate text to obtain a candidate probability of each candidate text; and correcting the candidate probability to obtain a corrected candidate probability corresponding to each candidate text under each instruction signal segment, identifying a voice signal and performing voice interaction. According to the invention, the accuracy of voice recognition is improved, so that the intelligent customer service can more accurately understand and respond to the voice instruction of the user, and the accuracy of voice interaction is improved.
Owner:HUAZE ZHONGXI (BEIJING) TECH DEV CO LTD

Device control method and device based on voice interaction

The embodiment of the invention provides a voice interaction method and device. The method can be applied to voice receiving equipment, and comprises the following steps: receiving a voice instruction of a user; determining a target regional space and a control instruction according to a target position in the voice instruction or indication information used for indicating the regional space in the voice instruction; sending area space information used for indicating the target area space and a voice instruction to the control equipment; receiving an execution result sent by the control equipment, wherein the execution result is a result that the controlled equipment executes the voice instruction control instruction; and providing feedback information according to the execution result. According to the method, when the user performs voice interaction with the household equipment, the household equipment in a certain specific area space expected to be controlled by the user can be controlled through a simple voice instruction.
Owner:HUAWEI TECH CO LTD

Voice conversation interaction method and system for industrial equipment

The invention provides a voice dialogue interaction method and system for industrial equipment, and relates to the technical field of man-machine interaction, and the method comprises the following steps: obtaining a voice signal of a user, and carrying out the voice enhancement processing of the voice signal of the user through an incremental adaptive filtering algorithm, and obtaining an enhanced voice; according to the enhanced voice, using an information gain transfer learning method to identify an interaction intention of the user, and generating a to-be-interacted voice based on the interaction intention of the user; based on the enhanced voice, utilizing a time delay estimation method to identify an interaction position of the user, and performing position prediction on the position of the user; according to the position prediction result, an industrial equipment horn output strategy is constructed; and outputting the to-be-interacted voice based on the industrial equipment loudspeaker output strategy so as to realize voice dialogue interaction between the user and the industrial equipment. According to the invention, convenience, high efficiency and accuracy of interaction between the user and the equipment in an industrial scene are greatly improved, and intelligent development of industrial production is facilitated.
Owner:CHINA APPLIED TECH CO LTD

Semiautomated relay method and apparatus

A captioning method for presenting captions to an assisted user (AU) during communication with a hearing user (HU) where the assisted user uses a captioned device and the hearing user uses a hearing user's device to facilitate the communication, the captioned device including a display screen and a speaker for presenting captions and broadcasting the hearing user's voice signals, respectively, the method comprising the steps of during an ongoing call between the AU and the HU, using an automated speech recognition (ASR) engine to generate initial ASR captions associated with the HU's voice signal, assessing at least one caption quality factor associated with prior initial ASR captions generated during the ongoing call, delaying broadcast of HU voice signal to the AU and based on the at least one caption quality factor, adjusting a duration of the HU voice signal broadcast delay.
Owner:ULTRATEC INC

Quality evaluation method for vehicle voice interaction function, electronic equipment and medium

The invention discloses a vehicle voice interaction function quality evaluation method, electronic equipment and a computer readable storage medium. The method comprises the steps of obtaining a test case set; according to a test case in the test case set, generating a voice request corresponding to the test case; the quality of the voice interaction function of the vehicle is evaluated according to the test case and the response result of the vehicle for the voice request, and the response result comprises a system operation log, a voice recognition result, execution operation, voice feedback and response time. Therefore, according to the test case in the test case set, the voice request corresponding to the test case is generated, and the quality of the voice interaction function of the vehicle is evaluated according to the response result of the voice request, so that the voice interaction function of the vehicle can be updated, maintained and the like based on the quality evaluation result; robust operation of the vehicle voice interaction function is guaranteed to a certain extent, and then the voice interaction function and the use experience of the vehicle are guaranteed.
Owner:GUANGZHOU XIAOPENG MOTORS TECH CO LTD

Dynamically activated graphical user interface with voice keyboard for electronic devices

1. Name of the product of this design: Dynamically activated graphical user interface with voice keyboard for electronic devices. 2. Purpose of this design product: an electronic device. 3. The key design point of this design product lies in the graphical user interface. 4. The picture or photo that best illustrates the key points of the design: Design 3 interface change state diagram 5. 5. Designate Design 3 as the base design. 6. Purpose of the graphical user interface: The purpose of the interface is to select and start the dynamic change process of the voice keyboard according to user operations, collect the user's voice and convert it into text. 7. Human-computer interaction method of graphical user interface: In Design 1, the main view is the main interface displayed by the electronic device. When the user clicks the input box in the main view, the interface jumps to interface change state diagram 1, interface change state diagram 2, interface change state diagram 3 and interface change state diagram 4 in sequence. In Design 2, the main view is the main interface displayed by the electronic device. When the user long presses the circular microphone icon in the main view, the interface jumps to Interface Change State Diagram 1, Interface Change State Diagram 2, and Interface Change State Diagram 3 in sequence. In Design 3 and Design 4, the main view is the main interface displayed by the electronic device. When the user clicks the input box in the main view, the interface jumps to Interface Change State Diagram 1 and Interface Change State Diagram 2 in sequence; when the user long presses the circular microphone icon in Interface Change State Diagram 2, the interface jumps to Interface Change State Diagram 3, Interface Change State Diagram 4 and Interface Change State Diagram 5 in sequence.
Owner:HUAWEI DEVICE CO LTD

Vending machine intelligent shopping guide method and system based on voice interaction

The invention discloses a vending machine intelligent shopping guide method and system based on voice interaction. The objective of the invention is to realize accurate recognition of user demands through voice recognition and natural language processing technologies, and perform intelligent recommendation in combination with commodity big data. According to the method, through integration of a high-sensitivity microphone and voice recognition, a voice instruction of a user is converted into text information, and a purchase intention behind the user is analyzed. Meanwhile, detailed information of commodities sold by the vending machine is collected, and after standardization processing and feature extraction, the detailed information serves as training data of the deep learning model. Personalized commodity recommendation can be generated according to user requirements through a trained and optimized model, and the personalized commodity recommendation is fed back to a user through a high-resolution display screen or voice output. According to the invention, the shopping experience of the user is improved, and the purchase conversion rate is improved through intelligent recommendation.
Owner:SHANGHAI QUZHI NETWORK TECH CO LTD

VR + AI rehabilitation training system

The invention discloses a VR + AI rehabilitation training system, belongs to the technical field of rehabilitation training, and aims to solve the problem of insufficient compliance of active rehabilitation training caused by physiological pain and psychological fear of a postoperative patient. Comprising a VR device, a doctor operation terminal, a motion capture device and a cloud server, and the VR device comprises a virtual scene management module which is used for receiving background music selected by a patient and a virtual scene selected by the patient from preset virtual scenes; the training parameter setting module is used for receiving a target distance, a target activity duration and a motion amplitude which are set by a doctor through the doctor operation terminal; the motion data acquisition module is used for receiving the motion amplitude acquired by the motion capture equipment and the motion direction acquired by the VR equipment; the motion state control module is used for controlling the motion state of the patient in the virtual scene according to the motion direction and the motion amplitude; and the AI voice interaction module is used for receiving voice input of the patient, generating answering content and outputting the answering content through voice interaction.
Owner:THE FIRST AFFILIATED HOSPITAL OF BENGBU MEDICAL COLLEGE

Intelligent AI voice control system and method for laboratory

The invention relates to the field of laboratory control, in particular to an intelligent AI voice control system and method for a laboratory. The method comprises the following steps: collecting a user voice instruction through an array microphone, analyzing the user voice instruction, and determining a user spatial position; acquiring space coordinates of a plurality of devices in a laboratory environment, and dynamically determining a target device which the user intends to control according to the user space position and the space coordinates; acquiring a control instruction set of the target equipment, and generating a structured control instruction according to the control instruction set and the user voice instruction; and controlling the target equipment to operate according to the structured control instruction. Target equipment mistaken selection caused by distance attenuation is effectively avoided, target equipment selection logic is strongly associated with the user sight line direction, pickup beams are dynamically adjusted according to the user space position, noise in a non-target area is effectively restrained, an execution structured control instruction is sent to the target equipment, and high-risk operation is effectively intercepted.
Owner:GUANGZHOU KELAOSISHIYANSHI INSTR EQUIP CO LTD

Cooking robot control system based on voice processing

The invention discloses a cooking robot control system based on voice processing, and relates to the technical field of intelligent robotics.The cooking robot control system comprises a steady-state locking subsystem, a dynamic detection subsystem, a change analysis subsystem and a feedback subsystem, firstly, a cooking robot sends a rotating speed control instruction to a motor corresponding to a stirring paddle by recognizing a voice command of a user, and the rotating speed control instruction is sent to the motor corresponding to the stirring paddle; the method comprises the following steps: collecting pot wall temperature data in a pot cavity, monitoring operation data of a motor of a stirring paddle in the pot cavity to determine a steady-state control mode, identifying the time for performing centrifugal inrush flow on food in the pot cavity based on the steady-state control mode and the push resistance rate monitored in real time to form centrifugal inrush flow and sauce redistribution operation, and then acquiring the pot wall temperature data in real time to obtain a sauce redistribution result. The temperature difference change before and after the short-time pulse acceleration is executed is analyzed to obtain an execution change temperature difference, an adjustment mechanism is triggered based on the execution change temperature difference, and finally, after the adjustment mechanism is executed, the temperature difference change is continuously fed back until the cooking operation is completed.
Owner:SHENZHEN HONGBO ZHICHENG TECH CO LTD

One time voice passphrase to protect against man-in-the-middle attack

Embodiments described herein provide for automatically authenticating operation requests and end-users who submit operation requests during contact events. A server obtains an operation request for an operation originated at an end-user device. The server generates a voice-based one-time password (OTP) using contextual information associated with the requested operation. The server generates and transmits an OTP prompt having text representing the OTP for display at a user interface of the user device. The server receives a response including an audio signal that contains the recording of the user speaking the OTP text aloud. The server uses the audio signal to authenticate the user and the operation request based on the speaker's voice, the accuracy of the user speaking the OTP, and liveness or fraud detection features extracted from the audio signal or metadata from the user device.
Owner:PINDROP SECURITY INC

System and method for adaptive beamforming for boomless headphones with fixed position microphones

A set of headphones comprising a digital signal processor, a headphone power management unit (PMU) to provide power to the digital signal microprocessor (DSP), a voice pick-up sensor to detect vibrations at a user's head caused by the user talking, and the DSP to determine where the user's voice audio input is picked up via an array of microphones formed into the set of headphones and wherein, when the DSP does not detect the user's voice audio input at the array of microphones, the DSP executes computer-readable program code of a beamforming module to recalibrate an angle of a voice detection zone at which the array of microphones detect the user's voice. The DSP to further calibrate the beamforming angle of the voice detection zone once the user's voice audio input is detected to meet an amplitude threshold level or a signal to noise level threshold level.
Owner:DELL PROD LP

Interaction method and system of intelligent glasses and translation machine

The embodiment of the invention belongs to the field of intelligent interaction, and relates to an interaction method of intelligent glasses and a translator, which comprises the following steps: the intelligent glasses synchronously acquire voice information of a user through a plurality of microphone matrixes; the voice information is transmitted to a translator in real time through a wireless communication protocol, a timestamp is added to each frame of voice information, and the translator detects transmission delay according to the timestamps and dynamically adjusts the transmission rate of the voice information; the translation machine divides the voice information into different subunits for processing according to syllables, intonations and grammar by adopting a preset acoustic model segmentation processing algorithm; the translation machine translates the voice information by using a vocabulary selection algorithm based on a language environment; and the translation machine feeds back a translation result to the intelligent glasses through an incremental feedback algorithm. The invention further provides an interaction system of the intelligent glasses and the translation machine. The objective of the invention is to realize low-delay and high-accuracy translation result feedback while ensuring high-precision speech recognition and rapid translation.
Owner:深圳目渡科技有限公司

Automatic regulation and control method of box equipment based on virtual digital human, medium and equipment

The invention relates to an automatic regulation and control method for box equipment based on virtual digital humans, which comprises the following steps of: acquiring consumption data of a user and song requesting data in a box, acquiring behavior data of the user in the box through image acquisition equipment, acquiring sound data of the user through audio acquisition equipment, and sending the sound data to the user through the audio acquisition equipment; generating multi-modal data of the user according to the consumption data, the song requesting data, the behavior data and the sound data; inputting the multi-modal data of the user into the trained neural network model to obtain a scene mode corresponding to the current box, and determining an interaction strategy of the virtual digital human according to the scene mode corresponding to the current box; and carrying out automatic regulation and control on box equipment based on an interaction strategy of the virtual digital human. According to the scheme, the box scene mode can be determined according to the multi-modal data, then the interaction strategy is determined based on the box scene mode, and the box equipment is automatically regulated and controlled through the virtual digital human, so that the user operation can be effectively simplified, and the interaction experience of the user is improved.
Owner:FUJIAN KAIMI NETWORK TECH CO LTD

System, program and method

To provide a novel system for evaluating human vocalization.SOLUTION: A system comprises at least one computer device. The system comprises specification means that uses a machine-learned prediction model that takes video information containing images of a person's lip during vocalization, or information regarding the position of a specified point on an image of a person's lip during vocalization as input information, and takes evaluation information related to the vocalization as output information so as to identify evaluation information related to the user's vocalization based on video information containing an image of the user's lip at the user's vocalization, or information related to the position of predetermined point on an image of the user's lip at the user's vocalization.SELECTED DRAWING: Figure 5
Owner:OKUCHY INC