Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

450 results about "S Voice" patented technology

S Voice is an intelligent personal assistant and knowledge navigator which is only available as a built-in application for the Samsung Galaxy S III, S III Mini (including NFC Variant), S4, S4 Mini, S4 Active, S5, S5 Mini, S II Plus, Note II, Note 3, Note 4, Note 10.1, Note 8.0, Stellar, Mega, Grand, Avant, Core, Ace 3, Tab 3 7.0, Tab 3 8.0, Tab 3 10.1, Galaxy Camera, and other 2013 or later Samsung Android devices. The application uses a natural language user interface to answer questions, make recommendations, and perform actions by delegating requests to a set of Web services. It is based on the Vlingo personal assistant.

Semiautomated relay method and apparatus

ActiveUS12482458B2Special service for subscribersTelephone sets with user guidance/featuresElectrophonic hearingAutomatic speech
A captioning method for presenting captions to an assisted user (AU) during communication with a hearing user (HU) where the assisted user uses a captioned device and the hearing user uses a hearing user's device to facilitate the communication, the captioned device including a display screen and a speaker for presenting captions and broadcasting the hearing user's voice signals, respectively, the method comprising the steps of during an ongoing call between the AU and the HU, using an automated speech recognition (ASR) engine to generate initial ASR captions associated with the HU's voice signal, assessing at least one caption quality factor associated with prior initial ASR captions generated during the ongoing call, delaying broadcast of HU voice signal to the AU and based on the at least one caption quality factor, adjusting a duration of the HU voice signal broadcast delay.
Owner:ULTRATEC INC

One time voice passphrase to protect against man-in-the-middle attack

Embodiments described herein provide for automatically authenticating operation requests and end-users who submit operation requests during contact events. A server obtains an operation request for an operation originated at an end-user device. The server generates a voice-based one-time password (OTP) using contextual information associated with the requested operation. The server generates and transmits an OTP prompt having text representing the OTP for display at a user interface of the user device. The server receives a response including an audio signal that contains the recording of the user speaking the OTP text aloud. The server uses the audio signal to authenticate the user and the operation request based on the speaker's voice, the accuracy of the user speaking the OTP, and liveness or fraud detection features extracted from the audio signal or metadata from the user device.
Owner:PINDROP SECURITY INC

Dual-layered artificial intelligence system with large language models and different virtual agents

Embodiments of the present disclosure may include a dual-layered artificial intelligence system including a leading virtual agent and a set of other virtual agents: the leading agent, equipped with vast general knowledge, interfaces with the user and enforces guidelines in the overarching goal and progress, branding, guiderails, regulatory compliance, and the system's voice, while the other agents contain vast knowledge in a specific domain. These other agents only communicate with the leading agent and are called upon by the leading agent when their respective expertise is needed to solve the goal.
Owner:BITHUMAN INC

Voice control method and device, electronic equipment and storage medium

The invention provides a voice control method and device, electronic equipment and a storage medium, and relates to the technical field of smart home. The method comprises the steps of obtaining a to-be-recognized voice signal of a user; inputting a to-be-recognized voice signal into the instruction simplification model to obtain a target control instruction output by the instruction simplification model; the instruction simplification model comprises a voice recognition model and a language simplification model, and the voice recognition model is used for converting a to-be-recognized voice signal into a text instruction; the language simplification model is used for simplifying the text instruction according to a preset semantic rule and a vocabulary mode to obtain a target control instruction; and executing the target control instruction. According to the method, the target control instruction is generated through voice recognition and language simplification, when a user sends out a tedious to-be-recognized voice signal with low accuracy, the language simplification model can filter redundant vocabularies and reduce semantic ambiguity, and then the problem that execution of equipment is affected due to the fact that the accuracy of the to-be-recognized voice signal sent by the old is low is solved.
Owner:GREE ELECTRIC APPLIANCE INC OF ZHUHAI +1

Teaching note generation method and device based on large language model and medium

The embodiment of the invention discloses a teaching note generation method and device based on a large language model and a medium, belongs to the technical field of artificial intelligence, and solves the problems that classroom notes generated in the prior art are lack of content structuring and knowledge relevance, poor in readability and difficult to be used for effective review. The method comprises the following steps: acquiring a teacher voice signal in a classroom teaching environment in real time through pickup equipment, and converting the teacher voice signal into an initial text stream through a preset voice recognition model; taking the initial text flow as query input, and performing semantic similarity retrieval on the initial text flow and a preset course knowledge base to obtain a plurality of related knowledge fragments; generating an enhanced prompt based on the initial text stream and the plurality of related knowledge fragments, and generating a structured classroom note based on the enhanced prompt and a preset large language model; and outputting the structured classroom note to a user interface, and optimizing a classroom teaching note generation process based on feedback information and modification information of the user.
Owner:天元大数据信用管理有限公司

Method for automatic voice tuning and sound system using the method

This disclosure provides a method of automatic voice tuning for a sound system and the sound system using the method. The method may comprise: obtaining, via a microphone, an input signal representative of a user's voice; obtaining gender information; obtain a detected pitch by performing a pitch detection on the input signal; and applying a tuning control to the input signal based on the gender information and the detected pitch.
Owner:HARMAN INT IND INC +1

system

We provide the system. [Solution] A means for converting the user's voice into text using a speech recognition engine, A means of analyzing text data using natural language processing technology to understand user intent, A server device that generates a response based on intent and provides the generated response as text data, A terminal device that converts text data into speech and plays it back to the user, A system that includes this.
Owner:SOFTBANK GROUP CORP

System

PendingJP2026029847AData processing applicationsMoodSpeaking style
An object of a system according to an embodiment is to detect an emotion of a user and provide appropriate advice on the basis of the detected emotion.SOLUTION: A system according to an embodiment includes a smart emotion detection speaker, an emotion detection AI, and a mentoring generation AI. The smart speaker detects emotions from the user's voice and speech. The emotion detection AI generates words based on emotions. The mentoring generation AI listens to the user's distress and provides advice.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

System

A system is provided.SOLUTION: A system comprising: means for inputting profile information of a target person; means for transmitting the profile information of the target person to a server; means for collecting relevant information on the Internet based on the transmitted profile information and generating a customized conversation script; means for transmitting the generated customized conversation script to a terminal; and means for recognizing a user's voice and responding based on the customized conversation script.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

Method, device, mobile user device, computer program for controlling an audio system of a vehicle

A method for controlling an audio system of a vehicle, comprising: Determining, in particular, predicting a parameter of an auditory communication in the interior of a vehicle, and - Changing at least one parameter of an audio signal of the vehicle's audio system depending on the determined, in particular predicted, parameter of the auditory communication, where - the audio signal is a multimedia content played via a multimedia system of the vehicle, a part of a sound design of the vehicle and / or an artificially generated engine noise, and the change in the parameter of the audio signal includes - a modification of the audio signal in the frequency domain such that signal components of the audio signal relating to the frequency range of the speaker's voice are modified such that the signal energy is distributed over one or more frequency ranges which do not relate to the frequency range of the speaker's voice or relate to it to a lesser extent.
Owner:BAYERISCHE MOTOREN WERKE AG

Outbound call processing methods, devices, and computer program products

This application discloses a method, apparatus, and computer program product for processing outbound call services. Relating to the field of artificial intelligence, the method includes: executing an outbound call task to a user; processing the target service according to a preset service processing procedure; collecting the user's voice signal during the target service processing; extracting voice features and voice content from the voice signal; determining the current service processing node corresponding to the voice content in the preset service processing procedure; determining the service path map corresponding to the preset service processing procedure; determining a path deviation index value based on the current service processing node and the service path map; inputting the voice features into a target model to obtain the user's behavior recognition result; determining a risk index value for the user interrupting the target service based on the behavior recognition result and the path deviation index value; and adjusting the preset service processing procedure according to the risk index value. This application solves the problem of poor accuracy in recognizing user intent in outbound call services in related technologies.
Owner:INDUSTRIAL AND COMMERCIAL BANK OF CHINA

Multi-room audio distribution method and system based on voice recognition

The invention provides a multi-room audio distribution method and system based on voice recognition, and the method comprises the steps: receiving a voice instruction of a user, and converting the voice instruction into text information; semantic analysis and intention recognition are conducted on the text information, and specific operation instructions and operation parameters are extracted; the operation instructions and the operation parameters are distributed to at least one function module related to the multi-room audio system, and the function module comprises an input source selection module, an input source volume control module, a multi-path output volume control module, a multi-path input and output routing module and a sound mixing module; the at least one function module executes corresponding operation according to the received operation instruction and the operation parameter so as to realize distribution of multi-room audios; and generating voice feedback information corresponding to the operation, and returning the voice feedback information to the user.
Owner:ZHONGKE WANYING (BEIJING) TECH CO LTD

System

A system is provided.SOLUTION: A system comprising: camera means for acquiring a body shape of a user; microphone means for recognizing a voice instruction of the user; display means for displaying a selection menu of clothes based on the voice instruction of the user; communication means for transmitting information of the clothes selected by the user to a generation and AI; generation and AI means for generating a fitting image when the user wears the selected clothes; display means for displaying the generated fitting image on a mirror; and communication means for performing a purchase procedure based on the voice instruction of the user.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

System

A system is provided.SOLUTION: Means for collecting voice data of a user, means for transmitting the collected voice data to a server, means for analyzing the voice data of the user in the server and extracting features such as a pitch, a rhythm, a vibrato, and a long tone, means for generating ideal singing data generated on the basis of an analysis result, means for transmitting the generated singing data to a user terminal, means for comparing singing of the user with the generated singing data, means for generating specific feedback on the basis of a comparison result, means for displaying feedback content to the user, and means for collecting new voice data for practice of the user, the system includes means for transmitting to the server again, and means for generating new scores and improvements based on the re-analysis results.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

The interactive hypnotherapeutic computer-implemented system based on artificial intelligence and a related computer program product

The invention discloses an interactive hypnotherapeutic computer-implemented system based on artificial intelligence, comprising a microphone for receiving voice commands from a user; a display device for displaying dynamic pictorial representations generated by the system; a computing device operatively connected to the display device and the microphone. The computing device is configured to execute a computer program implementing a hypnotherapeutic interaction with the user based on received voice commands, whereby the computing device comprises: a word recognition module configured to recognize predefined induction words or phrases from the received voice commands; an artificial intelligence, AI, response generation module trained to respond to the recognized induction words or phrases by generating corresponding dynamic pictorial representations, created by moving particles, on the display device, and configured to record interactions and optimize responses based on user-specific feedback. The system transitions through three stages represented by distinct visual patterns of dynamic pictorial representations, induced by the predefined induction commands, wherein the dynamic pictorial representations vary in parameters defining the behavior of pictorial representation's particles, in particular vary in color and / or speed and / or movement pattern and / or scattering of representations' particles; and wherein the visual representations generated by the system are semiotically inconsistent and do not match the semantic signification of the user's voice commands. The invention also discloses a computer program product comprising instructions that cause the computer system to: a. recognize specific induction words or phrases from voice commands received via a microphone; b. in response to the recognized induction words or phrases, generate corresponding dynamic pictorial representations, comprising particles, on a display device using an artificial intelligence response generation module, the pictorial representations being distinct for three stages of the system and varying in color, speed, movement pattern and scattering of representations' particles; c. record interactions and optimize responses based on user-specific feedback; wherein the visual representations are semiotically inconsistent and do not match the semantic signification of the user's voice commands.
Owner:UNIV SLASKI W KATOWICACH

Information processing system, information processing method, and program

The distribution of special offers will be made more entertaining. [Solution] The system includes an image acquisition unit that captures images of a user receiving guidance via an avatar, an audio acquisition unit that acquires the user's voice, an operation information acquisition unit that acquires operation information of the operator operating the avatar, and a reward selection unit that selects rewards to distribute to the user based on the images, the voice, and the operation information.
Owner:TOPPAN HOLDINGS INC

Semantic perception real-time voice endpoint detection method and device

The semantic perception real-time voice endpoint detection method is applied to a voice interaction system, and comprises the following steps: receiving a voice stream of a user and a voice stream sent by the system from at least two audio channels; performing causal voice coding on the voice stream of the user and the voice stream sent by the system to obtain corresponding double-path voice representation; a causal two-way modeling architecture is constructed, and the causal two-way modeling architecture is realized by adopting an autoregressive prediction network and is used for receiving two-way voice representation and modeling a speech round conversion dynamic state between a user and a system through an embedded causal constraint mechanism and a cross-channel interaction mechanism; based on the output of the autoregressive prediction network, jointly predicting the category of a speech round conversion event and the remaining time from the current moment to the end of the speech round; and a smoothing processing mode is used to correct the predicted speech wheel conversion event type, and a final endpoint detection result is determined. According to the method, the endpoint detection stability in a complex dialogue scene can be improved.
Owner:INST OF ACOUSTICS CHINESE ACAD OF SCI

Robot voice configuration method and device and electronic equipment

The embodiment of the invention provides a robot voice configuration method and device and electronic equipment, and relates to the technical field of robots, the method comprises the following steps: in response to a configuration operation for a configuration item in a configuration area of a voice interaction function, displaying a configuration parameter set by the configuration operation for the configuration item in the configuration area, the configuration area being located in a configuration page, the configuration page further comprises an interaction preview area; receiving an interaction request for simulating voice interaction with the robot, displaying the interaction request in the interaction preview area, responding to the interaction request based on the configuration parameters of the configuration items in the configuration area, and displaying a response result in the interaction preview area; and in response to the configuration application operation, issuing configuration parameters of each configuration item in the configuration area to the target robot. By applying the scheme provided by the embodiment of the invention, a user can flexibly configure the voice interaction function of the robot according to actual requirements.
Owner:BEIJING GALBOT AI CO LTD

Vehicle and control method thereof

Disclosed are a vehicle and a control method thereof that may allow an occupant as well as a driver to conveniently use a speech recognition system by utilizing a microphone of a mobile device as a source of speech input, that may utilize a microphone of a mobile device as a source of noise collection, when a driver uses a speech recognition system. A vehicle may include a microphone; a speaker; a communication module configured to communicate with a mobile device; and a controller. In a first mode, the communication module is configured to receive a voice signal of a user from the mobile device, and in the first mode, the controller is configured to perform noise cancellation on the user's voice signal received from the mobile device, based on an audio signal input to the microphone.
Owner:HYUNDAI MOTOR CO LTD +1

Voice conversation method and device based on deep learning, medium and electronic equipment

The invention relates to the technical field of artificial intelligence, and discloses a voice dialogue method and device based on deep learning, a medium and electronic equipment, and the method comprises the steps: extracting original text information corresponding to a user voice signal, and generating target text information in combination with a long short-term memory network and an attention mechanism; performing deep semantic analysis on the target text information, calculating a long-distance dependency weight in combination with a Transform structure used for voice semantic analysis, and generating a user intention vector; inputting the user intention vector into a dialogue generation model to obtain original dialogue text content; and optimizing the original dialogue text content by utilizing a generative adversarial network, generating optimized dialogue content, and converting the optimized dialogue content into voice dialogue content to be output. The speech recognition accuracy can be improved, the ability of natural language processing to understand complex semantics and contexts is improved, dialogues are more personalized and emotion, and the requirements of people for high-quality and humanized speech dialogues are met.
Owner:SICHUAN YIYUN INTELLIGENT NETWORKED AUTOMOBILE TECHNOLOGY CO LTD

Electronic apparatus for performing an operation corresponding to a user's voice and control method thereof

An electronic apparatus is provided. The electronic apparatus includes a memory configured to store a plurality of nodes corresponding to a plurality of user interface (UI) types, for each application, a display, and a processor connected with the memory and the display and controls the electronic apparatus, wherein the processor is configured to identify a first UI graph corresponding to the target application, acquire information on a target node that will perform the user command among a plurality of first nodes included in the first UI graph based on the user command and the at least one parameter, identify the current node corresponding to a UI displayed through the display, identify an action sequence from the current node to the target node based on the information on the current node and the target node, and perform an action corresponding to the user voice based on the action sequence.
Owner:SAMSUNG ELECTRONICS CO LTD

A voice and visual interaction control method for safe driving

This invention discloses a safe driving voice and visual interaction control method. The method includes: synchronously collecting and preprocessing the driver's visual data and voice interaction data; extracting key visual features based on the preprocessed visual data and calculating visual state feature values; triggering standardized voice interaction based on the visual state feature values, calculating a voice activity score in conjunction with the preprocessed voice interaction data, and determining the driver's voice response delay level based on the voice activity score; matching based on a predefined 3D virtual guide action sequence according to the voice response delay level, triggering a linkage response after matching to form a non-intrusive driving reminder with light, sound, and shape linkage; and completing closed-loop feedback control based on the execution state of the 3D virtual guide action sequence and the non-intrusive driving reminder with light, sound, and shape linkage. The method provided by this invention can reduce the monitoring misjudgment rate and achieve non-intrusive safety reminders.
Owner:SHANGHAI CHANGXING SOFTWARE CO LTD

Test method and device of vehicle, processing method and device of test data

This invention provides a vehicle testing method and apparatus, a test data processing method and apparatus, and two computer-readable storage media. The testing method includes the following steps: in response to a test function being triggered, sending collected vehicle data to a test terminal for parsing; making a voice call to the test terminal and determining whether the vehicle's voice call function is normal based on the received prompt tone; and in response to the determination that the voice call function is normal, acquiring the parsed data provided by the test terminal and determining whether the vehicle's function under test is normal based on the parsed data. By implementing this testing method, this invention can automatically acquire test data for the function under test and determine the test results, thereby improving the testing efficiency of the function under test.
Owner:SHANGHAI QINGGAN INTELLIGENT TECH CO LTD

Vehicle interaction methods, devices, vehicles, and storage media

This application relates to a vehicle interaction method, device, vehicle, and storage medium. The method includes: receiving a user's voice control command in a visible-to-speak mode; determining the user's actual intent based on the voice control command, and if the actual intent is a control intent for the central control screen, converting the current display content of the vehicle's central control screen into corresponding text content using a preset large model; after voice-reading the text content, controlling the central control screen to perform corresponding control actions based on the user's response to the text content, and after the control actions are completed, voice-reading the updated display content of the central control screen. This solves the problems of not being able to recognize image information, difficulty in accurately representing screen content, and the high cost of requiring third-party application developers to control related interactions.
Owner:CHERY AUTOMOBILE CO LTD

system

PendingJP2026105499APersonalizationData set
We provide the system. [Solution] A data processing means that receives a photograph and analyzes the shape and physical characteristics of an individual from the photograph, Based on the analyzed characteristics, a proposal method is used to compare them with past data sets and provide personalized clothing suggestions. An information display means that generates and presents purchase instructions for the proposed clothing items, A mechanical means that receives a user's voice command and automatically takes a photograph and makes a suggestion based on the command, A system that includes this.
Owner:SOFTBANK GROUP CORP

Registration recommendation methods, model training methods, devices, electronic equipment and media

This application discloses a registration recommendation method, model training method, device, electronic device, and medium, relating to the field of medical service equipment technology. In the registration recommendation model processing, the target primary department can be used as prior knowledge to improve the accuracy of the model. An attention mechanism layer helps the model understand the contextual information of the second feature and selects the feature that contributes most to the registration recommendation, thus improving accuracy. Then, through a fully connected layer and a classifier, the target secondary department is obtained, and both the target primary and secondary departments are displayed to the user. The target secondary department is the department recommended for the user's registration. Thus, this application can automatically generate recommended secondary departments for registration based on the user's voice-based medical information, facilitating registration with high accuracy.
Owner:ZHUHAI QUANSHITONG INFORMATION TECH CO LTD

robot

ActiveJP1811320SNursing careEngineering
This item is an autonomous robot capable of two-way communication with the user. The item propels itself using wheels attached to the bottom of its body. The body is equipped with communication functions, cameras, and other sensors, allowing it to autonomously converse and change its facial expressions and arm movements in response to the user's voice and facial expressions. The user can also remotely control the item using a controller. The item aims to enrich people's lives by autonomously communicating its will, emotions, and thoughts with the user, just like a real person, building a relationship of trust. The purpose and use of the item is up to the user. For example, it can be used as a nursing companion for the sick, or as a pet or toy for children to play with, and is not suited to being limited to a specific industrial use.
Owner:GROOVE X INC

Voice gateway equipment

The utility model provides a voice gateway device comprising an upper housing, and the surface of the upper housing is provided with heat radiation fins. The bottom shell is fixedly connected with the upper shell, a boss is arranged on the inner side of the bottom shell, and the upper shell and the bottom shell are fixedly connected to form a containing space; the main board is positioned in the accommodating space and is fixed on the boss; the battery is electrically connected with the main board, and the battery is fixed on the main board; the light guide part is fixed on the inner side of the bottom shell and is electrically connected with the mainboard; the antenna part is fixed on the outer side of the bottom shell and is electrically connected with the mainboard; a clamping groove is formed in the first side of the bottom shell; the second side of the bottom shell is provided with a plurality of interface holes. The voice gateway equipment is good in heat dissipation and simple to install, does not depend on an external power supply for power supply, and can ensure the continuity of voice communication under the condition of no external power supply for power supply.
Owner:SHENZHEN FLYINGVOICE NETWORK COMMUNICATION TECHNOLOGY CO LTD

system

We provide a system that enables more natural and accurate speech translation. [Solution] The system includes means for receiving the user's voice and acquiring audio data, means for providing speech recognition technology that analyzes the audio data and converts it into text data, means for motion recognition technology that captures the user's mouth movements and generates motion data, means for transmitting the text data and motion data to a server, means for the server to translate the text data into different languages ​​and convert the text in the different languages ​​into audio data, and means for transmitting the audio data to the user's terminal and playing it back.
Owner:SOFTBANK GROUP CORP

system

We provide the system. [Solution] A means of recording video of the screen being operated by the user, A means of recording the user's voice through a microphone, A means of sending recorded video files and recorded audio files to a server, A means of converting audio files to text on a server, A means of generating an operation manual based on converted text and video files, A means of providing the generated operation manual to the user, A system that includes this.
Owner:SOFTBANK GROUP CORP