Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

360 results about "S Voice" patented technology

S Voice is an intelligent personal assistant and knowledge navigator which is only available as a built-in application for the Samsung Galaxy S III, S III Mini (including NFC Variant), S4, S4 Mini, S4 Active, S5, S5 Mini, S II Plus, Note II, Note 3, Note 4, Note 10.1, Note 8.0, Stellar, Mega, Grand, Avant, Core, Ace 3, Tab 3 7.0, Tab 3 8.0, Tab 3 10.1, Galaxy Camera, and other 2013 or later Samsung Android devices. The application uses a natural language user interface to answer questions, make recommendations, and perform actions by delegating requests to a set of Web services. It is based on the Vlingo personal assistant.

Teaching note generation method and device based on large language model and medium

The embodiment of the invention discloses a teaching note generation method and device based on a large language model and a medium, belongs to the technical field of artificial intelligence, and solves the problems that classroom notes generated in the prior art are lack of content structuring and knowledge relevance, poor in readability and difficult to be used for effective review. The method comprises the following steps: acquiring a teacher voice signal in a classroom teaching environment in real time through pickup equipment, and converting the teacher voice signal into an initial text stream through a preset voice recognition model; taking the initial text flow as query input, and performing semantic similarity retrieval on the initial text flow and a preset course knowledge base to obtain a plurality of related knowledge fragments; generating an enhanced prompt based on the initial text stream and the plurality of related knowledge fragments, and generating a structured classroom note based on the enhanced prompt and a preset large language model; and outputting the structured classroom note to a user interface, and optimizing a classroom teaching note generation process based on feedback information and modification information of the user.
Owner:天元大数据信用管理有限公司

Method for automatic voice tuning and sound system using the method

This disclosure provides a method of automatic voice tuning for a sound system and the sound system using the method. The method may comprise: obtaining, via a microphone, an input signal representative of a user's voice; obtaining gender information; obtain a detected pitch by performing a pitch detection on the input signal; and applying a tuning control to the input signal based on the gender information and the detected pitch.
Owner:HARMAN INT IND INC +1

system

We provide the system. [Solution] A means for converting the user's voice into text using a speech recognition engine, A means of analyzing text data using natural language processing technology to understand user intent, A server device that generates a response based on intent and provides the generated response as text data, A terminal device that converts text data into speech and plays it back to the user, A system that includes this.
Owner:SOFTBANK GROUP CORP

System

PendingJP2026029847AData processing applicationsMoodSpeaking style
An object of a system according to an embodiment is to detect an emotion of a user and provide appropriate advice on the basis of the detected emotion.SOLUTION: A system according to an embodiment includes a smart emotion detection speaker, an emotion detection AI, and a mentoring generation AI. The smart speaker detects emotions from the user's voice and speech. The emotion detection AI generates words based on emotions. The mentoring generation AI listens to the user's distress and provides advice.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

System

A system is provided.SOLUTION: A system comprising: means for inputting profile information of a target person; means for transmitting the profile information of the target person to a server; means for collecting relevant information on the Internet based on the transmitted profile information and generating a customized conversation script; means for transmitting the generated customized conversation script to a terminal; and means for recognizing a user's voice and responding based on the customized conversation script.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

Outbound call processing methods, devices, and computer program products

This application discloses a method, apparatus, and computer program product for processing outbound call services. Relating to the field of artificial intelligence, the method includes: executing an outbound call task to a user; processing the target service according to a preset service processing procedure; collecting the user's voice signal during the target service processing; extracting voice features and voice content from the voice signal; determining the current service processing node corresponding to the voice content in the preset service processing procedure; determining the service path map corresponding to the preset service processing procedure; determining a path deviation index value based on the current service processing node and the service path map; inputting the voice features into a target model to obtain the user's behavior recognition result; determining a risk index value for the user interrupting the target service based on the behavior recognition result and the path deviation index value; and adjusting the preset service processing procedure according to the risk index value. This application solves the problem of poor accuracy in recognizing user intent in outbound call services in related technologies.
Owner:INDUSTRIAL AND COMMERCIAL BANK OF CHINA

Multi-room audio distribution method and system based on voice recognition

The invention provides a multi-room audio distribution method and system based on voice recognition, and the method comprises the steps: receiving a voice instruction of a user, and converting the voice instruction into text information; semantic analysis and intention recognition are conducted on the text information, and specific operation instructions and operation parameters are extracted; the operation instructions and the operation parameters are distributed to at least one function module related to the multi-room audio system, and the function module comprises an input source selection module, an input source volume control module, a multi-path output volume control module, a multi-path input and output routing module and a sound mixing module; the at least one function module executes corresponding operation according to the received operation instruction and the operation parameter so as to realize distribution of multi-room audios; and generating voice feedback information corresponding to the operation, and returning the voice feedback information to the user.
Owner:ZHONGKE WANYING (BEIJING) TECH CO LTD

System

A system is provided.SOLUTION: A system comprising: camera means for acquiring a body shape of a user; microphone means for recognizing a voice instruction of the user; display means for displaying a selection menu of clothes based on the voice instruction of the user; communication means for transmitting information of the clothes selected by the user to a generation and AI; generation and AI means for generating a fitting image when the user wears the selected clothes; display means for displaying the generated fitting image on a mirror; and communication means for performing a purchase procedure based on the voice instruction of the user.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

System

A system is provided.SOLUTION: Means for collecting voice data of a user, means for transmitting the collected voice data to a server, means for analyzing the voice data of the user in the server and extracting features such as a pitch, a rhythm, a vibrato, and a long tone, means for generating ideal singing data generated on the basis of an analysis result, means for transmitting the generated singing data to a user terminal, means for comparing singing of the user with the generated singing data, means for generating specific feedback on the basis of a comparison result, means for displaying feedback content to the user, and means for collecting new voice data for practice of the user, the system includes means for transmitting to the server again, and means for generating new scores and improvements based on the re-analysis results.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

Information processing system, information processing method, and program

The distribution of special offers will be made more entertaining. [Solution] The system includes an image acquisition unit that captures images of a user receiving guidance via an avatar, an audio acquisition unit that acquires the user's voice, an operation information acquisition unit that acquires operation information of the operator operating the avatar, and a reward selection unit that selects rewards to distribute to the user based on the images, the voice, and the operation information.
Owner:TOPPAN HOLDINGS INC

Semantic perception real-time voice endpoint detection method and device

The semantic perception real-time voice endpoint detection method is applied to a voice interaction system, and comprises the following steps: receiving a voice stream of a user and a voice stream sent by the system from at least two audio channels; performing causal voice coding on the voice stream of the user and the voice stream sent by the system to obtain corresponding double-path voice representation; a causal two-way modeling architecture is constructed, and the causal two-way modeling architecture is realized by adopting an autoregressive prediction network and is used for receiving two-way voice representation and modeling a speech round conversion dynamic state between a user and a system through an embedded causal constraint mechanism and a cross-channel interaction mechanism; based on the output of the autoregressive prediction network, jointly predicting the category of a speech round conversion event and the remaining time from the current moment to the end of the speech round; and a smoothing processing mode is used to correct the predicted speech wheel conversion event type, and a final endpoint detection result is determined. According to the method, the endpoint detection stability in a complex dialogue scene can be improved.
Owner:INST OF ACOUSTICS CHINESE ACAD OF SCI

Robot voice configuration method and device and electronic equipment

The embodiment of the invention provides a robot voice configuration method and device and electronic equipment, and relates to the technical field of robots, the method comprises the following steps: in response to a configuration operation for a configuration item in a configuration area of a voice interaction function, displaying a configuration parameter set by the configuration operation for the configuration item in the configuration area, the configuration area being located in a configuration page, the configuration page further comprises an interaction preview area; receiving an interaction request for simulating voice interaction with the robot, displaying the interaction request in the interaction preview area, responding to the interaction request based on the configuration parameters of the configuration items in the configuration area, and displaying a response result in the interaction preview area; and in response to the configuration application operation, issuing configuration parameters of each configuration item in the configuration area to the target robot. By applying the scheme provided by the embodiment of the invention, a user can flexibly configure the voice interaction function of the robot according to actual requirements.
Owner:BEIJING GALBOT AI CO LTD

Electronic apparatus for performing an operation corresponding to a user's voice and control method thereof

An electronic apparatus is provided. The electronic apparatus includes a memory configured to store a plurality of nodes corresponding to a plurality of user interface (UI) types, for each application, a display, and a processor connected with the memory and the display and controls the electronic apparatus, wherein the processor is configured to identify a first UI graph corresponding to the target application, acquire information on a target node that will perform the user command among a plurality of first nodes included in the first UI graph based on the user command and the at least one parameter, identify the current node corresponding to a UI displayed through the display, identify an action sequence from the current node to the target node based on the information on the current node and the target node, and perform an action corresponding to the user voice based on the action sequence.
Owner:SAMSUNG ELECTRONICS CO LTD

A voice and visual interaction control method for safe driving

This invention discloses a safe driving voice and visual interaction control method. The method includes: synchronously collecting and preprocessing the driver's visual data and voice interaction data; extracting key visual features based on the preprocessed visual data and calculating visual state feature values; triggering standardized voice interaction based on the visual state feature values, calculating a voice activity score in conjunction with the preprocessed voice interaction data, and determining the driver's voice response delay level based on the voice activity score; matching based on a predefined 3D virtual guide action sequence according to the voice response delay level, triggering a linkage response after matching to form a non-intrusive driving reminder with light, sound, and shape linkage; and completing closed-loop feedback control based on the execution state of the 3D virtual guide action sequence and the non-intrusive driving reminder with light, sound, and shape linkage. The method provided by this invention can reduce the monitoring misjudgment rate and achieve non-intrusive safety reminders.
Owner:SHANGHAI CHANGXING SOFTWARE CO LTD

Test method and device of vehicle, processing method and device of test data

This invention provides a vehicle testing method and apparatus, a test data processing method and apparatus, and two computer-readable storage media. The testing method includes the following steps: in response to a test function being triggered, sending collected vehicle data to a test terminal for parsing; making a voice call to the test terminal and determining whether the vehicle's voice call function is normal based on the received prompt tone; and in response to the determination that the voice call function is normal, acquiring the parsed data provided by the test terminal and determining whether the vehicle's function under test is normal based on the parsed data. By implementing this testing method, this invention can automatically acquire test data for the function under test and determine the test results, thereby improving the testing efficiency of the function under test.
Owner:SHANGHAI QINGGAN INTELLIGENT TECH CO LTD

system

PendingJP2026105499APersonalizationData set
We provide the system. [Solution] A data processing means that receives a photograph and analyzes the shape and physical characteristics of an individual from the photograph, Based on the analyzed characteristics, a proposal method is used to compare them with past data sets and provide personalized clothing suggestions. An information display means that generates and presents purchase instructions for the proposed clothing items, A mechanical means that receives a user's voice command and automatically takes a photograph and makes a suggestion based on the command, A system that includes this.
Owner:SOFTBANK GROUP CORP

Voice gateway equipment

The utility model provides a voice gateway device comprising an upper housing, and the surface of the upper housing is provided with heat radiation fins. The bottom shell is fixedly connected with the upper shell, a boss is arranged on the inner side of the bottom shell, and the upper shell and the bottom shell are fixedly connected to form a containing space; the main board is positioned in the accommodating space and is fixed on the boss; the battery is electrically connected with the main board, and the battery is fixed on the main board; the light guide part is fixed on the inner side of the bottom shell and is electrically connected with the mainboard; the antenna part is fixed on the outer side of the bottom shell and is electrically connected with the mainboard; a clamping groove is formed in the first side of the bottom shell; the second side of the bottom shell is provided with a plurality of interface holes. The voice gateway equipment is good in heat dissipation and simple to install, does not depend on an external power supply for power supply, and can ensure the continuity of voice communication under the condition of no external power supply for power supply.
Owner:SHENZHEN FLYINGVOICE NETWORK COMMUNICATION TECHNOLOGY CO LTD

system

We provide a system that enables more natural and accurate speech translation. [Solution] The system includes means for receiving the user's voice and acquiring audio data, means for providing speech recognition technology that analyzes the audio data and converts it into text data, means for motion recognition technology that captures the user's mouth movements and generates motion data, means for transmitting the text data and motion data to a server, means for the server to translate the text data into different languages ​​and convert the text in the different languages ​​into audio data, and means for transmitting the audio data to the user's terminal and playing it back.
Owner:SOFTBANK GROUP CORP

system

We provide the system. [Solution] A means of recording video of the screen being operated by the user, A means of recording the user's voice through a microphone, A means of sending recorded video files and recorded audio files to a server, A means of converting audio files to text on a server, A means of generating an operation manual based on converted text and video files, A means of providing the generated operation manual to the user, A system that includes this.
Owner:SOFTBANK GROUP CORP

A voice intercom device

The utility model discloses a voice intercom device, voice intercom device includes: voice emission module, voice receiving module, voice call button 110, MCU control module, MCU control module controls voice emission module to send call signal and user's voice signal, display screen module is used to show digit, first plug module, one end of first plug module is connected to computer interface, and the other end of first plug module is used to power supply voice emission module, voice receiving module, MCU control module and display screen module. The utility model has the beneficial effect that: after user presses voice call button 110, user communicates with the real -time voice of server through voice intercom device. User can obtain seat number or room number information from display screen module when communicating, thereby directly informs service personnel of server, avoids server service personnel after receiving multiple call signals simultaneously and can not confirm the position of user, improves the communication efficiency.
Owner:BOJING TIMES (DONGGUAN) TECHNOLOGY CO LTD

Action recognition device and action recognition method

To provide an action recognition device that enables more appropriate work management in production sites. [Solution] The behavior recognition device according to the present invention is a behavior recognition device that recognizes the behavior of a worker in a workspace for producing an object, and comprises: an authentication unit that authenticates the worker in the workspace based on skeletal information of the worker stored in advance; and a voice recognition unit that authenticates the worker's voice, recognizes the start and end of work in a predetermined process related to the object to be produced based on the worker's voice, and obtains the start time and end time of the work.
Owner:TOYOTA JIDOSHA KK

Electronic device capable of recognizing user and control method thereof

An electronic device is disclosed. The electronic device includes an interface, a memory, and one or more processors, wherein the one or more processors can: when a gesture for account registration is recognized based on sensing data acquired through the interface, provide one or more candidate words based on feature information of the gesture; when one or more of the candidate words are selected, store account information including the selected candidate word and feature information of the gesture in the memory; and when a user's voice and gesture are recognized, perform control operations based on account information corresponding to the user's voice and gesture from among a plurality of account information stored in the memory.
Owner:SAMSUNG ELECTRONICS CO LTD

A method for projecting a mobile phone using voice control

PendingCN122340212ADisplay deviceSpeech sound
This invention provides a method for voice-controlled mobile phone screen mirroring, comprising: a source device receiving a user's voice command, the voice command triggering a screen mirroring operation; the source device determining the current usage scenario based on the voice command, the determination distinguishing between video playback scenarios and non-video playback scenarios based on foreground application information and audio output status; the source device selecting a screen mirroring mode based on the determined usage scenario and performing screen mirroring on a target display device, the screen mirroring mode including a video mirroring mode and a mirrored mirroring mode. This invention accurately distinguishes between video playback scenarios and non-video playback scenarios, thereby adaptively selecting either a video mirroring mode or a mirrored mirroring mode to ensure efficient presentation of the mirrored content, and matches with voice commands from the vehicle's infotainment system, achieving cross-device interactive control.
Owner:RIVOTEK TECH (JIANGSU) CO LTD

Interaction information and operation button display area of a voice interaction graphical user interface of an electronic device

1. Name of the product in this design: Interactive information and operation button display area of ​​a voice-interactive graphical user interface for electronic devices. 2. Purpose of this design: An electronic device. 3. The key design features of this product are the graphical user interface in the electronic device for which protection is sought. 4. The image or photograph that best illustrates the key design points: Design 1 change state diagram. 5. Design 1 is designated as the basic design. 6. Purpose of Graphical User Interface: The graphical user interface is used for voice interaction, such as voice interaction with an intelligent agent. The part of the graphical user interface that needs protection is the interactive information and operation button display area of ​​the voice interaction graphical user interface. 7. Human-computer interaction method of graphical user interface: In Design 1, the user clicks the voice input button in the main view of Design 1, and the graphical user interface changes from the main view of Design 1 to the change state diagram of Design 1 to display the user's voice input status and the information of the recognized user voice input. In Design 2, when the user clicks the voice input button in the main view of Design 2, the graphical user interface changes from the main view of Design 2 to Design 2 Change State Figure 1 to display the user's voice input status and the recognized user voice input information. When the user slides up in Design 2 Change State Figure 1, the graphical user interface changes from Design 2 Change State Figure 1 to Design 2 Change State Figure 2 to display a prompt message to cancel sending voice. 8. Other situations requiring explanation: The portion of the graphical user interface shown by the dashed lines in the views of Design 1 to Design 2 does not constitute the part that is protected by this design.
Owner:BEIJING BAIDU NETCOM SCI & TECH CO LTD

Area displaying a status of an electronic device information display dynamic graphical user interface virtual assistant

1. The name of the design product: the area of the display virtual assistant of the dynamic graphical user interface of the electronic device information display. 2. The use of the design product: an electronic device. 3. The design points of the design product: the protected part of the graphical user interface. 4. The picture or photo that best indicates the design points: change state figure 3. 5. The use of the graphical user interface: the graphical user interface can be used for human-computer interaction and implementation of the functions of the electronic device, and can be used for information display, the protected part of the graphical user interface is the area of the display virtual assistant of the information display graphical user interface. 6. The human-computer interaction mode of the graphical user interface: the graphical user interface can be interacted by clicking, pressing, sliding or operating the graphical user interface, or by capturing the user's hand, eyes or voice. 7. The change state description of the graphical user interface: when the virtual assistant is activated by operating the electronic device and / or by capturing the user's voice, the appearance of the graphical user interface changes from the main view to change state figure 1, change state figure 2, change state figure 3 in turn. 8. Other circumstances that need to be explained: the dashed line does not constitute part of the appearance design claimed.
Owner:APPLE INC

Voice interaction methods, devices, equipment and storage media

PendingCN122313966AEngineeringPhysical therapy
This disclosure relates to a voice interaction method, apparatus, device, and storage medium, comprising: receiving a voice interaction command, executing a target task indicated by the voice interaction command, and exiting the voice interaction after the target task is completed. Compared with the prior art, the solution disclosed in this application executes the target task corresponding to the voice interaction command after the user issues a voice interaction command, and automatically exits the voice interaction after the target task is completed. This avoids the voice interaction function remaining in an awake state for a long time, continuously listening to and responding to the user's voice input, which could lead to erroneous responses to the user's voice, reducing unnecessary disturbance to the user, and thus improving user comfort and satisfaction.
Owner:SHANGHAI LIXIANG AUTOMOBILE CO LTD

system

We provide the system. [Solution] Means for detecting incoming calls, A means to switch to answering machine mode if there is no response within a certain period of time, A means of recording the caller's voice message, A means of converting recorded audio data into text, A means of notifying the user's device of the converted text data, A system that includes this.
Owner:SOFTBANK GROUP CORP

System

A system is provided.SOLUTION: A system comprising: means for obtaining a user's speech input and converting the speech data to text data; means for parsing the text data and generating an appropriate response using a natural language processing engine; and means for converting the generated response to speech and providing it to the user.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

System

A system is provided.SOLUTION: A system, comprising: means for converting a user's speech into a digital signal; means for transmitting the digital signal to a server; means for converting the speech digital signal into character string text on the server; means for translating the converted text into a specified language in real time; means for converting the translated text into speech data that retains characteristics of the original speech; means for transmitting the speech data to a terminal; and means for playing the speech data on the terminal.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

A method and system for multi-mode switching control of a "best friend" phone

This invention discloses a multi-mode switching control method and system for a smart terminal device, relating to the field of intelligent terminal control technology. The method includes: acquiring activity information related to the currently running task and pre-recorded main user voice data; determining the device task status based on the activity information; if the device task status is a high-priority core task, monitoring environmental changes; if the environmental change is a change in light, adjusting the image sensor and acquiring a video stream through the adjusted image sensor to maintain visual continuity during multi-mode switching; if the environmental change is a change in sound source, recognizing the main user's voice signal based on the pre-recorded main user voice data to maintain clear main user voice during multi-mode switching. This invention can combine activity information and pre-recorded main user voice data to adjust video or audio under different environmental changes, thereby achieving multi-mode switching control, improving accuracy and user experience.
Owner:FOSHAN CHENGYI TECHNOLOGY CO LTD