Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

8 results about "Received spoken" patented technology

A bidirectional sign language communication system with a multi-sensor glove and haptic feedback

A bidirectional sign language communication system (100) with multisensor glove and haptic feedback, comprising: a wearable glove with a variety of sensors configured to detect hand, finger and wrist movements of a user, wherein the sensors include at least curvature sensors, a pinch detector, a strain sensor, a pressure measuring arrangement and an inertial measurement unit; a processing unit that is functionally coupled with the sensors and configured to process sensor signals and derive sign language tokens using a machine learning model; a wireless communication interface configured to transmit the derived tokens to a companion device for playback as text or speech; wherein the companion device is further configured to receive spoken or textual inputs, convert the inputs into coded tactile patterns, and transmit the patterns to the wearable glove; and a distributed haptic feedback subsystem integrated into the wearable glove and configured to present tactile patterns to the user via fingertip actuators, vibration motors, or a tactile matrix, wherein the system (100) enables bidirectional real-time offline communication between a sign language user and a non-sign language user.
Owner:VIJ RASHVIN HIGHLANDS RANCH

Method and apparatus for generating regression-based gloss-free sign language pose using discrete representation

The present invention relates to a method and apparatus for generating a regression-based gloss-free sign language pose using discrete representation that generates a sign language pose on the basis of (gloss-free) spoken sentence data, and the method comprises the steps of: receiving spoken sentence data as an input and transforming the spoken sentence data into a discretized sign language pose token according to spatial and temporal characteristics by using a self-supervised learning model; and automatically recursively generating a sign language pose through a transformer-based encoder-decoder architecture by using the discretized sign language pose token as an input.
Owner:KOREA ADVANCED INST OF SCI & TECH

Systems and methods for automated speech-to-transaction in healthcare operations

A computer-implemented system and method enable real-time healthcare transaction processing based on ambient audio captured during physician-patient conversations. The system includes one or more microphones to receive spoken dialogue and a speech recognition engine to transcribe the audio into text. A natural language processing module analyzes the text to identify clinical intents corresponding to healthcare transactions. The identified intent is standardized into a structured query format and transmitted to an external system via a communications interface. In response, the system receives external data such as insurance coverage status, cost estimates, or authorization requirements. A user interface displays the external response in real time during the clinical encounter, allowing the physician or patient to make informed decisions. Additional modules may include context monitoring logic to suppress non-actionable utterances and a secure storage engine to log queries and responses. The system automates insurance verification and authorization workflows without requiring manual data entry.
Owner:BLACKSTONE MICHAEL

Eyeglass Translation System

An eyeglass translation system includes a wearable frame, an audiovisual input assembly, a microprocessor, and an audiovisual output assembly. The wearable frame includes rims, a pair of temples, and a pair of lenses in the rims. The audiovisual input assembly is positioned in the frame and is designed to receive spoken and written words. The microprocessor is positioned in the frame and is operatively connected to the audiovisual input assembly to receive input words from the audiovisual input assembly and translate input words into a selected language. The audiovisual output assembly is positioned in the frame and operatively connected to the microprocessor to produce visual text or audio of translated input words received from the microprocessor. The eyeglass translation system can detect spoken or written words, translate the words into a selected language, and then provide audio or visual images of the translated words.
Owner:CASTRO BIENVENIDO DANIEL

Natural language interactions using visual understanding

Techniques for performing an action with respect to displayed content are described. A natural language interpretation corresponding to a received spoken user input may be determined. Prior to receiving the spoken user input, content may be displayed to the user from which the spoken user input was received. The natural language interpretation may represent a request to perform an action with respect to a portion of the content currently being displayed. Content identifiers corresponding to content being displayed, may be determined, and embedding data representing at least one feature of the content may be determined using the content identifiers. The natural language interpretation and the embedding data may be processed to determine that the spoken user input relates to a first portion of the displayed content instead of a second portion of the displayed content. Based on the determination, an action responsive to the spoken user input may be performed.
Owner:AMAZON TECH INC

Input system and method

An input system comprises a spoken input processor configured to receive spoken inputs from a microphone operably coupled to the input system; a physical input processor configured to receive physical inputs from a peripheral device operably coupled to the input system, the physical inputs having a timing relative to the spoken inputs; and a speech recognition processor configured to recognise speech from the speech inputs; wherein the speech recognition processor uses at least some of the received physical inputs, and their timing relative to the spoken inputs, as part of the speech recognition process.
Owner:SONY INTERACTIVE ENTERTAINMENT LLC

Voice command and audio output to obtain built-in test equipment data

A voice processing system for obtaining data from built-in test equipment of an aircraft which includes an audio device adapted and configured to receive spoken instruction of a user, a voice processor adapted and configured to process the audio input and to translate the audio input into command terms and a memory having a command database programmed with command terms. The memory is adapted and configured to receive translated command terms from audio input, search the command database for translated command terms, map the translated command terms to programmed command terms, and output a command message for retrieval. The system also includes a command processor adapted and configured to receive the command message from the memory and to perform the command by searching stored built-in test equipment failure data in accordance with the command message.
Owner:HAMILTON SUNDSTRAND CORP

Natural language interactions using visual understanding

Techniques for performing an action with respect to displayed content are described. A natural language interpretation corresponding to a received spoken user input may be determined. Prior to receiving the spoken user input, content may be displayed to the user from which the spoken user input was received. The natural language interpretation may represent a request to perform an action with respect to a portion of the content currently being displayed. Content identifiers corresponding to content being displayed, may be determined, and embedding data representing at least one feature of the content may be determined using the content identifiers. The natural language interpretation and the embedding data may be processed to determine that the spoken user input relates to a first portion of the displayed content instead of a second portion of the displayed content. Based on the determination, an action responsive to the spoken user input may be performed.
Owner:AMAZON TECH INC