Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

6 results about "Received spoken" patented technology

A bidirectional sign language communication system with a multi-sensor glove and haptic feedback

A bidirectional sign language communication system (100) with multisensor glove and haptic feedback, comprising: a wearable glove with a variety of sensors configured to detect hand, finger and wrist movements of a user, wherein the sensors include at least curvature sensors, a pinch detector, a strain sensor, a pressure measuring arrangement and an inertial measurement unit; a processing unit that is functionally coupled with the sensors and configured to process sensor signals and derive sign language tokens using a machine learning model; a wireless communication interface configured to transmit the derived tokens to a companion device for playback as text or speech; wherein the companion device is further configured to receive spoken or textual inputs, convert the inputs into coded tactile patterns, and transmit the patterns to the wearable glove; and a distributed haptic feedback subsystem integrated into the wearable glove and configured to present tactile patterns to the user via fingertip actuators, vibration motors, or a tactile matrix, wherein the system (100) enables bidirectional real-time offline communication between a sign language user and a non-sign language user.
Owner:VIJ RASHVIN HIGHLANDS RANCH

Method and apparatus for generating regression-based gloss-free sign language pose using discrete representation

The present invention relates to a method and apparatus for generating a regression-based gloss-free sign language pose using discrete representation that generates a sign language pose on the basis of (gloss-free) spoken sentence data, and the method comprises the steps of: receiving spoken sentence data as an input and transforming the spoken sentence data into a discretized sign language pose token according to spatial and temporal characteristics by using a self-supervised learning model; and automatically recursively generating a sign language pose through a transformer-based encoder-decoder architecture by using the discretized sign language pose token as an input.
Owner:KOREA ADVANCED INST OF SCI & TECH

Systems and methods for automated speech-to-transaction in healthcare operations

A computer-implemented system and method enable real-time healthcare transaction processing based on ambient audio captured during physician-patient conversations. The system includes one or more microphones to receive spoken dialogue and a speech recognition engine to transcribe the audio into text. A natural language processing module analyzes the text to identify clinical intents corresponding to healthcare transactions. The identified intent is standardized into a structured query format and transmitted to an external system via a communications interface. In response, the system receives external data such as insurance coverage status, cost estimates, or authorization requirements. A user interface displays the external response in real time during the clinical encounter, allowing the physician or patient to make informed decisions. Additional modules may include context monitoring logic to suppress non-actionable utterances and a secure storage engine to log queries and responses. The system automates insurance verification and authorization workflows without requiring manual data entry.
Owner:BLACKSTONE MICHAEL

Natural language interactions using visual understanding

Techniques for performing an action with respect to displayed content are described. A natural language interpretation corresponding to a received spoken user input may be determined. Prior to receiving the spoken user input, content may be displayed to the user from which the spoken user input was received. The natural language interpretation may represent a request to perform an action with respect to a portion of the content currently being displayed. Content identifiers corresponding to content being displayed, may be determined, and embedding data representing at least one feature of the content may be determined using the content identifiers. The natural language interpretation and the embedding data may be processed to determine that the spoken user input relates to a first portion of the displayed content instead of a second portion of the displayed content. Based on the determination, an action responsive to the spoken user input may be performed.
Owner:AMAZON TECH INC

Input system and method

An input system comprises a spoken input processor configured to receive spoken inputs from a microphone operably coupled to the input system; a physical input processor configured to receive physical inputs from a peripheral device operably coupled to the input system, the physical inputs having a timing relative to the spoken inputs; and a speech recognition processor configured to recognise speech from the speech inputs; wherein the speech recognition processor uses at least some of the received physical inputs, and their timing relative to the spoken inputs, as part of the speech recognition process.
Owner:SONY INTERACTIVE ENTERTAINMENT LLC

Natural language interactions using visual understanding

Techniques for performing an action with respect to displayed content are described. A natural language interpretation corresponding to a received spoken user input may be determined. Prior to receiving the spoken user input, content may be displayed to the user from which the spoken user input was received. The natural language interpretation may represent a request to perform an action with respect to a portion of the content currently being displayed. Content identifiers corresponding to content being displayed, may be determined, and embedding data representing at least one feature of the content may be determined using the content identifiers. The natural language interpretation and the embedding data may be processed to determine that the spoken user input relates to a first portion of the displayed content instead of a second portion of the displayed content. Based on the determination, an action responsive to the spoken user input may be performed.
Owner:AMAZON TECH INC