Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

15 results about "Phonetic form" patented technology

In the field of linguistics, specifically in syntax, phonetic form (PF), also known as phonological form or the articulatory-perceptual (A-P) system, is a certain level of mental representation of a linguistic expression, derived from surface structure, and related to Logical Form. Phonetic form is the level of representation wherein expressions, or sentences, are assigned a phonetic representation, which is then pronounced by the speaker. Phonetic form takes surface structure as its input, and outputs an audible (or visual, in the case of sign languages), pronounced sentence.

Multi-modal sign language emotion interaction system and method based on agent architecture

The invention relates to the field of sign language emotion interaction, and provides a multi-mode sign language emotion interaction system and method based on an intelligent agent architecture, and the system comprises a recognition module which is used for recognizing a current task type based on multi-mode data of a user, and generating a task scheduling instruction; the large language model module is used for generating feedback content by utilizing a large language model according to the task scheduling instruction; the emotion module is used for extracting keywords from the dialogue history and environment context information of the user, judging the emotion of the user through an emotion recognition algorithm, generating prompt words with humor / comforting elements by utilizing a large language model in combination with the emotion and the keywords of the user, and sending the prompt words to the user; according to the cue word, utilizing a large language model to select proper language content to generate feedback content with emotion; and the generation module is used for displaying the feedback content with the emotion to the user in a sign language form and / or a voice form through a digital person. According to the method, the defects that an existing system is fragmented and cannot be expanded, and situations cannot be understood are overcome.
Owner:XIAN THERMAL POWER RES INST CO LTD +2

Instrument panel interface content generation method and device, electronic equipment and storage medium

The invention relates to an instrument panel interface content generation method and device, electronic equipment and a storage medium, and the method comprises the steps: obtaining a user instruction which is a natural language instruction in a text form or a voice form; a pre-trained natural language processing model is utilized to analyze the user instruction to obtain a user intention, and the user intention comprises key data items and intention labels obtained by the user intention; target data matched with the key data items are obtained, the target data are processed and displayed according to the intention label, instrument panel interface content is generated, and the instrument panel interface content comprises at least one of a chart and a text abstract. Therefore, the instrument panel interface content can be automatically generated according to the intention of the user, and the instrument panel interface content does not need to be generated by a professional through manual configuration or code writing by a professional tool, so that the generation efficiency of the instrument panel interface content is improved.
Owner:BEIJING QINGSONG YIKANG INFORMATION TECHNOLOGY CO LTD

Chat-bot assisted authentication

Systems and methods for chat-bot assisted authentication are disclosed. The system performing the steps of: receiving, from a user within an application, user input in the form of written language or audible language; receiving a request, from the user, to access an authenticated space within the application; requesting, from the user, user authentication credentials; receiving, from the user, the user authentication credentials; generating a unique token, the unique token comprising the user input and the user authentication credentials; requesting access to the authenticated space, wherein requesting access comprises presenting the unique token; and receiving access to the authenticated space. The system may further send a response to the user based on the user input from the unique token. The system may using natural language processing and / or other machine learning processing to analyze the user's requests.
Owner:TRUIST BANK

system

Provide a system. 【Solution means】 Means for receiving travel-related information from a user, Means for generating a travel plan based on the received information, Means for presenting the generated travel plan to the user and receiving feedback, Means for readjusting the travel plan by reflecting the user's feedback, Means for visualizing and displaying the travel plan on a map, Means for supporting reservations for accommodation facilities and transportation, Means for dynamically adjusting the plan in consideration of real-time information during the trip, Means for providing information in multiple languages, Means for accepting the user's travel request in natural language using voice recognition technology, Means for presenting the generated travel plan in voice using text-to-speech technology, A system including the above.
Owner:SOFTBANK GROUP CORP

Method and device for converting text into voice, medium, electronic equipment and program product

The invention relates to a text-to-voice method and device, a medium, electronic equipment and a program product, and the method comprises the steps: obtaining a session content, the session content comprises a first text and a second text outputted by a dialogue model for the first text, and the second text comprises a text organized by a serial number; according to the first text and the second text, identifying a target language used for playing the serial number in a voice form; and playing the serial number in the form of voice by using the target language, in the disclosure, the influence of the first text and the second text in the session content on the language adopted when the serial number is played in the form of voice is considered at the same time, so that the accuracy and the reliability of the language adopted when the serial number is played in the form of voice are improved.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

Pet language translator

1. The name of the design product: pet language translator. 2. The use of the design product: for identifying the call of a pet, translating into human language, and playing in the form of voice. 3. The design points of the design product: in shape. 4. The picture or photo that best shows the design points: design 1 perspective view 2. 5. Design 1 is designated as the basic design.
Owner:HANGZHOU OUDUN PILOT TECHNOLOGY CO LTD

System

A system is provided.SOLUTION: This system is provided with a means for receiving voice data, a means for converting the received voice data into text data, a means for generating interaction contents based on the text data, a means for outputting the generated interaction contents as voice, and a means for changing the kind of the voice based on the setting of a user.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

Voice conversion method and device based on virtual scene, equipment, medium and product

The application discloses a voice conversion method and device based on a virtual scene, equipment, a medium and a product, and relates to the field of machine learning. The method comprises the following steps: acquiring a natural language command in the form of voice, the natural language command being used for commanding a non-player character, a target virtual character being in a first virtual scene, and the target virtual character comprising at least one of the non-player character and a master virtual character; acquiring a plurality of first scene hot words corresponding to the first virtual scene in which the target virtual character is located, the plurality of first scene hot words being scene-related words of the first virtual scene; and converting the natural language command into a command analysis result in the form of text based on the plurality of first scene hot words. In the above manner, the command analysis result can be more matched with the first virtual scene, and the accuracy of the command analysis result can be improved. The application can be applied to a terminal game scene, a virtual reality scene, an augmented reality scene and other virtual scenes that need to perform a voice conversion process.
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

Method for Assisting a User in Implementing an Examination Workflow of a Magnetic Resonance Examination

The disclosure relates to a method for assisting a user in implementing a workflow of a magnetic resonance examination on a patient. The method may include receiving a query by the user, wherein the input is made in text form or as voice input; determining output information corresponding to the query by means of a large language model (LLM), and providing the output information; and outputting the output information in text form or as voice output.
Owner:SIEMENS HEALTHINEERS AG

Intelligent glasses based on multi-modal vision-language model and environment perception method

The invention discloses intelligent glasses based on a multi-mode vision-language model and an environment perception method, and relates to the technical field of visual impairment assistance, and the intelligent glasses comprise a glasses main body, a camera unit, an edge calculation unit, a bone conduction audio unit, a touch interaction unit, a power supply management unit and a wireless communication unit; the edge calculation module is internally provided with a lightweight multi-modal vision-language model and is used for carrying out preprocessing and semantic analysis on the acquired environment image and generating environment description information, and the bone conduction audio unit broadcasts the environment description information to a wearer in a voice form; the multi-mode vision-language model is used for synchronously processing vision and language features, obstacles, traffic signals, character signboards and scene semantics are effectively recognized, the perception dimension is more comprehensive, rich environment semantic description can be provided, accurate recognition of multiple elements such as obstacles, identifiers and characters in a complex environment is achieved, and the recognition efficiency is improved. And the safety of navigation guidance for visually impaired people is improved.
Owner:GUANGDONG PHARMA UNIV

system

We provide the system. [Solution] A speech recognition means that receives voice input from the user and converts that voice into text, An analysis method that analyzes text data using a natural language processing engine and classifies the content of inquiries into specific categories, A model selection means for selecting a generative artificial intelligence model that generates an appropriate response based on the analysis results, A response generation means that generates a response to a query using a selected generative artificial intelligence model, A response output means that provides the generated response to the user in text or voice, A system that includes this.
Owner:SOFTBANK GROUP CORP

system

We provide the system. [Solution] A means of receiving user instructions as voice using natural language, A speech recognition means for converting the aforementioned instructions from speech to text, A language analysis means for analyzing the transcribed instructions and understanding their intent, Based on the understood intent, a means of using a knowledge base to search for information, A means for generating a response in natural language based on the aforementioned searched information, Means for providing the generated response in audio and visual formats, A system that includes this.
Owner:SOFTBANK GROUP CORP

Interactive language learning system based on semantic scene generation

The invention relates to the technical field of language learning, in particular to an interactive language learning system based on semantic scene generation, and the system comprises a concept building module which is used for building the concept of a target language in the cognition of a learner through the association of visual information and / or auditory information and a unified voice form; the note recording number learning module is used for associating a unified voice form with a corresponding note recording number on the basis of the concept construction module, and training a learner to master the conversion between the voice and the note recording number; and the character learning module is used for training the learner to master spelling rules between the note numbers and the corresponding writing forms according to the note characters. According to the interactive language learning system generated based on the semantic scene, the memory efficiency and durability of vocabularies and syntax are greatly improved, more importantly, the language intuition and direct application ability of a learner are cultivated, and the key transformation from learning of knowledge about languages to learning of the languages is achieved.
Owner:赖诚诚

Quick matching method for fuzzy keywords based on Trie tree

The invention belongs to the technical field of computer information retrieval and data structure application, and relates to a quick matching method of fuzzy keywords based on a Trie tree, comprising a collaborative architecture of a variant Trie tree, a global quick failure bitmap and a multi-modal similarity calculation module, and a variant Trie tree node integrated local quick reachable cache to reuse a successful matching path; the global fast failure bitmap records invalid state-character pairs through 64-bit key values to realize cross-query path pruning; in the multi-modal similarity calculation, Chinese character phonetic form and font characteristics are fused, and an initial confusion matrix and a stroke difference punishment mechanism are combined, so that the matching precision is improved. Through a double-layer cache mechanism and a composite similarity model, the method significantly reduces the calculation redundancy on the premise of not sacrificing the accuracy. The method effectively improves the efficiency and expandability of fuzzy matching, and is suitable for a large-scale text processing scene.
Owner:ASPIRE INFORMATION TECH BEIJING