Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

15 results about "Lexical database" patented technology

A lexical database is a lexical resource which has an associated software environment database which permits access to its contents. The database may be custom-designed for the lexical information or a general-purpose database into which lexical information has been entered.

Construction method of bilingual vocabulary data knowledge graph

The invention relates to the technical field of vocabulary data, in particular to a construction method of a bilingual vocabulary data knowledge graph. The method comprises the following steps: acquiring a vocabulary database; performing data vocabulary preprocessing on the vocabulary database to obtain unstructured vocabulary data; performing nested relationship extraction according to the unstructured vocabulary data to obtain semi-structured vocabulary data; performing label analysis on the semi-structured vocabulary data to obtain initial vocabulary data; performing context coding on the initial vocabulary data by using a preset BERT model to obtain a vocabulary embedding vector; therefore, by introducing a multi-step collaborative mechanism of context coding, cross-language alignment and graph semantic mapping, the technical problems of incomplete structure, semantic disjunction, poor language transformation ability and the like of a traditional vocabulary construction method are solved, and the automatic construction ability and expression depth of a multi-language semantic knowledge system are improved.
Owner:HAINAN VOCATIONAL COLLEGE OF SCI & TECH

Methods and systems for generating textual outputs from images

Embodiments of the present disclosure provide systems and methods for performing text extraction from an image including textual data. The method performed by a processor includes extracting machine-readable textual data from the image. The machine-readable textual data includes one or more words. The method includes comparing each of the one or more words with a dataset including a domain lexicon database and a language dictionary database to determine a first set of words and a second set of words. The first set of words is words successfully matching with words available in the dataset, and the second set of words is words with no successful match with words available in the dataset. Further, the method includes splitting at least one word of the second set of words into two or more words to determine a third set of words and generating a textual output associated with the image.
Owner:MAERSK AS

An LLM text processing method

The present invention relates to an LLM text processing method, belonging to the technical field of text processing. The method includes: obtaining input question information; inputting the input question information into a word segmentation database to output question words; inputting the question words into a preset vocabulary database to output synonymous words, mutually exclusive words, and abbreviated words, and defining the question words, synonymous words, mutually exclusive words, and abbreviated words as query words; outputting final question information according to the query words; using a preliminary retrieval method according to the final question information to find out the reference answer information and the corresponding path; inputting the reference answer information and the path into a priority database to match the priority; determining the preliminary retrieval answer information from the matched priority; using a refined retrieval method according to the preliminary retrieval answer information to find out the refined retrieval answer information; taking the refined retrieval answer information as the final answer text and uploading and displaying it. This application has the effect of improving the accuracy of LLM retrieval.
Owner:SHANGHAI JILIXUN INFORMATION TECH CO LTD

Efficient zero shot event extraction with context definition

A method for zero-shot event extraction, performed by a computer device. The method includes training a context encoder, a first definition encoder, and a second definition encoder with auto extracted context-definition alignment data; retrieving a plurality of verbal synsets from a lexical database; refining a representation model based on the context-definition alignment data and the plurality of verbal synsets; encoding a plurality of candidate event type definitions; encoding the refined representation model with the trained context encoder; and determining whether the encoded representation model belongs to one of the plurality of candidate event type definitions based on a cosine similarity between the encoded representation model, the trained context encoder, the first trained definition encoder, and the second trained definition encoder.
Owner:TENCENT AMERICA LLC

Chinese knowledge text word segmentation method and question answering method

The invention discloses a Chinese knowledge text word segmentation method and question and answer method, and the method comprises the steps: obtaining a Chinese knowledge corpus material corresponding to a business type, and segmenting the Chinese knowledge corpus material to obtain initial text fragments; for each initial text fragment, performing vocabulary extraction on the initial text fragment according to a pre-generated template vocabulary database to obtain template vocabularies in the initial text fragment; extracting non-template vocabularies according to the template vocabularies in the initial text fragment, and analyzing the non-template vocabularies to determine a professional vocabulary corpus; for each specialized vocabulary corpus, constructing a context relation chain based on the specialized vocabulary corpus, and merging the specialized vocabulary corpus according to nodes in the context relation chain to generate segmented vocabularies; the non-template vocabularies in the initial text fragment are merged according to the word segmentation vocabularies, and the word segmentation result of the initial text fragment is determined based on the merging result and the template vocabularies, so that the problem of inaccurate word segmentation of the Chinese text corpus is solved, and the word segmentation accuracy is improved.
Owner:中移信息技术有限公司 +1

Data processing method and system capable of improving efficiency of chinese vocabulary learning

The present invention provides a data processing method and system capable of improving efficiency of Chinese vocabulary learning. The method comprises: a data storage unit storing a vocabulary database related to Chinese language learning; segmenting each word into two or more sub-words; binding each sub-word to a balloon graphic, and assigning a color to the balloon; displaying four or more balloons on a client display screen, wherein one sub-word is displayed on each balloon, and the position of the balloon in the screen changes randomly; detecting an instruction for dragging one balloon in the client display screen so that the one balloon at least partially overlaps another balloon; upon detection of said instruction, indicating that phrase matching is performed on the sub-words bound to the selected balloons; and traversing the data storage unit, if there is a corresponding phrase in the data storage unit, indicating that the matching is successful, removing the balloons from the client display screen, and displaying the phrase in a set area of the client display screen, otherwise, retaining the balloons on the client display screen until all balloons in the client display screen are removed.
Owner:HONG KONG CHINESE COMPUTER SOCIETY CO LTD

Construction method of vocabulary database, speech ability testing method and system

The embodiment of the invention provides a construction method of a vocabulary database. Determining a parent set corresponding to the child set of each age group, feeding back whether each vocabulary in the first data set belongs to the vocabulary capable of being recognized by the corresponding child or not by parents according to the cognitive performance of the corresponding child, and adding the vocabulary into a second data set corresponding to the age group; respectively using the second data set to generate speech ability test questions of the children of the corresponding age groups; and for a tested vocabulary, adding the vocabulary into the target data set corresponding to the age group of which the test accuracy is higher than the target accuracy, and constructing a vocabulary database. By adopting the scheme provided by the invention, a vocabulary testing tool related to the language ability of the Chinese children can be well established.
Owner:BEIJING XIYUE TONGLE EDUCATION TECH CO LTD

Legal translation method and system based on neural network

The invention relates to a legal translation method and system based on a neural network, and is applied to the technical field of legal translations, and the method comprises the steps: receiving a to-be-translated legal document, preprocessing the to-be-translated legal document, and obtaining a to-be-translated document; mapping the to-be-translated literature with a preset vocabulary database, and determining an unambiguous clear vocabulary and a multi-meaning complex-meaning vocabulary; performing association degree dynamic matching on the meaning of the clear vocabulary and the meaning of the remeaning vocabulary according to a preset language model, and determining an optimal vocabulary meaning corresponding to the remeaning vocabulary; and according to a preset functional verification model, carrying out integration verification on the clear vocabulary and the complex-meaning vocabulary, and outputting a translated document.
Owner:CHENGDU LUYOUYOU TRANSLATION SERVICE CO LTD

Text translation method and device, computer equipment and storage medium

The invention relates to the technical field of natural language processing, and discloses a text translation method and device, computer equipment and a storage medium, and the method comprises the steps: obtaining a to-be-translated text, segmenting the to-be-translated text, and obtaining a plurality of segmentation objects corresponding to the to-be-translated text; coding each segmented object to obtain a segmented object coding vector of each segmented object; predicting a translation vocabulary of the coding vector of each segmented object to obtain a translation vocabulary prediction result of each segmented object; determining domain vocabularies in the translation vocabulary prediction result based on the translation vocabulary prediction result and a preset domain vocabulary database; and outputting the translated text of the to-be-translated text based on the translated vocabulary prediction result and the domain vocabulary, thereby improving the vocabulary translation accuracy.
Owner:YUNNAN POWER GRID CO LTD ELECTRIC POWER RES INST

Semantic frame identification using capsule networks

Semantic frame identification involves associating identified target words in the sentential context of their natural language source with semantic frames from a frame lexical database. The disclosed invention leverages the CapsNet architecture for improved semantic frame identification of a target word in a natural language input. This includes deriving the features of a target word identified in the sentence and extracting the features of the word units and the thematic words around the target word. Through dynamic routing of capsules, the CapsNet is able to filter the candidate frames for the target word to reduce the search space and apply the CapsNet prediction to identify a frame from a frame lexical database.
Owner:COGNIZER INC

A classification prediction method driven by label information enhancement and strong negative samples

This invention discloses a classification prediction method driven by label information enhancement and strong negative samples. The method comprises the following steps: first, constructing high-quality text by truncation of text and matching the content of a workbook; second, considering the informativeness of sentiment labels, using a vocabulary database and a professional vocabulary to enrich the sentiment label representation; third, using strong negative sampling to distinguish each text from strong negative samples; and finally, using a prediction loss function that incorporates sentiment label information and a contrastive loss function that incorporates strong negative sampling to ensure model accuracy. This method integrates text truncation, sentiment label information enhancement, and strong negative sampling to improve the accuracy of classification prediction.
Owner:ANHUI UNIV

Service system optimization method and device and electronic equipment

The invention provides a business system optimization method and device and electronic equipment, and the method comprises the steps: constructing a vocabulary database which at least comprises professional vocabularies, emotion vocabularies and modification vocabularies; a text input by a user is obtained, the vocabulary database is matched with the text input by the user, target professional vocabularies, target emotion vocabularies and target modification vocabularies in the text input by the user are extracted, a service system corresponding to the target professional vocabularies is determined, and the service system is used for handling financial services; obtaining the emotion intensity of the target emotion vocabulary and the modification intensity of the target modification vocabulary, determining the emotion of the user according to the emotion intensity and the modification intensity, and optimizing the service system according to the emotion of the user, so that the optimized service system meets the requirements of the user. Through the method, the user emotion can be accurately determined, so that the service system can be optimized according to the user emotion, and the service system can better meet the requirements of the user.
Owner:中国邮政储蓄银行股份有限公司

System and method for modification, personalization and customization of search results and search result ranking in an internet-based search engine

A computer server system and method are disclosed for personalization and customization of network search results and rankings, such as for Internet searching. A representative server system comprises: a network interface to receive a query from a user and transmit return queries and search results; a data storage device having a first, lexical database having one or more compilations and templates; and one or more processors configured to access the first database and search a selected compilation using the query to generate initial search results; to comparatively score each selected parsed phrase of the initial search results, for each classification of a selected template and a selected compilation, and to output initial and final search results arranged according to the classifications and the predetermined order of the template. A representative embodiment may also include use of a second, semantic database having multi-dimensional vectors corresponding to parsed phrases, paragraphs, or clauses.
Owner:XMENTIUM INC

Device for medical assistance and inclusion of vulnerable people in natural language.

UndeterminedCO20260008848U1Electric networkBody movement
This invention falls within the field of electronics, computer science, and artificial intelligence related to medical conditions and disability. The invention aims to address the needs of people with disabilities or vulnerabilities, such as the blind, mute, or ill, enabling them to ask questions and receive alerts in natural language, even while sleeping, regarding health variables related to the heart, lungs, and temperature. This information can be accessed in real time without physical contact, internet access, or electricity. It is particularly useful for people living in remote, vulnerable areas without access to these services, or during emergencies such as power outages and weather events that disrupt public grid power or internet or artificial intelligence services for diagnosing and monitoring vital functions and medical conditions.To solve the problem, the Device for medical assistance and inclusion of vulnerable people in natural language comprises a sound sensor (7) connected to a natural language processing unit (4) which is connected to a database of words (8) related to medical emergencies, where the natural language processing unit (4) is configured as a means to transform speech into text and is connected to a millimeter wave processor (2) configured to obtain variables of heart rate, respiration and body movement among a plurality of other variables.
Owner:GALVEZ RENDON WALTER ROLANDO