Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

8 results about "Function word" patented technology

In linguistics, function words (also called functors) are words that have little lexical meaning or have ambiguous meaning and express grammatical relationships among other words within a sentence, or specify the attitude or mood of the speaker. They signal the structural relationships that words have to one another and are the glue that holds sentences together. Thus they form important elements in the structures of sentences.

Deep learning-based ancient modern text machine translation method

The invention discloses an ancient modern text machine translation method based on deep learning, and belongs to the technical field of natural language processing. The invention provides a multi-task collaborative optimization framework aiming at the problems of domain knowledge deficiency, frequent word activation, corpus scarcity and the like of an existing neural network model in ancient language translation. Precise domain knowledge injection is realized by constructing a closed ancient language parallel corpus retrieval library and combining historical term alignment and logic chain reasoning; a probability weighted part-of-speech embedding mechanism is adopted, and grammar feature distribution of virtual word ambiguity and word type activity is dynamically fused, so that the grammar analysis robustness is improved; and semantic generation, retrieval alignment and part-of-speech constraint tasks are jointly optimized through a fixed weight strategy. According to the method, the semantic accuracy and the logic continuity of ancient language translation are improved, and extended application of complex ancient language styles is supported. Experimental results verify that the performance of the method is improved in terms of BLEU and ROUGE-L indexes, and efficient technical support is provided for ancient book digitization and cross-time explanation.
Owner:ZHONGBEI UNIV

A text pronunciation optimization method and system for speech synthesis

PendingCN122290565AFunction wordSpoken language
This invention discloses a text pronunciation optimization method and system for speech synthesis, belonging to the field of speech synthesis management technology. In this method, after emotional processing, long sentences are split into segments based on a triple rule system of semantic blocks, classical Chinese function words, and character length. The English portion of the text undergoes layered processing, distinguishing between pure uppercase letter combinations and regular English words, adding splitting and prosodic markers to pure uppercase letter combinations, and generating optimized text pronunciation information. This invention breaks through the limitations of existing single-rule adaptation in speech synthesis text processing, pioneering a multi-dimensional layered optimization framework that integrates technologies such as semantic parsing of numbers and operators, scene-based matching of polyphonic characters, emotional markers for literary texts, and semantic long sentence splitting. It designs multiple exclusive optimization rules for the speech synthesis and reading needs of literary and everyday spoken texts, solving the core pain point of existing technologies that emphasize generality but neglect specific scenarios.
Owner:DEEP THINKING (HANGZHOU) DATA CO LTD

Method and device for training language model, equipment and medium

The invention provides a method and device for training a language model, equipment and a medium. In one method, a plurality of training samples are received, a training sample of the plurality of training samples including a reference cue word and a reference response, the reference cue word including a dotted word and a notional word, and the reference response including a dotted word and a notional word. A first language model is obtained, the first language model describes the incidence relation between the cue word and the response for the cue word, and the first language model comprises a plurality of nodes; in the pre-training stage, a plurality of parameters of a plurality of nodes of a first language model are updated based on a plurality of training samples to obtain a second language model. And in the updating process, based on the plurality of training samples, updating a first group of parameters of a first group of nodes corresponding to the virtual words in the plurality of nodes. And based on the plurality of training samples, updating a second group of parameters of a second group of nodes corresponding to the notional word in the plurality of nodes, the second number of the second group of nodes being less than or equal to the first number of the first group of nodes.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

Function word extraction system and method based on multi-modal adaptive learning

The invention relates to an efficacy word extraction system and method based on multi-modal adaptive learning. According to the system, multi-modal data such as texts, images and audios are received through the input module, and feature fusion is carried out through the preprocessing module. A hybrid encoder module is utilized, local n-gram features are firstly extracted through a convolutional neural network layer, then global context information is extracted through a Transform encoder, the comprehensiveness of feature representation is guaranteed through a cooperation mechanism, and the limitation of a traditional single model in the aspect of feature representation precision is overcome. And then, the knowledge distillation module optimizes the text semantic representation data to generate efficient lightweight semantic representation information. The decoder module integrates a vocabulary library, adopts a strategy of combining generation and copying, and utilizes the thought of a pointer-generation network, so that the problem of unregistered words is effectively solved, the coverage rate and accuracy are ensured while generalization is kept, and efficient, accurate and robust efficacy word extraction is realized.
Owner:QIZHI TECH CO LTD

Visual language model training methods, image and text analysis methods, and related devices

This invention belongs to the field of image and text processing, and discloses a visual language model training method, image and text analysis method, and related apparatus. It obtains the function word mask of text samples based on a pre-set function word dictionary; utilizes the visual encoder and text encoder of the visual language model to obtain the corresponding image features and complete text features of image and text samples, as well as the function word text features obtained from the aforementioned function word mask; calculates the text-image cross-attention and function word-image cross-attention using the aforementioned image features, complete text features, and function word text features, and then obtains the final differential cross-attention through difference; then calculates the training loss and updates the parameters through backpropagation gradients; by repeating the above steps, the visual language model is trained, resulting in a visual language model with low training cost, almost lossless accuracy, and improved robustness, thereby enhancing the image and text analysis capabilities of the visual language model and providing an optimized foundation for the security protection of the visual language model system.
Owner:XI AN JIAOTONG UNIV

Tibetan rhythm structure prediction method based on grammar information

The invention discloses a Tibetan rhythm structure prediction method based on grammar information, relates to the technical field of speech synthesis, and is applied to the field of rhythm structure prediction in Tibetan speech synthesis. The syntactic information-based Tibetan rhythm structure prediction method comprises the following steps of S1, completing Tibetan analysis and virtual word continuous quantization, extracting a grammar probability, a hierarchical focus and a pause difference, and storing and constructing a Tibetan rhythm prediction database after preprocessing; s2, based on the strength of the virtual word continuing relation, the influence range hit mark and the word order distance, boundary discrimination is carried out; s3, entropy feature analysis is carried out through the grammar role probability data and the boundary judgment data; s4, generating a three-layer rhythm boundary type sequence according to boundary forming, and inhibiting over-dense pause; and S5, carrying out rhythm intensity evaluation through the pause duration difference and the grammar information entropy data. The problems of pause, confusion and increased understanding burden caused by dislocation of rhythm and semantic levels in Tibetan long speech synthesis are solved.
Owner:TIBET UNIV

Data interaction method and apparatus, and computer device and storage medium

The present application relates to a data interaction method and apparatus, and a computer device and a storage medium. The method includes: constructing function words and command words, which are related to communication information configured for performing interaction with a software system; setting a channel for publishing and subscribing to a remote dictionary server database monitoring the function words and the command words on the channel in real time to identify a changed function word and / or command word; determining whether the changed function word and / or command word is data information required or maintained by a target process; if yes, maintaining corresponding communication information to a storage space of the target process; writing the obtained communication information into the remote dictionary server database; and publishing the function words and the command words onto the channel for use by other processes.
Owner:INSPUR SUZHOU INTELLIGENT TECH CO LTD

Tibetan dialect language conversion method

PendingCN120542377ANatural language data processingFunction wordSpoken language
The invention discloses a rule-based Tibetan dialect language conversion method. The method comprises the following steps: S1, constructing a dialect and virtual word knowledge base, spoken words and a virtual word rule base by using dialect data and Tibetan grammar; s2, establishing a conversion model by using the constructed virtual word rule base; s3, collecting data by using the conversion model, and carrying out data preprocessing; and S4, based on the processed data, using a conversion model to carry out dialect language recognition, and then outputting a written language. Compared with the prior art, the rule-based Tibetan dialect language conversion method has the advantages that the detailed Tibetan language and virtual word rule base is constructed, so that the conversion model is formed.
Owner:TIBET UNIV