Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

9 results about "Chinese word" patented technology

Rule-based enhanced chinese word segmentation and semantic unit parsing method and system

PendingCN122311199AEngineeringChinese word
This application relates to the field of data processing technology and discloses a method and system for Chinese word segmentation and semantic unit parsing based on rule enhancement. The method includes: constructing a multi-level domain rule base from forced matching rules, pattern template rules, and context constraint rules to generate a set of rule triples; loading the set of rule triples into an AC automaton to linearly scan the input text to obtain a set of candidate segments and trigger word position indices; constructing a candidate directed acyclic graph, and obtaining a structured word segmentation sequence and a pruned candidate pool through dynamic programming after dynamic modulation and fusion scoring; verifying the integrity of downstream slots, and locating backup candidate edges in the pruned candidate pool using gap signals when failure occurs, and obtaining a corrected word segmentation sequence through renegotiation and scoring. This application improves the completeness of professional terminology recognition in professional vertical scenarios and the self-correction capability of cross-domain word segmentation results.
Owner:TIANJIN FEIPENG SHENGYUAN TECHNOLOGY DEVELOPMENT CO LTD

A brain-computer interface system for recognizing the intention of Chinese oral language based on a sound-meaning integration double model

ActiveCN121560160BSpoken languageStereotaxis
The application provides a Chinese spoken language intention recognition brain-computer interface system based on a sound-meaning integration double model, belongs to the technical field of biomedical engineering, and relates to language brain-computer interface technology. Taking sound-meaning integration as the core, the stereotactic intracranial electroencephalogram (sEEG) technology is adopted to collect neural signals of the brain articulatory motor coding area and the semantic concept organization coding area. The system comprises a voice initiation decoder, a speech decoder, a semantic decoder and a Chinese word speech-semantic fusion synthesizer, the target decoder is constructed by extracting high gamma band features of key brain areas of the frontal lobe (left inferior frontal gyrus, premotor cortex, etc.), the temporal lobe (anterior temporal lobe, dorsolateral temporal lobe, etc.). At the same time, a visual and auditory induction training paradigm is matched, three tasks of listening to sound to group words, looking at words to group words and word association are set, and the subjects are supported to generate words independently. The system effectively solves the homonym and near homonym word ambiguity problem in Chinese spoken language recognition, and provides a precise interactive tool for ALS and other speech disorder patients.
Owner:BEIJING TIANTAN HOSPITAL AFFILIATED TO CAPITAL MEDICAL UNIV

A method for generating a pinyin bucket word library of a four-level index

PendingCN122452553ADigital dataData integrity
The present application relates to the technical field of electronic digital data processing, and discloses a four-level index pinyin bucket word library generation method, acquires a Chinese word library text, extracts the first Chinese character and the first pinyin of each line of words as a classification key, and generates a triple; the triple is classified into a corresponding pinyin bucket according to the pinyin, the words in the bucket are sorted in descending order of word frequency, an independent Chinese character block is generated for each unique Chinese character, and the words are solidified according to the word length by using a first character multiplexing mechanism; a four-level static index of an initial letter statistical area and a pinyin index area and a Chinese character word index area and a word storage area is sequentially constructed, each index area uses fixed-length entries and absolute offset addressing; a metadata area containing a file length double-semantic field is constructed, a high-frequency-first deterministic truncation is performed on the word library according to a preset threshold, each data area is spliced and a data integrity check value is appended, and an embedded binary word library is generated. The problems of high storage redundancy and uncontrollable memory are solved, and the purposes of deterministic analysis, resource adaptation and high security are achieved.
Owner:SICHUAN HAIGE HENGTONG PRIVATE NETWORK TECH CO LTD

A multi-language word vector weighting alignment method based on orthogonal pach analysis

This invention relates to the field of natural language processing technology and discloses a multilingual word vector weighted alignment method based on orthogonal Protodyakonov analysis. The method includes the following steps: S1: Randomly initialize an orthogonal matrix R, and randomly rotate the Chinese word vectors, i.e., X < -XR; substitute the weight matrix A, calculate the weighted orthogonal transformation matrix W according to the provided weighted Protodyakonov analysis, and obtain the weighted aligned Chinese word vector XW2. The Chinese word vectors are downloaded through a given word vector link to obtain Chinese and English word vectors X and Y. A given alignment dictionary L is used to weight sentiment words, with the goal of weighted alignment from X to Y. The method provided by this invention can perform weighted alignment according to the needs of downstream tasks, further improving the performance of downstream tasks, realizing a shared word vector space, and providing a mathematical proof of the weighted orthogonal Protodyakonov analysis.
Owner:GUANGZHOU UNIVERSITY

A Power Grid Transfer Verification Method Based on Digital Twins

PendingCN122338809ATransient statePower grid
This invention discloses a power grid transfer verification method based on digital twins, relating to the field of power grid dispatching technology. The method includes the following steps: First, a dual-track isolation architecture of associated and asynchronous coexisting modes is constructed. Data sections are captured and extracted from the associated mode. Then, unstructured dispatching operation text is processed through Chinese word segmentation, stop word removal, normalized object standardization, and action vector extraction to generate a structured underlying operation script matrix. Subsequently, in the asynchronous coexisting mode, single-step dispatching operations are executed line by line according to the script matrix, and topology analysis, electromechanical transients, and dynamic power flow cross-verification are performed after each micro-state switching action. When overload or hidden risks are identified, safety blocking, full-network load rebalancing, rewriting of the original dispatching command control script, and re-verification are executed. This invention can improve the accuracy and closed-loop handling of transfer verification and enhance the consistency of on-site execution without interfering with the monitoring of the physical power grid operation.
Owner:XUANCHENG POWER SUPPLY OF ANHUI ELECTRIC POWER CORP

A standardized dark chain sample construction and label annotation system and method thereof

PendingCN122293404AIp addressTranscoding
This invention discloses a standardized dark link sample construction and labeling system and method, including a task scheduling module, a memory management module, a DNS caching module, a web page asynchronous download module, and a content recognition and detection module. The main functions are: retrieving the website attributes to be detected from the detection task table, including the URL address and scheduling frequency; retrieving the network address corresponding to the URL using DNS caching technology; issuing an alarm if an anomaly is detected; constructing an HTTP request to the server and downloading the web page; analyzing the web page content, performing encoding recognition on the web page content, and using Chinese and English word segmentation recognition algorithms to perform Chinese word segmentation on the transcoded content; issuing alarms for dark links; and issuing alarms for typos. This invention employs a method where, upon program startup, the scope of sites to be monitored is obtained from the task scheduling node based on the local IP address and computing power; simultaneously, the detection frequency of each site is obtained, and alarms are issued for detected abnormal content.
Owner:陈昭奕

Nuclear power valve design document information recognition and collection method and system

ActiveCN120913222BPart of speechSoftware engineering
The application provides a nuclear power valve design file information recognition and collection method and system, the collection method comprising receiving a nuclear power valve design file, and pre-processing the nuclear power valve file based on a CNN valve design file image recognition model to meet the format and geometric requirements, to obtain a target valve design file; identifying the information of the to-be-identified area of the target valve design file based on a CRNN valve drawing information recognition model; creating a corpus analysis library using the recognition result of step 3, and carrying out part-of-speech tagging and Chinese word segmentation processing on the corpus analysis library; carrying out text analysis on the processed corpus result of the nuclear power valve of step 4, extracting high-frequency parameters and key parameters as the processing basis for subsequent text keyword extraction, and carrying out text vectorization processing; filtering invalid information according to the text vectorization similarity, judging the correctness of the key information that has completed information integration, and then operating the correct information into the valve drawing information database.
Owner:SHANGHAI NUCLEAR ENGINEERING RESEARCH & DESIGN INSTITUTE CO LTD

Emotion triple extraction method and device

The application provides a kind of sentiment triple extraction method and device, wherein the method comprises: obtaining a text to be evaluated;The text to be evaluated is input into the extraction model, and the sentiment triple extracted by the extraction model is obtained;Wherein, the extraction model is trained based on the fragment text sample, the text combination formed by the fragment text sample and the collocation label corresponding to the text combination, and the collocation label is determined according to the text combination in advance;The extraction model is used for extracting the sentiment triple of the text to be evaluated based on the semantic features and Chinese word segmentation features of the text to be evaluated.The sentiment triple extraction method and device provided in the embodiment of the application can combine the semantic features and Chinese word segmentation features of the text to be evaluated, and improve the accuracy of Chinese text sentiment triple extraction.
Owner:TSINGHUA UNIVERSITY