Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

11 results about "WordNet" patented technology

WordNet is a lexical database for the English language. It groups English words into sets of synonyms called synsets, provides short definitions and usage examples, and records a number of relations among these synonym sets or their members. WordNet can thus be seen as a combination of dictionary and thesaurus. While it is accessible to human users via a web browser, its primary use is in automatic text analysis and artificial intelligence applications. The database and software tools have been released under a BSD style license and are freely available for download from the WordNet website. Both the lexicographic data (lexicographer files) and the compiler (called grind) for producing the distributed database are available.

Network threat report attack knowledge graph automatic construction method and system based on large language model, storage medium and program product

The invention relates to a network threat report attack knowledge graph automatic construction method and system based on a large language model, a storage medium and a program product, and the method comprises the steps: carrying out the iterative processing of extracted entities and relationships through a clustering method based on the large language model, and generating an initial attack knowledge graph; according to the invention, aggregation processing is carried out on a plurality of technology example threat reports belonging to the same attack technology through an attack technology graph template generation mechanism, a standardized attack technology standardized template library is established, and automatic attack technology tagging of new threat reports is realized by adopting an attack technology alignment method based on Word2Vec and WordNet, so that the automatic attack technology tagging of the new threat reports is realized. And a complete attack technology knowledge graph is constructed, so that a security analyst can quickly understand an attack path and a key threat point, and the automation degree and the accuracy of threat intelligence analysis are remarkably improved.
Owner:STATE GRID SHANGHAI MUNICIPAL ELECTRIC POWER CO +1

Refined password structure generation method

The invention discloses a refined password structure generation method. The method comprises the following steps: inputting a password and identifying a special character string in the password; segmenting and marking the unmarked password according to the type to obtain a letter segment L, a number segment D and a symbol segment S; finely classifying the divided letter segments L; and obtaining a fine structure of the password according to a fine classification result. According to the scheme, the WordNet dictionary system is utilized to perform full-coverage semantic recognition and classification on the English character strings, the problem of insufficient model precision caused by semantic information loss in a traditional PCFGs password structure analysis method is solved, and a high-practicability refined structure representation and generation method is provided for password analysis.
Owner:UNIV OF ELECTRONICS SCI & TECH OF CHINA

A signature-based set semantic similarity connection method

The application relates to a signature-based set semantic similarity connection method, and belongs to the fields of databases and information retrieval. The method comprises four parts: firstly, a classification tree construction step: a classification tree is constructed according to a WordNet knowledge base given a data set; secondly, a set signature step: each set in the data set is signed to obtain a corresponding signature data set; thirdly, a data preprocessing step: the sets in the signature data set are sorted to obtain a sorted data set; and finally, a connection processing step: self-connection is performed on the sets in the sorted data set to obtain a set of semantic similarity results. The method is based on signature prefix filtering technology and length filtering technology, and finally realizes the set semantic similarity connection method, so that the set semantic connection efficiency can be effectively improved.
Owner:KUNMING UNIV OF SCI & TECH

A method, system, storage medium and program product for automatically constructing a network threat report attack knowledge graph based on a large language model

The present invention relates to a method, system, storage medium and program product for automatically constructing a network threat report attack knowledge graph based on a large language model. The method generates an initial attack knowledge graph by iteratively processing the extracted entities and relationships through a clustering method based on a large language model. At the same time, the present invention aggregates multiple technical example threat reports belonging to the same attack technology through an attack technology graph template generation mechanism, establishes a standardized template library for standardized attack technologies, and adopts an attack technology alignment method based on Word2Vec and WordNet to realize automatic attack technology labeling of new threat reports, construct a complete attack technology knowledge graph, facilitate security analysts to quickly understand attack paths and key threat points, and significantly improve the automation and accuracy of threat intelligence analysis.
Owner:STATE GRID SHANGHAI MUNICIPAL ELECTRIC POWER CO +1

Hallucination detection system, hallucination detection method, and electronic device

An electronic device comprises a hallucination detection system. The hallucination detection system can execute a hallucination detection method. The hallucination detection method comprises implementing a conversion step, an authenticity detection step, and a result output step. The conversion step comprises: generating a plurality of output wordnets according to the output text. The authenticity detection step comprises: generating multiple pieces of authenticity information by a trained graph embedding model. Each of the multiple pieces of the authenticity information indicates whether one of the output wordnets is determined to be non-hallucinated or hallucinated by the graph embedding model. The result output step comprises: generating result of hallucination detection according to the multiple pieces of the authenticity information. The graph embedding model is trained through a plurality of positive wordnets and a plurality of negative wordnets generated from domain knowledge data.
Owner:INSTITUTE FOR INFORMATION INDUSTRY

A method and system for improving completion parameters for understanding user input

PendingCN122114174AAccurately capture potential needsavoid misreadingDigital data information retrievalNatural language data processingUser inputEngineering
The application relates to the technical field of industrial manufacturing, and discloses a method and system for improving the completion parameters of understanding user input, which comprises the following steps: inputting an industrial standardized file, generating a semantic label, reasoning through the industrial standardized file and the semantic label, obtaining an optimized industrial standardized file, and obtaining an industrial rule set and a text parameter set input by a user. The application realizes accurate semantic label generation based on a WordNet synonym set, solves the polysemy ambiguity problem, calculates the similarity between a candidate parameter and a user conversation through an improved Levenshtein distance algorithm and a compensation mechanism, accurately captures the potential demand of the user, improves the consistency with the actual input intention of the user, multiplies the initial weight value, the click rate and the time decay factor to sort the candidate parameters, balances the instant production demand and the historical experience in the industrial scene, and significantly improves the accuracy, adaptability, efficiency and interpretability in four dimensions.
Owner:BEIJING INFORMATION TECH BOTE INTELLIGENT TECH CO LTD

Method and device for measuring word semantic similarity by integrating SAO and Bayesian model

ActiveCN119692358BSemantic analysisMeasure wordData mining
The present invention provides a method and apparatus for measuring the semantic similarity of words by integrating the SAO and Bayesian models, which relates to the technical field of calculating the semantic similarity of words. This method is an innovative method for measuring the semantic similarity of words, which integrates the extraction of the SAO (subject-action-object) structure and the Bayesian model. First, the SAO structure is extracted from the text, and then the occurrence frequencies of similar words are statistically counted based on these structures, which is used as a measure of the similarity of words. At the same time, semantic knowledge bases such as WordNet are also utilized to obtain the semantic similarity parameters of words, and these parameters and statistical results are integrated through the Bayesian model to calculate the posterior probability of the semantic similarity of words. The innovation of this method lies in that it not only utilizes the advantages of the statistical-based method but also combines the knowledge-based method, providing a new perspective to understand and quantify the similarity between words.
Owner:XIAMEN UNIV OF TECH

A Weighted Bayesian Classifier and Ontology Mapping Method for Semantic Ontology

The present invention claims protection for a weight Bayesian classifier and an ontology mapping method for semantic ontology, which comprises the following steps: S1: Parsing the ontology: Parsing the ontology participating in the ontology mapping through the Jena package in Java, extracting concepts, and using the WordNet semantic dictionary to complete the expansion of concepts; S2: Determining weights using the harmony search algorithm: Introducing a dynamic weight mechanism to reduce the impact brought by its assumption of conditional independence of attributes; S3: Constructing a weight Bayesian classifier: Establishing a weight Bayesian classifier through the weights obtained in the previous step, converting the ontology mapping problem into a classification problem, calculating the maximum posterior probability to complete the classification; finally, saving the output classification result into the AllergoGraph database. The present invention adopts the harmony search algorithm, uses the classification error rate as the objective function, calculates the global optimal weight and assigns it to the weighted naive Bayesian classifier; the present invention proposes a new way of generating solutions to improve the optimization ability of the harmony search algorithm.
Owner:CHONGQING UNIV OF POSTS & TELECOMM

A zero-shot event detection method based on a text entailment recognition model

The application discloses a zero sample event detection method based on a text entailment recognition model, and comprises the following steps: S1, obtaining initial text data from a NYT corpus and performing pretreatment, and obtaining separate sentence texts after the treatment; S2, using a WordNet to expand seed keywords; S3, using the expanded keyword set to screen the text data, when a text sentence contains keywords corresponding to a certain event type, the sentence has a considerable probability to express the event type; S4, using a text entailment recognition model to mark the data screened in the step S3; S5, further using the text entailment recognition model to screen the data obtained in the step S4; S6, using the data obtained in the step S5 to train the text entailment recognition model, and using the trained model to perform event detection on an ACE data set. The application focuses on optimizing the application of the text entailment recognition model in zero sample event extraction.
Owner:SOUTH CHINA UNIV OF TECH

A method and system for generating medical text data

The present invention belongs to the field of information technology and relates to a method for generating medical text data, which includes obtaining a medical record sample, obtaining a new medical record sample by using MetaMap for professional medical name matching and replacement, and at the same time using the WordNet dictionary for synonym search and obtaining the current medical record sample by using the synonym replacement with the highest similarity, thus expanding the medical record sample data and solving the problem of data shortage in medical text classification training.
Owner:CHONGQING UNIV OF POSTS & TELECOMM

Image Clustering Method Based on Image-Text Pretrained Model

The present invention provides an image clustering method based on a text-image pre-trained model, belonging to the technical field of image processing. The method includes: according to all nouns obtained from WordNet, using the text-image pre-trained model to select candidate words in the text modality, retrieving the candidate words for each image, and constructing a representation corresponding to it in the text modality; according to the image representation and the representation corresponding to it in the text modality, using the method of cross-modal mutual distillation to obtain the image clustering result. The present invention solves the problem that the text modality of the text-image pre-trained model cannot be effectively utilized in the case of unknown class names, resulting in poor image clustering effect and large computational overhead, and avoids consuming a large amount of human resources for image classification.
Owner:SICHUAN UNIV