Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

196 results about "Semantic integrity" patented technology

Semantic integrity. Semantic integrity ensures that data entered into a row reflects an allowable value for that row. The value must be within the domain, or allowable set of values, for that column. For example, the quantity column of the items table permits only numbers.

Hierarchical semantic-driven retrieval enhancement generation method and system

The invention discloses a hierarchical semantic-driven retrieval enhancement generation method and system, a natural hierarchical relationship and a semantic boundary of a document are effectively reserved by constructing a tree hierarchical structure based on a document chapter title, and a recursive semantic boundary splitting strategy is adopted to refine overlong text nodes, so that the semantic integrity is ensured, and the retrieval enhancement generation efficiency is improved. And the model input length limitation is met, and the semantic information is prevented from being lost. Meanwhile, node knowledge point extraction and abstract generation are achieved through a large language model, top-down multi-level title path transmission and bottom-up content aggregation are combined, the structural perception and semantic expression ability of nodes is enhanced, and in the retrieval stage, based on similarity distribution of query and node semantic expression, an adaptive retrieval threshold value is dynamically calculated, and the retrieval efficiency is improved. A fixed top-k retrieval strategy is replaced, intelligent screening of different query and hierarchical nodes is achieved, information coverage and redundancy suppression are balanced, and retrieval efficiency and accuracy are remarkably improved.
Owner:NORTH CHINA UNIVERSITY OF TECHNOLOGY

Semantic enhancement adaptive partitioning method and system for natural resource large model questions and answers

The invention provides a semantic enhancement adaptive partitioning method and system for natural resource large model questions and answers, and aims to solve the problems of difficulty in term boundary recognition, damage to semantic integrity and the like. According to the method, three core technologies including theme perception coarse-grained paragraph division, self-adaptive sliding window theme hierarchy division and embedded perception context self-adaptive text segmentation are fused. The system firstly analyzes a natural resource long text structure, identifies titles and theme levels and aligns associated contents; paragraphs are extracted according to a theme perception strategy and are subdivided into sentence sets according to grammar rules; an improved sliding window mechanism is adopted to divide sentences into window sentence block groups. The method is characterized in that a dynamic aggregation threshold mechanism is introduced, the semantic association degree between adjacent sentence blocks is calculated through an embedded perception context semantic segmentation technology, whether the sentence blocks are combined or not is judged by combining a similarity distribution change trend and a dynamic adjustment threshold, self-adaptive delimitation of semantic boundaries is achieved, and text blocks which are clear in structure and coherent in semantics are generated.
Owner:HUBEI PROVINCIAL DEPT OF NATURAL RESOURCES INFORMATION CENT +1

PDF text extraction method and system based on large language model

The invention relates to the field of document processing and data extraction, and particularly discloses a PDF text extraction method and system based on a large language model.The method includes the steps that content of all pages of a target PDF document is positioned and marked to obtain a first to-be-recognized area and a second to-be-recognized area, and noise interference features of the to-be-recognized areas are removed; formulating a multi-level text logic reconstruction strategy to complete reconstruction of a logic sequence of the target PDF document, preliminarily outputting a first-level PDF document, and performing primary image-text association degree analysis to output first association strength; performing intelligent anomaly recognition and correction on the content of the first-level PDF document on a semantic structure through a large language model to obtain a second-level PDF document, and outputting second association strength; judging whether the secondary PDF document is qualified or not based on the first association strength and the second association strength; according to the method, the logic sequence and the semantic integrity of the document can be recovered, and the text purity and the structural integrity are improved.
Owner:NANJING WEISHIDE SOFTWARE CO LTD

Data processing method, data processing model acquisition method, data processing model acquisition device, data processing equipment and medium

The invention relates to the technical field of data processing, and discloses a data processing method, a data processing model obtaining method, a data processing model obtaining device, data processing equipment and a medium. According to the data processing model obtaining method, preprocessing is carried out on a multi-source heterogeneous data set, a standard data set is obtained, multi-modal feature extraction processing is carried out on the basis of the standard data set, and a multi-modal feature set is obtained; performing cross-modal alignment processing on the multi-modal feature set to generate a cross-modal alignment feature set, performing fusion processing based on the modal alignment feature set to generate a fusion feature set, and inputting the fusion feature set into the to-be-trained model to perform training to obtain a data processing model. According to the method provided by the invention, the adaptation capability of the data processing model to multi-source heterogeneous data and the expression capability of the data processing model to a high-dimensional structure are enhanced while the multi-modal semantic integrity is ensured, so that the semantic understanding and reasoning judgment performance of the model in a complex scene is remarkably improved.
Owner:SHENZHEN KUAILU TECH CO LTD

Intelligent text retrieval method and system, storage medium and program product

The invention provides an intelligent text retrieval method and system, a storage medium and a program product, and relates to the field of intelligent text retrieval, the method comprises the following steps: segmenting each original document into a plurality of semantic segments according to semantic similarity; extracting a text feature of each semantic fragment to generate a text fragment vector; selecting a related fragment from each text fragment vector, wherein the first similarity between the related fragment and the retrieval vector is greater than a preset first threshold value; calculating a second similarity between each related fragment and the adjacent fragment; combining the adjacent segments with the second similarity greater than a preset second threshold value with the corresponding related segments to obtain stitching segments; based on a retrieval result corresponding to a historical retrieval request of a user, determining a retrieval granularity which is firstly displayed; and according to the retrieval granularity, adjusting the display content of the stitching fragment, and returning a text retrieval result. By implementing the method, the semantic integrity and the result relevancy of the text retrieval result can be improved.
Owner:NANJING WEISHIDE SOFTWARE CO LTD

Intelligent agent visual language navigation method and system based on task completion prediction

The invention provides an agent visual language navigation method and system based on task completion prediction. The method comprises the step of constructing a dual-drive structure composed of a self-adaptive mixed pooling mechanism and a task completion analysis module. Firstly, in the visual information processing process, a dynamic weight distribution strategy is adopted to carry out multi-scale adaptive mixed pooling on panoramic features, so that the fusion effect of local and global semantic information is optimized, and the retention capability and semantic integrity of navigation historical information in a dynamic topological map are improved. And then, inspired by a human navigation cognitive behavior mechanism, a task completion analysis module is designed and introduced, and the task execution progress is dynamically estimated based on the recognition condition of a key landmark in a navigation path, so that an intelligent agent is driven to preferentially select a key path node and invalid exploration is reduced. And finally, realizing efficient understanding and execution of the natural language instruction by the intelligent agent through a multi-round cyclic cross-modal reasoning and action prediction mechanism.
Owner:FUZHOU UNIV

Speech semantic analysis method and device and storage medium

The invention discloses a speech semantic analysis method and device, and a storage medium. The method comprises the steps of determining a real-time speech recognition result corresponding to frame-by-frame input speech; inputting the real-time voice recognition result and the context information into a semantic integrity judgment model to judge whether the real-time voice recognition result forms a complete semantic unit or not; and determining a real-time semantic analysis result corresponding to the real-time voice recognition result based on the streaming semantic analysis engine and the context information under the condition that the real-time voice recognition result is detected to form a complete semantic unit. Therefore, by introducing a frame-by-frame real-time speech recognition mechanism and a semantic integrity discrimination model, a processing link of staged dependence and waiting in a traditional speech processing system is broken, and closer dynamic cooperation between speech recognition and semantic analysis is realized.
Owner:AISPEECH CO LTD

Method and system for generating mediation document

The invention provides a mediation document generation method and system, relates to the technical field of data processing, and comprises the step of forming structured case data through text preprocessing, named entity recognition and legal information association based on case text data and legal corpus information. Then, generating a preliminary mediation document with placeholders by utilizing case classification, semantic matching and a sequence-to-sequence model; after standard auditing and logic auditing are carried out on the documents, optimization is carried out based on reinforcement learning, language style migration and a text generative adversarial network, and the structural rationality, law term normativity and semantic integrity are improved. And finally, a mediator preference template is constructed through Few-shot Learning and meta learning, the format and language style of the document are adjusted in a personalized manner, and a mediation document conforming to the habits of a mediator is generated. According to the method, the intelligence and the accuracy of mediation document generation are improved, the law compliance and the logic preciseness are ensured, and the mediation working efficiency is improved.
Owner:SICHUAN XINYUNDIAO TECHNOLOGY SERVICE CO LTD

TF-IDF and cross entropy-based cue word compression method and system

The invention discloses a cue word compression method and system based on TF-IDF and cross entropy, belongs to the technical field of large model cue word compression, and aims to solve the problems that redundant information is introduced into long cue words, the model efficiency is reduced and the cost is increased. To-be-compressed content is divided into sentences at the sentence level and then converted into embedded vectors, and the Euclidean distance is calculated in combination with problem vectors so as to screen related sentences; calculating a TF-IDF value at the word level through a word frequency and an inverse document frequency to extract keywords and recombine sentences; and selecting a reference model and a basic model at the Token level, identifying the key Token based on a cross entropy loss difference value, and splicing the key Token in sequence to generate a compressed cue word. According to the method, a complex calculation structure is avoided, the inference efficiency is improved while the semantic integrity is maintained, and the resource consumption is reduced.
Owner:ARTIFICIAL INTELLIGENCE INNOVATION RES INST OF ZHEJIANG UNIV OF TECH BINJIANG DISTRICT HANGZHOU

Knowledge question-answering method and system based on topic knowledge graph retrieval enhancement

The invention discloses a knowledge question-answering method and system based on topic knowledge graph retrieval enhancement, and the method comprises the steps: firstly extracting a local topic represented in a triple form based on an original document through employing a large language model, carrying out the clustering, and generating a global topic triple set representing the global perspective of the whole document; secondly, on the basis of the global topic triple set, topic-guided entity and relation extraction is adopted, and a mixed knowledge graph is constructed; secondly, providing a semantic perception personalized PageRank algorithm, matching query semantics with semantics of edges in the mixed knowledge graph, and dynamically adjusting the weight of score propagation between nodes; and finally, designing a three-level progressive retrieval mechanism, retrieving multi-level information related to user query from the mixed knowledge graph, and inputting the multi-level information into the large language model to generate a final answer. According to the method, the semantic integrity and retrieval precision of the knowledge graph are remarkably improved, and the accuracy, comprehensiveness and enabling performance of generated answers are ensured.
Owner:HANGZHOU DIANZI UNIV

Text abstract generation method and system based on sparse attention acceleration

The invention discloses a text abstract generation method and system based on sparse attention acceleration, and the method comprises the steps: reading long text data, carrying out word segmentation and embedded coding processing, extracting a sequence feature vector, and mapping the sequence feature vector into a query matrix Q, a key matrix K and a value matrix V; constructing an abstract generation network, wherein the abstract generation network comprises a sparse attention calculation module, a feedforward calculation module, a prediction head module and a key value cache module; inputting the query matrix Q, the key matrix K and the value matrix V into an abstract generation network, passing through a backbone network formed by stacking a sparse attention calculation module and a feedforward calculation module for multiple times, processing by a prediction header module, and caching a historical decoding state in real time through a key value caching module to obtain an initial text feature vector; and based on the initial text feature vector, executing an autoregressive decoding process through an abstract generation network, and outputting a final abstract result. According to the method, the decoding process can be accelerated, and the semantic integrity and coherence of the generated abstract are ensured.
Owner:ZHEJIANG UNIV

Recommendation system method for keeping semantic integrity based on large language model

The invention discloses a recommendation system and method for keeping semantic integrity based on a large language model. According to the method, user-article interaction data and text information are fused, a prompt template of a user and an article is constructed, a configuration file with rich semantics is generated by utilizing a large language model, and initial embedded representation is extracted. Then, two-stage dimensionality reduction transformation is carried out through principal component analysis and a multi-layer perceptron, semantic information is reserved, and low-dimensional embedding is generated; on the basis of the embedding, cosine similarity is calculated, a user-user and article-article similar graph is constructed, and final embedding representation is generated through graph convolutional network coding. Meanwhile, collaborative filtering is combined to capture an interaction relationship and optimize a joint learning target, including recommendation loss, cross-modal comparison loss and regularization terms, so as to align semantics and collaborative filtering embedding, and finally generate a high-precision personalized recommendation result. The method effectively improves the semantic comprehension ability and recommendation accuracy of a recommendation system, and is suitable for various recommendation scenes.
Owner:NANJING UNIV OF AERONAUTICS & ASTRONAUTICS

Retrieval enhancement generation method and system based on multivariate fusion

The embodiment of the invention provides a retrieval enhancement generation method and system based on multivariate fusion, and the method comprises the steps: firstly structuring a knowledge base document, generating a knowledge graph, and importing an ES to establish an index; inputting question sentences, carrying out word segmentation and other processing, querying instance nodes from the map, and outputting meeting conditions according to answer sentence patterns; if not, fragmenting and blocking the document according to a title level, obtaining candidate results through vector, ES and atlas retrieval, normalizing scores, merging and optimizing the scores of the candidate retrieval results according to a retrieval source, calculating comprehensive scores, and outputting the comprehensive scores in a descending order; and finally, carrying out semantic integrity aggregation on the combined candidate results, and inputting into a large language model to obtain a final answer. According to the method, the relevance of the retrieved content is greatly improved, the semantic integrity of the retrieved content is improved, the model magic view problem is reduced, and the question and answer accuracy and the answer quality of a knowledge base are improved.
Owner:BEIJING ZHITONG YUNLIAN TECH CO LTD

Directory perception-based long document knowledge base construction method and program product

The invention discloses a long document knowledge base construction method based on directory perception and a program product, and belongs to the field of artificial intelligence and natural language processing. According to the scheme, an original long document is sequentially subjected to preprocessing, directory structure analysis, mixed blocking, double-tag generation, tag intelligent optimization, vectorization and meta-information mounting, and finally automatic construction and high-quality retrieval enhancement generation of a knowledge base are achieved. According to the method, the semantic integrity is guaranteed by fully utilizing the perceptual ability of the directory structure, the label quality and the retrieval efficiency are improved by combining a double-label system and intelligent optimization circulation, and the construction efficiency, the retrieval accuracy and the result traceability of the knowledge base are remarkably improved; the method is suitable for intelligent processing and application of long complex structure documents such as academic specialities, technical documents, policies and regulations and the like.
Owner:SOUTHEAST UNIV

Global knowledge extraction method and system based on large model and RAG technology

The invention discloses a global knowledge extraction method and system based on a large model and an RAG technology, and belongs to the technical field of large model data processing. According to the global knowledge extraction method and system based on the large model and the RAG technology, local sensitive hashing is adopted for duplicate removal, similar texts are combined through dynamic thresholds, and redundant data interference is reduced; a text segmentation strategy based on separator confidence ensures that segmented text blocks maintain semantic integrity and are adaptive to large model input length limitation, similarity retrieval and full-text retrieval are fused, results are reordered in combination with an RRF algorithm, semantic relevance and accuracy are considered, the multi-scene knowledge construction requirement is met, and the multi-scene knowledge construction efficiency is improved. High-correlation entity pairs are screened through mutual information, context knowledge is retrieved from a vector database in combination with an RAG system, semantic reasoning is conducted through a large language model, hidden logic relations among entities are generated, and jumping from data association to knowledge generation is achieved.
Owner:COMMUNICATION UNIVERSITY OF CHINA

Intelligent document information real-time retrieval method and system based on RAG technology

The invention relates to an intelligent document information real-time retrieval method and system based on the RAG technology, and the method comprises the steps: analyzing documents of various formats, extracting a text, maintaining the content continuity through adaptive semantic partitioning processing, and building a character-level position index at the same time; after vectorizing the text blocks, generating a plurality of rewriting queries for the original query; respectively carrying out mixed retrieval (combining keywords and semantic retrieval) for each rewriting query, and fusing and reordering results to obtain candidate text blocks; after correlation filtering, inputting a large language model according to correlation to generate an answer, and if the result is negative, triggering secondary retrieval and reordering; and finally, outputting a structured answer containing position information and supporting front-end visualization. According to the method, the semantic integrity maintenance, the multi-format document processing efficiency and the key information retrieval accuracy are effectively improved.
Owner:ECCOM NETWORK SYST CO LTD

Large language model multi-source inference mapping knowledge domain question and answer method and system

The invention discloses a large language model multi-source inference knowledge graph question answering method and system, and the method comprises the steps: extracting entities in a question, linking the entities with knowledge graph entities, and determining a subject entity set; searching paths in the knowledge graph by adopting a graph constraint reasoning method, and obtaining candidate answers and reasoning paths; detecting incomplete path coverage or answer dimension deficiency by using a large language model, triggering question decomposition to generate logically complementary sub-questions, and obtaining answers; establishing an inference source coordination layer, a coordination graph constraint inference source, a planning-retrieval inference source and a sub-problem inference source for multi-source evidence; and standardizing the output of each inference source into a structured evidence triple, sorting and selecting the first K evidences according to relevance, and generating a final answer by utilizing a large language model to carry out inductive inference. According to the method, reasoning can be dynamically adjusted, multi-source evidences can be effectively aggregated, logic consistency and semantic integrity are guaranteed, and the integrity and accuracy of knowledge graph multi-hop questions and answers are improved.
Owner:BEIFANG UNIV OF NATITIES

Dynamic anti-fraud decision-making method based on adversarial evolution

The invention relates to the technical field of intelligent anti-fraud, in particular to a dynamic anti-fraud decision-making method based on adversarial evolution, which comprises the following steps: firstly, constructing an original fraud sample data set as the basis of adversarial sample generation and model training; the attacker model is based on a language generation model, generates a confrontation sample with sensitive word disguise, word order disturbance and semantic concealment features through guiding prompt, and simulates a real fraud variant. And mixing the confrontation sample and the original sample to construct a training set, inputting the training set into a defender model for training, and optimizing an attacker model prompt with recognition accuracy to realize dynamic iterative lifting. A dynamic purification and semantic compensation mechanism is introduced into the defender model, a gradient sensitive area is inhibited, semantic integrity is repaired, the stability and recognition capability of the model in a complex context are enhanced, and therefore accurate recognition and interception of diversified novel fraud texts are achieved.
Owner:SOUTHWEST UNIVERSITY OF POLITICAL SCIENCE AND LAW +1

Medical document slicing and retrieval method and device, equipment and storage medium

The invention discloses a medical document slicing and retrieval method and device, equipment and a storage medium, and the method comprises the steps: generating a high-fidelity semantic summary text and a child node of a structured keyword set through a father node containing an original complete paragraph text, and storing the child node to a vector database; storing the semantic summary text and the structured keyword set into a keyword database; and when a user query instruction is received, performing dual-channel retrieval through the vector database and the keyword database, obtaining related target child nodes, extracting target original contents of a target father node from the target child nodes, transmitting the target original contents to the large language model, generating a final answer, and sending the final answer to the user. According to the method, the relevance and accuracy of retrieval results can be remarkably improved, instant and accurate decision support basis is provided for doctors, semantic integrity can be reserved, efficient multi-modal retrieval can be achieved, and the speed and efficiency of medical document slicing and retrieval are improved.
Owner:XIEHE HOSPITAL ATTACHED TO TONGJI MEDICAL COLLEGE HUAZHONG SCI & TECH UNIV

Video multi-language conversion method and system based on semantic segmentation

The invention provides a video multi-language conversion method and system based on semantic segmentation, which are used for solving the technical problems of semantic segmentation and low processing efficiency in the existing video multi-language conversion. The method comprises the following steps: firstly, performing audio-video separation preprocessing on an input video, and extracting an audio stream and a video stream; then voice feature extraction is carried out based on the audio stream, phoneme, pause and intonation features are obtained, and voice recognition is carried out to obtain text features; semantic segmentation points are determined according to the voice and text features, and semantic fragments are generated; the semantic fragments are converted into independent tasks, and multi-language conversion is executed through a parallel processing framework; and finally, synthesizing the processed semantic fragments to generate a video of a target language. According to the method, the semantic integrity is ensured through a semantic segmentation technology, the conversion efficiency is improved by adopting a parallel processing framework, and high-quality video multi-language conversion is realized.
Owner:SHENZHEN SHUNCHI ELECTRONIC TECHNOLOGY CO LTD

Robust noise reduction processing method and system for sound wave signal self-supervised learning enhancement

ActiveCN121415799ASpeech analysisPhysical realisationTime domainProbability propagation
The invention provides a sound wave signal self-supervised learning enhanced robust noise reduction processing method and system, and relates to the technical field of signal processing, and the method comprises the steps: carrying out the feature enhancement of an initial time-frequency representation through a dynamic adaptive mask strategy, constructing a self-supervised reconstruction task based on the mask time-frequency representation, separating noise and signal subspaces in a semantic manifold space, and carrying out the self-supervised learning enhanced robust noise reduction. And establishing a probability propagation network in combination with time domain continuity characteristics to model a local dependency relationship, and finally generating a noise reduction weight and realizing semantic fidelity optimization. The method can effectively improve the noise reduction effect and semantic integrity of sound wave signals in a noise complex environment.
Owner:BEIJING GUANYU INFORMATION TECHNOLOGY CO LTD

Long text reading understanding method based on dynamic partitioning and selection

The invention discloses a long text reading understanding method based on dynamic partitioning and selection, which is characterized in that the method adopts dynamic partitioning to dynamically divide long text input into discrete text blocks, selects and screens out irrelevant text segments by using selection partitioning, and splices the remaining text segments according to an original sequence, so as to obtain a long text reading understanding result. In order to conform to context window limitation predefined by a large language model, the method specifically comprises the steps of text preprocessing, dynamic blocking, block selection, large model output and the like. Compared with the prior art, the method has the advantages that the semantic coherence and the understanding accuracy are improved, the processing capability of the model on the super-long text is enhanced, the internal semantic integrity of each block is ensured, the semantic ambiguity caused by the block is reduced, the semantic coherence damage caused by the block with the fixed length is avoided, and the semantic coherence and the understanding accuracy are improved; and the processing capability of the model on the super-long text is enhanced, the data utilization efficiency is high, and the method has an important value and a good application prospect in practical application.
Owner:EAST CHINA NORMAL UNIV

Multi-language text adaptive configuration method and electronic equipment

The invention discloses a multi-language text self-adaptive configuration method and electronic equipment, and relates to the technical field of computers, and the method comprises the following steps: obtaining a source text set in response to a translation instruction received when a page runs, and translating the source text set according to a target language identifier and a constraint strategy to obtain a translated text set; according to a constraint strategy, carrying out adaptability detection on each translation in the translation set to obtain an adaptive set and a non-adaptive set; performing semantic rewriting on each non-adaptive translation in the non-adaptive set to obtain a target candidate set; and aggregating the adaptation set and the target candidate set in the same rendering frame, and rendering and displaying the adaptation set and the target candidate set in batches, so that the problems of lack of multi-language dynamic adaptation capability and insufficient semantic equivalent compression are solved, the accumulated layout offset is remarkably reduced on the premise of ensuring semantic integrity and readability, and the layout efficiency is improved. And the page stability and the user experience are improved.
Owner:INSPUR SUZHOU INTELLIGENT TECH CO LTD

Target-oriented video semantic communication system based on visual model

The invention provides a target-oriented video semantic communication system based on a visual model, and the system comprises a semantic extractor which is based on an SAM2 model and is used for processing an original video, generating a segmentation mask and extracting semantic information; the ViMama encoder is used for carrying out channel encoding on the output of the semantic extractor; the channel adaptation module is used for optimizing a coding sequence ViMama decoder based on the signal-to-noise ratio information of a physical channel, and is used for carrying out channel decoding to obtain a feature sequence; and the semantic reconstruction device is used for performing semantic reconstruction based on the feature sequence, recovering data and outputting a target video. According to the method, the problems of large redundant semantic information interference, insufficient deep semantic coding capability and poor communication robustness in a complex channel environment in a video are solved, efficient compression and robust transmission of video data are realized on the premise of ensuring semantic integrity, and the overall performance and adaptability of semantic communication are greatly improved.
Owner:湖南工商大学

Intelligent question answering method and system based on AI

The invention relates to the technical field of semantic processing and knowledge maps, and provides an AI-based intelligent question-answering method and system, and the method comprises the steps: collecting a knowledge map of an intelligent question-answering library, and obtaining an inquiry text of a user; performing word segmentation on the inquiry text to obtain a plurality of vocabularies; obtaining a reply design scale of the inquiry text; determining the number of analysis queues and establishing a plurality of analysis channels; obtaining the semantic intensity of each marked vocabulary and obtaining a high-intensity semantic vocabulary sequence; quantizing the analysis contribution of each mark vocabulary at each moment and the initial analysis limit at each moment, and further obtaining the buffer semantic integrity retention degree at each moment; judging and establishing a buffer space in real time, adjusting an analysis process, and finally outputting a semantic weight of each vocabulary; and screening a plurality of core vocabularies, obtaining a relation path in combination with the knowledge graph, and further generating a reply text. The invention aims to solve the problem that semantic weight acquisition is inaccurate due to word order interference in text keyword analysis.
Owner:HENAN KUKE CULTURE SCI & TECH CO LTD

Image fusion method and system based on multi-semantic guidance and mixed experts

The invention discloses an image fusion method and system based on multi-semantic guidance and mixed experts. The problems that in the prior art, robustness is insufficient, and visual fidelity and semantic integrity are difficult to balance are effectively solved. According to the method, infrared and visible light images to be fused and task identifiers are obtained, and firstly, a CLIP network is utilized to extract high-level semantic features as guide vectors; and then inputting the image and the guide vector into a pre-trained hybrid expert (MoE) fusion network. According to the network, intermediate features are extracted through an encoder, a gating network dynamically activates part of expert subnets according to task identifiers and semantic vectors and calculates routing weights, adaptive nonlinear transformation and weighted fusion are carried out on the features, and finally a high-quality fusion image is reconstructed through a decoder. The method can adapt to different task requirements, and the calculation efficiency is remarkably improved while the image fusion quality and the semantic consistency are improved.
Owner:XIDIAN UNIV

Semantic perception adaptive compression coding method and device based on saliency detection

The invention discloses a semantic perception adaptive compression coding method and device based on saliency detection. The method comprises the following steps: performing saliency detection on an original image based on a frequency coordination FT algorithm to obtain a saliency detection result; the saliency detection result comprises a saliency value of each pixel in the original image; dividing the original image into a plurality of image blocks, and performing importance grading on the image blocks according to a saliency detection result to obtain an importance level of each image block; performing adaptive compression coding on the image blocks according to the importance levels of the image blocks to obtain coded image blocks; and sending the coded image block to a channel for transmission. According to the method, collaborative optimization of the coding efficiency and the visual quality is realized on the premise of ensuring the semantic integrity of the saliency region of the image.
Owner:NANJING UNIV OF INFORMATION SCI & TECH

Multi-industry information system integrated heterogeneous data fusion processing platform

The invention relates to the technical field of heterogeneous data fusion processing, in particular to a heterogeneous data fusion processing platform for multi-industry information system integration, which comprises an industry knowledge graph module for acquiring business data of each industry, extracting a core entity relationship of each industry and constructing a field knowledge graph of at least two industries; when the system is used, the limitation of fixed weighted average is broken through by constructing a lightweight domain knowledge graph for each industry, extracting a core entity relationship and dynamically adjusting the fusion weight of each data feature, so that the purpose of dynamically adapting to the semantic feature difference of different industries is achieved; the method can be applied to cross-department data fusion of smart cities conveniently, semantic conflict intensity is accurately quantified by using a dual-channel neural network, adaptive strategy grading processing is combined, efficient resolution is performed through unit conversion during low conflicts, semantic integrity is guaranteed by using field isolation fusion during high conflicts, and conflict resolution efficiency and accuracy can be improved.
Owner:HUAIAN XINGMINGCHUANG INFORMATION TECHNOLOGY CO LTD

Real-time audio-video synchronous generation method and system based on semantic analysis

The invention relates to the technical field of audio and video generation, in particular to a real-time audio and video synchronous generation method and system based on semantic parse, and the system and method respectively decompose texts, audios and videos into minimum semantic units, and combine with a pre-training model to ensure that each modal unit is complete in semantics and accurate in granularity. And meanwhile, a cross-modal context consistency coefficient is introduced to filter pseudo associations with similar semantics but irrelevant scenes, and a matching uniqueness punishment mechanism is matched to avoid generation of conflicts, so that the problems that semantic associations are fuzzy and contents deviate from requirements in traditional video and audio generation are effectively solved. Video and audio synthesis based on an optimal text-audio and video semantic unit matching combination output by game optimization and core strategy parameters is a key advantage of guaranteeing generation quality. The accurate corresponding relation between the text and the audio and video unit is defined through the optimal matching combination; the core strategy parameters provide a dynamic adaptation basis for the synthesis process, and audio and video generation modules can be guided to adjust quality parameters and resource allocation according to scene requirements.
Owner:YONGBAO JIAFU (SHANGHAI) IND CO LTD

Network security sensitive information desensitization method and system for data leakage

The invention provides a network security sensitive information desensitization method and system for data leakage, and the method comprises the steps: obtaining a to-be-processed original data set, extracting the multi-dimensional context features, including a semantic association feature and a time sequence association feature, of the to-be-processed original data set; and calling a pre-trained sensitive information identification model to carry out sensitive information positioning processing on the multi-dimensional context features, generating a sensitive information identifier set, generating a corresponding desensitization processing scheme according to the sensitive information identifier set, including a replacement rule and a mask strategy, executing an effect verification operation on the desensitization processing scheme, and obtaining a desensitization result of the desensitization processing scheme. And a desensitization result verification report is generated, and the retention degree of the desensitized data set on the semantic integrity of the original data is evaluated, so that sensitive information can be comprehensively and accurately identified, accurate desensitization is realized, the semantic integrity of the data is retained, and the network security protection level is improved.
Owner:HEDUN DIGITAL (SHANGHAI) INFORMATION TECHNOLOGY CO LTD