Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

170 results about "Text stream" patented technology

Large model prompt project optimization system and method fusing domain knowledge graph

The invention discloses a large model prompt project optimization system and method fusing a domain knowledge graph. The system comprises an analysis module, a template generation engine module, a large model interaction interface module, a feedback analysis module and an optimization strategy module. And the analysis module forms a constraint coding signal containing an entity attribute incidence matrix. The template generation engine module forms an enhanced prompt text stream with a reservoir physical property parameter slot; the large model interaction interface module receives the enhanced prompt text stream and generates a question and answer response data stream containing geological terminologies; the feedback analysis module forms a feedback signal containing semantic deviation measurement through a semantic error vector calculation algorithm; and the optimization strategy module forms a parameter optimization instruction signal and transmits the parameter optimization instruction signal to the analysis module to complete iterative updating of the constraint condition. According to the large model prompt project optimization system fusing the domain knowledge graph, the problem of low answer accuracy of a large model in the oil-gas exploration field due to lack of professional constraints can be solved.
Owner:KARAMAY HONGYOU SOFTWARE

System and method for dynamic token estimation and buffer management in text-to-text variational autoencoder models

A method is provided for estimating the number of distinct tokens in a text stream using a modified text-to-text variational autoencoder (T5VQVAE) model. The method includes receiving a continuous input of a text stream; dynamically maintaining a buffer that stores a probabilistic subset of tokens from the text stream; calculating a sampling probability for each token based on a condition related to the current state of the buffer; updating the buffer based on the sampling probability to include or exclude tokens; encoding the buffered tokens into a latent space using the T5VQVAE model; and estimating the number of distinct tokens in the text stream based on the tokens in the buffer and the corresponding sampling probabilities.
Owner:LEPTUDE INC

Text block dynamic segmentation method and system based on RAG

The invention provides an RAG-based text block dynamic segmentation method and system, and the method comprises the steps: analyzing a multi-level directory structure of a document, and segmenting a text block by taking a directory node as a reference; for a document without a directory or with an incomplete directory, a mixed segmentation mode of rules and semantics is switched; performing potential Dirichlet distribution topic modeling on the continuous text stream, and calculating a topic distribution vector of each text segment in real time; performing adaptive segmentation on the detected topic boundary, inserting a hard segmentation mark at a topic mutation point, and performing soft segmentation on a gradual change topic area; the initial window size of the sliding window is set according to the document type, and the semantic density in the window is monitored in real time; and layering, blocking and recombining. The system comprises a segmentation mode module, a mark confirmation module and a block recombination module. According to the method, the RAG retrieval accuracy is improved, the memory occupation is reduced, and meanwhile, the streaming throughput is supported.
Owner:北京三维天地科技股份有限公司

Document processing method and system based on text content extraction

The invention relates to a document processing method and system based on text content extraction. The method comprises the steps that an original document containing text, image and format information is received, the encoding format of the document is automatically detected, character set conversion is executed, and hierarchical indexes including page numbers, paragraphs and tables are established for an unstructured document; the method comprises the following steps: synchronously processing text content and visual layout through a pre-trained visual-language model, extracting word-level and sentence-level semantic features by a text stream embedding layer, analyzing spatial distribution features of document elements by a visual encoder, and fusing text and visual features through a cross-modal attention mechanism; and loading the domain knowledge graph matched with the document type, and executing entity linking to associate the text mentions to the knowledge nodes. According to the document processing method and system based on text content extraction, through the synergistic effect of vision-text joint coding and knowledge enhancement, the accuracy of financial contract key clause recognition tasks is improved, the error rate is lower than that of industry benchmark products, and the semantic understanding precision is remarkably improved.
Owner:WIN THE BID HUIKANG TECH CO LTD

Method and apparatus for emotion recognition in real-time based on multimodal

At least one aspect of the present disclosure provides an emotion recognition method using an audio stream performed by an emotion recognition apparatus including receiving an audio signal having a preset unit length to generate the audio stream corresponding to the audio signal; converting the audio stream into a text stream corresponding to the audio stream; and inputting the audio stream and the converted text stream to a pre-trained emotion recognition model to output a multi-modal emotion corresponding to the audio signal.
Owner:SK TELECOM CO LTD

Short video production method and system based on artificial intelligence

The invention discloses a short video production method and system based on artificial intelligence, and relates to the technical field of multimedia processing, and the method comprises the steps: inputting a split script into a multi-modal generation engine, employing a lower triangle space-time attention matrix to generate a visual stream, and generating a text stream and an audio stream; calculating a neural conduction delay amount between the visual stream and the audio stream based on an audio-visual perception delay prediction model, and generating a corrected audio stream through time axis forward compensation; calculating a cross-modal attention weight matrix according to the node association strength of the dynamic causal graph, and projecting the text stream and the corrected audio stream to a joint feature space to generate a multi-modal fusion feature; and inputting the multi-modal fusion features into a format adaptation engine, and after time-space consistency verification, embedding a neural implicit watermark and encoding the neural implicit watermark into a short video. According to the invention, the semantic consistency and coordination among the text, the audio and the visual content are improved, so that high-quality and high-intelligence automatic generation of the short video is realized on the whole.
Owner:BEIJING YIJIABANG TECH CO LTD

Automatic conference recording and abstract generating method for intelligent conference

The invention relates to the technical field of data processing, in particular to an automatic conference recording and abstract generating method for an intelligent conference, which comprises the following steps of: capturing voice streams of multiple speakers in real time through a directional microphone array, generating an original text stream with a time sequence mark based on real-time voiceprint clustering, and synchronously extracting intention intensity parameters in the voice streams; performing intention-driven dynamic segmentation processing on the original text stream; performing agenda-perceived abstract block generation on the segmented text, wherein decision expressions meeting a semantic density threshold value are extracted as abstract core blocks; and assembling the abstract core blocks into a structured abstract document according to the hierarchical structure of the agenda template. Compared with a traditional linear transcription mode based on voice recognition, the method has the advantages that the mapping precision between the speech content and the identity of the participant is remarkably improved, and a solid foundation is provided for semantic segmentation and decision tracing.
Owner:广东公信智能会议股份有限公司

Low-delay streaming voice interaction system with interruption processing function

The invention provides a low-delay streaming voice interaction system with an interruption processing function, and relates to the technical field of artificial intelligence. Necessary preprocessing and acoustic feature extraction are performed through a real-time acoustic processing module, and distortion and complex environmental noise introduced by an interaction channel are resisted through a robustness enhancement technology; the streaming acoustic decoding module performs acoustic modeling, language model application and decoding in real time in parallel, and outputs an ultra-low-delay text transfer result stream; the real-time acoustic processing module is combined with a signal processing technology and is responsible for detecting user voice activity in a high-precision and ultra-low-delay manner, and particularly judging the real-time voice activity state of a user through the user voice activity in an AI voice playing period; an efficient and low-delay bidirectional streaming network transmission mode is adopted between modules of the system and between the modules and a communication platform, and it is ensured that audio streams, acoustic feature streams, text streams and control signals can be transmitted and processed in real time with extremely low end-to-end delay.
Owner:GUANGDONG CHAOTENG INFORMATION TECHNOLOGY CO LTD

AI streaming output rich text real-time rendering method based on minimum rendering unit recognition

The invention provides an AI streaming output rich text real-time rendering method and system based on minimum rendering unit recognition. The method comprises the steps that an AI model outputs a text stream to an AI agent; defining a minimum rendering unit (MRU), wherein the minimum rendering unit is self-consistent in format and complete; the AI agent analyzes the text streams in real time, the analysis process continuously judges whether the currently accumulated text streams form a complete MRU or not, and if yes, a rendering trigger mark is added for output; if not, directly outputting; the client receives the output result and the text stream and judges whether current rendering is triggered or not, and if yes, the rendering result is displayed; and if not, continuing to receive the output result until the rendering result is displayed after the rendering is triggered. According to the method, real real-time rendering can be realized, and a user can see a formatted result to appear on a screen step by step and smoothly while AI generates contents, so that response speed perception and interaction experience are greatly improved, and the pain point of blank screen waiting is solved.
Owner:SHANGHAI GREAT WISDOM SHENJIU INFORMATION TECH CO LTD

Voice recognition processing method, system and equipment based on conference scene and medium

The invention relates to a voice recognition processing method, system and device based on a conference scene and a medium, and belongs to the technical field of voice processing. The voice recognition processing method comprises the following steps: acquiring an original conference audio stream collected by a microphone array; performing signal preprocessing on the original conference audio stream collected by the main channel, and outputting a pure voice signal; generating a sound source orientation thermodynamic diagram based on the original conference audio stream; extracting multi-dimensional voiceprint feature vectors from the pure voice signals, performing dynamic grouping, outputting a voice fragment set marked with voiceprint IDs, and generating an initial transcription text; dynamically correcting the initial transliteration text, and outputting a transliteration text stream with an industry term tag; and performing periodic memory enhancement processing on the transliteration text stream, outputting and analyzing a long text, and generating structured conference summary data. According to the invention, the automation level and accuracy of conference voice processing can be improved.
Owner:CHINA TRANSPORT INFORMATION TECH GRP CO LTD

NLP-based customer service dialogue quality detection method and system

The invention provides an NLP-based customer service dialogue quality detection method and system, and relates to the technical field of natural language processing, and the method comprises the steps: obtaining an e-commerce platform customer service dialogue text stream in real time; semantic understanding is conducted on the dialogue text flow, a semantic understanding result is output, and the semantic understanding result comprises a user consultation intention, key question elements and customer service inquiry information entity integrity; based on the semantic understanding result, performing sentiment analysis on the dialogue text stream, and outputting a structured sentiment analysis result; setting an initial state point and a termination state point in the dialogue processing path based on the semantic understanding result and the sentiment analysis result; the initial state point is a semantic vector fusing core problem elements fed back by the user for the first time and a current user emotional state value. According to the invention, through full-process intelligent processing, dynamic monitoring and accurate optimization of customer service quality are realized, and user satisfaction, customer service efficiency and business normalization are improved.
Owner:MEGAVIEW INTELLIGENCE TECH LTD

AIGC content security monitoring system and method based on dynamic reasoning and context awareness

The invention relates to the technical field of natural language processing, in particular to an AIGC content safety monitoring system and method based on dynamic reasoning and context awareness, and the system comprises a semantic graph construction module which is used for extracting entity nouns and predicate verbs in a text according to a received AIGC interaction text flow, generating a semantic concept node set, and sending the semantic concept node set to a database; and performing directed connection and hierarchical nesting on the concepts in the semantic concept node set according to a logic direction according to a subject-predicate-object dependency relationship rule. According to the method, the curvature value of the semantic track is calculated, and the similarity between the direction vector and the center of the sensitive semantic cluster is combined for double verification, so that sudden turning of an intention in a dialogue process or progressive induction to a sensitive field can be perceived, and abnormal mutation can be recognized through a curvature pulse form; therefore, hostile attack behaviors are accurately captured in real time in dynamic interaction, and the defense capability for context dependent attacks and implicit induction behaviors is improved.
Owner:XINGXUAN DIGITAL TECHNOLOGY (SHANGHAI) CO LTD

Hatred speech and privacy information identification system and method based on large model

The invention relates to the technical field of natural language processing and information safety crossing, in particular to a hatred speech and privacy information recognition system and method based on a large model, and the system comprises a data collection layer, a model processing layer, a result output layer and a feedback optimization layer. The system has the beneficial effects that social media text streams are acquired through the data acquisition layer, a hatred detection engine (based on a Transform hybrid neural network) and a privacy recognition engine (a rule base and a lightweight CNN model) in the model processing layer perform processing, the result output layer generates structured output and executes dynamic desensitization, and the feedback optimization layer realizes manual auditing and incremental learning.
Owner:SHANDONG LANGCHAO YUNTOU INFORMATION TECH CO LTD

PDF drawing data extraction method and system based on intelligent identification

The invention relates to the field of drawing recognition, in particular to a PDF drawing data extraction method and system based on intelligent recognition. Comprising the following steps: reading an internal structure of a PDF engineering drawing to obtain a native text stream, a vector path and a grating image; identifying the native text flow through a shunt preprocessing framework to form structured text data; rendering the vector path and the grating image to obtain a background image; analyzing the structured text data by utilizing the intelligent recognition model through the character recognition and extraction sub-model, obtaining drawing metadata and recording the position, and obtaining a character recognition result; analyzing the background image through a graphic element recognition and classification sub-model, recognizing and classifying component elements, and obtaining a graphic recognition result; and performing fusion according to the visual space corresponding relation to form a drawing analysis result. According to the method, the adaptive capacity of engineering drawings with various sources and different qualities is improved through the shunting preprocessing framework and the intelligent identification model.
Owner:TAIZHOU HUAWEI INFORMATION TECH CO LTD

Multi-modal fusion short message compliance and security dual-auditing method and multi-modal fusion short message compliance and security dual-auditing system

The invention discloses a multi-modal fusion short message compliance and security dual auditing method and system, and the method comprises the steps: encrypting account information through employing an irreversible algorithm based on SHA256 and a random salt value, and decomposing the content of a short message into a text stream, a link stream and a symbol stream through employing a regular expression; traversing each character in the short message content by adopting a prefix mode based on a Trie tree in combination with an AC automaton algorithm to detect sensitive words; performing symbol semantic classification mapping and analysis on the special symbol feature data extracted from the symbol stream; carrying out sending behavior analysis on the text feature data extracted from the text stream; performing special detection on link feature data extracted from the link stream; based on the factors corresponding to the sensitive words, the semantics, the behaviors and the links and the weights of the factors, a multi-modal feature fusion decision risk assessment algorithm is adopted to output a risk score value and a risk decision rule of the risk score value so as to execute short message interception operation or short message release operation. According to the invention, full-dimension perception and dynamic defense of risks can be realized.
Owner:JIANGXI TIANLI TECH INC

Lightweight member marketing method and device for real-time conversation, terminal and storage medium

The invention discloses a lightweight member marketing method and device for real-time conversation, a terminal and a storage medium, and the method comprises the steps: responding to a received user voice stream, continuously generating an incomplete intermediate text stream through streaming voice recognition, and inputting the intermediate text stream into a lightweight marketing response model; the lightweight marketing response model predicts the potential intention of the user in real time based on the intermediate text stream, and pre-assembles one or more candidate response frameworks from the fragmented marketing corpus; when it is detected that the utterance of the user ends, the real intention of the user is determined based on a complete final text stream generated by streaming speech recognition; and arbitrating the candidate reply framework according to the weighted decision factor set to obtain an arbitration result, determining final reply content according to the arbitration result, converting the final reply content into an instant response call, and outputting the instant response call. According to the method, the time delay of the real-time call marketing process can be reduced, and the concurrent processing capability is improved.
Owner:SUZHOU XIYUAN DIGITAL TECH CO LTD

Simultaneous interpretation method and system based on large model and electronic equipment

The invention discloses a simultaneous interpretation method and system based on a large model and electronic equipment, and the method comprises the steps: extracting bilingual parallel corpora related to terms from professional resources based on a standardized professional dictionary, obtaining qualified corpora through data enhancement processing and manual screening, and constructing a multi-level corpus according to the levels of words, sentences and paragraphs; the method comprises the following steps: receiving an input audio stream in real time, extracting acoustic features through preprocessing, inputting a pre-established large-scale speech recognition model, and carrying out incremental decoding on the acoustic features in a sliding window mode; and calling a sentence boundary prediction network to judge a pause point, and outputting a text stream with a timestamp. Performing fine tuning on the large-scale speech recognition model by using a multi-level corpus, translating a text stream based on the fine-tuned large-scale speech recognition model, and constraining term translation according to a standardized professional dictionary; and synchronously displaying the audio output in the translation result and the subtitles. According to the scheme, the terminology recognition and translation accuracy is improved, and simultaneous interpretation delay is reduced.
Owner:TONGFANG KNOWLEDGE DIGITAL PUBLISHING TECH CO LTD

RAG content generation method and system based on dynamic slicing and adaptive fusion

The invention discloses an RAG content generation method and system based on dynamic slicing and self-adaptive fusion, and relates to the field of natural language process.The method comprises the steps that in response to a user query instruction, semantic fusion processing is conducted on a plurality of related text blocks, and a preliminary text sequence is traversed; detecting a logic breakpoint between adjacent text blocks through a context association identifier, recalling a middle connection text block corresponding to the logic breakpoint, and inserting the middle connection text block into a sequence to generate a fusion text stream; performing adaptive information density adjustment on the fused text stream to obtain optimized context content; generating structured prompt information based on the optimized context content and the user query instruction; and calling the target large language model to generate a user query result based on the structured prompt information. The answer accuracy and logic continuity of the target large language model for user query are effectively improved, and the utilization efficiency of model computing resources is optimized.
Owner:HANGZHOU WEIMING XINKE TECH CO LTD +1

Real-time video processing method and system based on artificial intelligence

The invention discloses a real-time video processing method and system based on artificial intelligence, and belongs to the technical field of computers. Through deep fusion of multi-source information and a regional adaptive coding mechanism, the problem of collaboration between content generation and network resource allocation in a high-interactivity live broadcast scene is solved. In the feature extraction link, video, audio and text streams are synchronously processed, so that three-dimensional perception of a live broadcast scene is realized; in the multi-mode driving instruction set generation process, audio event features and text interaction features are subjected to semantic association, so that content generation and real-time interaction are highly synchronized, and response delay caused by information splitting in a traditional scheme is avoided; according to the regional differentiation coding processing, resource allocation is dynamically adjusted according to the regional importance of visual semantic feature recognition and the network state, and the coding quality of a digital human region is preferentially guaranteed, so that the stability of core visual elements is maintained under the network fluctuation condition, and the intelligent level of real-time video processing and the user experience are improved.
Owner:HANGZHOU QUKAN TECH CO LTD +1

Security transaction instruction intelligent analysis method and system based on large model

The invention relates to the technical field of financial science and technology, and discloses a security transaction instruction intelligent analysis method and system based on a large model, and the method comprises the steps: carrying out the compliance screening of an instruction text flow and an instruction feature chart, and obtaining an initial text flow and an initial feature chart; analyzing a transaction feature factor corresponding to the initial feature chart, extracting a space-time structure feature vector from the initial feature chart, and analyzing a structure description text corresponding to the space-time structure feature vector to generate a chart analysis text corresponding to the initial feature chart; calculating segmentation entropy loss corresponding to the initial segmented text, and performing segmentation iterative optimization on the initial text stream to obtain a target segmented text; generating an instruction target text corresponding to the security transaction instruction, calculating a structured likelihood between instruction tags, and determining an association aggregation tag in the instruction tags; and analyzing the transaction intention corresponding to the security transaction instruction, and generating an analysis result corresponding to the security transaction instruction. According to the invention, the accuracy of security transaction instruction analysis can be improved.
Owner:CHINALIN SECURITIES CO LTD

Multi-language voice content recognition method and system

The invention relates to a multi-language voice content recognition method and system, and belongs to the technical field of voice signal processing, and the recognition method comprises the steps: collecting original audio stream data, executing the noise reduction filtering processing, and segmenting the original audio stream data into a plurality of audio segments; extracting acoustic feature vectors of the audio clips; inputting the audio clip into the speech recognition model group, and obtaining a text clip and a confidence score; fusing all the text fragments to generate a preliminary recognition text, segmenting the preliminary recognition text into text blocks, and translating the text blocks into a target language text block by block through a streaming translation model in combination with a semantic feedback tag received in real time; executing bidirectional translation verification on the text blocks of which the confidence scores are lower than a set threshold value; inputting the translated text stream and the acoustic feature vector into a semantic analysis model, and fusing an output result to obtain a semantic feedback tag; and constructing a cross-modal association graph through the graph neural network, outputting a key information abstract and triggering an alarm signal. The speech recognition accuracy can be improved while the recognition efficiency is guaranteed, and risk response is carried out.
Owner:BEIJING HIZHI TECH CO LTD

Intelligent conference auxiliary method and system

The invention provides an intelligent conference auxiliary method and system, and belongs to the field of network communication. The method comprises the following steps: excluding target application window data from shared stream data in a process of generating the shared stream data based on collected screen shared data, and sending the target application window data to a conference terminal for visual interface feedback; acquiring voice data of the conference terminal in real time, and transliterating the voice data to generate voice text stream data; in response to the screenshot instruction, performing screenshot analysis on the selected screen image to generate image text stream data; and inputting the generated voice text stream data and / or image text stream data into a pre-training language model, fusing the context information of the current session, and generating a structured response content conforming to a preset rule. According to the intelligent conference assisting method and system, it can be ensured that a user selectively hides a specific application window in screen sharing, and intelligent assistance can be provided for a video conference of the user, so that the personalized and scene-aware interaction experience of the user is improved.
Owner:SHANXI JUEPA CLOUD TECHNOLOGY CO LTD

Audio interaction processing methods and system, server, client and electronic device

Audio interaction processing methods and system, a server, a client and an electronic device. A method is executed by an audio processing server, and comprises: converting an input audio sent by a client into a text to be replied to (S102); sending to a text processing server the text to be replied to, so as to generate a reply text stream (S104); receiving a request of the client for a reply audio stream corresponding to the reply text stream (S106); on the basis of a message identifier, acquiring the reply text stream from the text processing server (S108); generating a reply audio stream on the basis of the reply text stream (S110); and sending the reply audio stream to the client for playing (S112).
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

Dynamic knowledge retrieval enhancement method based on large language model

The invention relates to the technical field of retrieval enhancement, in particular to a big language model-based dynamic knowledge retrieval enhancement method, which comprises the following steps of: acquiring theoretical regulations and actual operation data of an unmanned aerial vehicle, and driving a model generation request; monitoring the reasoning text stream in real time, intercepting a logic triple to perform causal verification, and constructing negative feedback prompt blocking and resetting for unsupported logic; executing double-track system dynamic retrieval aiming at regulations and data, evaluating content conformity by utilizing reflection Token and executing conflict resolution correction; and generating a fault hypothesis by adopting tree decoding, constructing an anti-fact query to retrieve the evidence of the security, and pruning the path with the evidence of the security to determine a final path. According to the method, illusion is effectively blocked through real-time logic verification and a negative feedback mechanism, double-track retrieval and anti-fact pruning are combined, strict alignment between the reasoning process and physical facts and theoretical regulations is ensured, and the accuracy and credibility of fault diagnosis are remarkably improved.
Owner:ZHEJIANG FULIN TECH CO LTD

Voice real-time question and answer processing method based on large model and domain knowledge base

The invention discloses a voice real-time question and answer processing method based on a large model and a domain knowledge base, which belongs to the technical field of computer data processing, and comprises the following steps: acquiring a voice input stream of a user, carrying out intelligent sound wave deconstruction on the voice input stream, generating a text stream, and carrying out dependency syntactic analysis and semantic role labeling; generating intention information and key information, dynamically accessing a domain knowledge base, executing predictive loading, generating a preloaded data subset, performing fusion processing by combining the intention information, the key information and the preloaded data subset, generating a fusion result, performing accuracy verification and correction, generating a natural language answer, and converting the natural language answer into voice output. The technical scheme of combining knowledge predictive loading driven by intention analysis, context fusion and answer traceability correction is adopted, and low-delay response, high-precision intention understanding and high-factuality answer of voice questions and answers can be achieved.
Owner:STATE GRID SHANDONG ELECTRIC POWER CO

Teaching note generation method and device based on large language model and medium

The embodiment of the invention discloses a teaching note generation method and device based on a large language model and a medium, belongs to the technical field of artificial intelligence, and solves the problems that classroom notes generated in the prior art are lack of content structuring and knowledge relevance, poor in readability and difficult to be used for effective review. The method comprises the following steps: acquiring a teacher voice signal in a classroom teaching environment in real time through pickup equipment, and converting the teacher voice signal into an initial text stream through a preset voice recognition model; taking the initial text flow as query input, and performing semantic similarity retrieval on the initial text flow and a preset course knowledge base to obtain a plurality of related knowledge fragments; generating an enhanced prompt based on the initial text stream and the plurality of related knowledge fragments, and generating a structured classroom note based on the enhanced prompt and a preset large language model; and outputting the structured classroom note to a user interface, and optimizing a classroom teaching note generation process based on feedback information and modification information of the user.
Owner:天元大数据信用管理有限公司

Multi-modal fusion document image classification method based on multi-branch deep convolution and hierarchical semantic modeling

The invention relates to the field of computer vision and artificial intelligence, in particular to a multi-modal fusion document image classification method based on multi-branch deep convolution and hierarchical semantic modeling. According to the method, visual and text information of a document is co-processed through a double-flow architecture; a visual semantic feature of a document image is efficiently extracted from a visual flow by adopting a lightweight convolution operation and multi-scale feature fusion mechanism; an original text is obtained from a text stream through an optical character recognition technology, text error correction and word embedding extraction are carried out through a fine tuning language model, word-level and sentence-level context modeling is realized by inputting a hierarchical semantic coding network, and document-level features are generated. And high-precision document image classification is realized by fusing the double-flow feature vectors and combining a full-connection layer. Experimental results show that the method has an excellent multi-modal feature fusion effect, relatively high classification precision and strong robustness to complex document scenes in a document image classification task, and is suitable for application scenes of various document image classification.
Owner:GUILIN UNIV OF ELECTRONIC TECH

Data compression method and system

The invention relates to the technical field of data compression, in particular to a data compression method and system, which comprises an LZ module, a text stream module, a sequence stream module and an interval entropy coding framework, and is characterized in that the LZ module is used for performing repeated data screening on original data to generate a text stream and a sequence stream; the LZ module screens duplicated data through a hash verification method, specifically, hash values are calculated for input data blocks, the duplicated data are determined based on hash value comparison, and direct data content comparison is avoided; and the character flow module is used for carrying out entropy coding on the character flow. According to the method, an interval entropy coding framework is innovatively proposed, a better probability model is constructed through artificial intelligence and a mathematical method, a better finite state machine is trained, the optimal state machine is used in the compression and decompression process, and the problem that the decompression speed of a neural network method is low is efficiently solved.
Owner:LANZHOU UNIV

Risk public opinion analysis method based on natural language processing

The invention discloses a natural language processing-based risk public opinion analysis method, which comprises the following steps of: capturing a public opinion text in real time, and desensitizing a user identity field by adopting a Hash pseudo identifier to obtain an original text stream; constructing a double-classification weak supervision framework, calculating a risk confidence coefficient, and when the risk confidence coefficient is greater than or equal to a first threshold value, writing a corresponding text into a labeling pool to form a risk corpus; mapping the risk corpus into a dynamic heterogeneous risk knowledge graph according to an entity-relationship-time slice meta-path rule; applying a multi-head space-time diffusion Transform to a sub-graph intersecting with the current window in the graph to obtain a risk propagation vector of an event node; and inputting the risk propagation vector and the historical baseline vector into a causal comparative analysis model, outputting a risk early warning index, and when the early warning index is greater than or equal to a second threshold value, pushing a risk alarm including an event path to a background management and control center. According to the invention, the automation degree, accuracy and reliability of public opinion risk monitoring and early warning are effectively improved.
Owner:CHANGZHOU JIADO HIGH-TECH CO LTD

A data compression method and system

The present application relates to the technical field of data compression, in particular to a data compression method and system, comprising: an LZ module, a text stream module, a sequence stream module, an interval entropy encoding framework, the LZ module is used for screening repeated data of original data, generating a text stream and a sequence stream; the LZ module screens repeated data through a hash verification method, specifically: calculating a hash value for the input data block, determining repeated data based on hash value comparison to avoid direct data content comparison; the text stream module is used for entropy encoding the text stream; the present application innovatively proposes an interval entropy encoding framework, constructs a more optimal probability model through artificial intelligence and mathematical methods, trains a more optimal finite state machine, and uses the optimal state machine in the compression and decompression process, efficiently solving the problem of slow decompression speed of neural network type methods.
Owner:LANZHOU UNIV